Data Scientist
Austin, TX - USA
Job Summary
Data Scientist role
Candidates must be local to Austin TX (living within a 25-mile radius of Austin TX currently no relocating candidates will be considered).
On-site work is required 4-5 days per week for this role at the TX office.
The interview process will be 1 technical interview and then an offer will be made.
Please let me know if there are any questions.
Job Description:
| TxDOT work to be accomplished | Resource with strong skills in Data architecture and expertise in dimensional data modeling preferably in a bigdata or an enterprise data warehouse environment. Experience with Informatica Power Exchange for CDC based replication or similar tool is a plus. Expertise in development of data ingestion to Hadoop based data lake (HDFS Hive) and S3 storage complex data transformations and data security. Programming expertise in a major programming/scripting language including but not limited to Java C Scala python Go. Hands on delivery capability working in a team across the full lifecycle Data Source Analysis Design Development Testing and Implementation. This includes providing input into requirements platform development technical design of the project level technical architecture for ETL big data application design and development testing and deployment of the proposed solution. Proficient in Hadoop ecosystem components such as Hive Yarn Knox Ranger Sqoop Oozie HiveQL Spark SQL/data access. Experience in ETL development using Informatica BDM Kafka or comparable software. Understanding and skills in data governance security architecture load balancing and troubleshooting. Well-versed and deep development experience in Big Data Platform clusters on AWS HDFS storage optimization Data Lake/Data Warehouse foundation capabilities. |
| Minimum Yrs of Experience Skills and Qualifications | 6 years of Cloud Data Warehousing with specialization in Snowflake and Python |
| Preferred Skills and Qualifications | 1 year of prior experience with State Agencies |