Job Summary The Senior Data Engineer will serve as the primary technical owner responsible for the ongoing maintenance enhancement operational support and reliability of Nebraskas enterprise data processing framework supporting Medicaid Encounter Processing and enterprise data ingestion.
This mission-critical platform processes provider member reference and encounter data used by downstream Medicaid systems and reporting. The application framework is developed using Scala Apache Spark Hive and Drools and operates on the Cloudera Data Platform (CDP).
Following the States migration from an Apache open-source Hadoop environment to Cloudera Data Platform this position will assume primary responsibility for maintaining the production code base implementing business and regulatory changes supporting production operations resolving application defects and performing routine operational administration of the data processing environment.
This position serves as the single primary technical owner of the application framework and is expected to maintain sufficient expertise across the complete technology stackincluding Scala Apache Spark Hive Drools and the Cloudera Data Platformto ensure the ongoing reliability maintainability and operational continuity of this mission-critical system.
Primary Ownership
Own the Scala/Spark application framework supporting Medicaid Encounter Processing.
Own the Drools business rule implementation and ongoing rule maintenance.
Own enterprise data ingestion processes for master data and source system files.
Own Spark and Hive batch processing workflows.
Own production issue investigation and application defect resolution.
Own production releases version management and deployment coordination.
Own application performance tuning and optimization.
Own technical documentation operational procedures and knowledge transfer.
Provide basic operational administration and monitoring of the Cloudera Data Platform while coordinating with infrastructure and cloud operations teams.
Required Qualifications
Minimum 7 years of experience developing enterprise-scale distributed data processing applications.
Strong hands-on development experience with Scala.
Experience developing applications using Apache Spark 2.x and/or Spark 3.x.
Experience developing and maintaining applications running on Cloudera Data Platform (CDP) 7.x or equivalent enterprise Hadoop distributions.
Experience with Apache Hive 3.x for data processing and Hive table management.
Experience implementing business rules using the Drools Rules Engine.
Strong SQL development and query optimization skills.
Experience supporting Linux-based production environments.
Experience troubleshooting distributed Spark applications in production.
Experience using Git and modern version control practices.
Preferred Qualifications
Medicaid or healthcare experience.
Experience with Medicaid Encounter Processing.
Experience with Cloudera Manager HDFS and YARN.
Experience integrating with IBM DataStage.
Familiarity with AWS infrastructure supporting Cloudera.
Minimum 7 years of experience developing enterprise-scale distributed data processing applications.
Advanced (7-9 Years)
Yes
Skills
Others
Drools
Experience implementing business rules using the Drools Rules Engine
Proficient (4-6 Years)
Yes
Skills
Others
Scala
Strong hands-on development experience with Scala.
Proficient (4-6 Years)
Yes
Skills
Others
SQL and Linux
Strong SQL development and query optimization skills Experience supporting Linux-based production environments
Yes
Skills
Others
Apache experience
Experience developing applications using Apache Spark 2.x and/or Spark 3.x. Experience with Apache Hive 3.x for data processing and Hive table management.
Proficient (4-6 Years)
Yes
Skills
Others
Cloudera
Experience developing and maintaining applications running on Cloudera Data Platform (CDP) 7.x or equivalent enterprise Hadoop distributions
No
Required Skills:
Apache
Job Title:Senior Data Engineer- 66129 Duration:12 Months Location: Remote / Lincoln Nebraska Job SummaryThe Senior Data Engineer will serve as the primary technical owner responsible for the ongoing maintenance enhancement operational support and reliability of Nebraskas enterprise data processing f...
Job Title:Senior Data Engineer- 66129
Duration:12 Months
Location: Remote / Lincoln Nebraska
Job Summary The Senior Data Engineer will serve as the primary technical owner responsible for the ongoing maintenance enhancement operational support and reliability of Nebraskas enterprise data processing framework supporting Medicaid Encounter Processing and enterprise data ingestion.
This mission-critical platform processes provider member reference and encounter data used by downstream Medicaid systems and reporting. The application framework is developed using Scala Apache Spark Hive and Drools and operates on the Cloudera Data Platform (CDP).
Following the States migration from an Apache open-source Hadoop environment to Cloudera Data Platform this position will assume primary responsibility for maintaining the production code base implementing business and regulatory changes supporting production operations resolving application defects and performing routine operational administration of the data processing environment.
This position serves as the single primary technical owner of the application framework and is expected to maintain sufficient expertise across the complete technology stackincluding Scala Apache Spark Hive Drools and the Cloudera Data Platformto ensure the ongoing reliability maintainability and operational continuity of this mission-critical system.
Primary Ownership
Own the Scala/Spark application framework supporting Medicaid Encounter Processing.
Own the Drools business rule implementation and ongoing rule maintenance.
Own enterprise data ingestion processes for master data and source system files.
Own Spark and Hive batch processing workflows.
Own production issue investigation and application defect resolution.
Own production releases version management and deployment coordination.
Own application performance tuning and optimization.
Own technical documentation operational procedures and knowledge transfer.
Provide basic operational administration and monitoring of the Cloudera Data Platform while coordinating with infrastructure and cloud operations teams.
Required Qualifications
Minimum 7 years of experience developing enterprise-scale distributed data processing applications.
Strong hands-on development experience with Scala.
Experience developing applications using Apache Spark 2.x and/or Spark 3.x.
Experience developing and maintaining applications running on Cloudera Data Platform (CDP) 7.x or equivalent enterprise Hadoop distributions.
Experience with Apache Hive 3.x for data processing and Hive table management.
Experience implementing business rules using the Drools Rules Engine.
Strong SQL development and query optimization skills.
Experience supporting Linux-based production environments.
Experience troubleshooting distributed Spark applications in production.
Experience using Git and modern version control practices.
Preferred Qualifications
Medicaid or healthcare experience.
Experience with Medicaid Encounter Processing.
Experience with Cloudera Manager HDFS and YARN.
Experience integrating with IBM DataStage.
Familiarity with AWS infrastructure supporting Cloudera.
Minimum 7 years of experience developing enterprise-scale distributed data processing applications.
Advanced (7-9 Years)
Yes
Skills
Others
Drools
Experience implementing business rules using the Drools Rules Engine
Proficient (4-6 Years)
Yes
Skills
Others
Scala
Strong hands-on development experience with Scala.
Proficient (4-6 Years)
Yes
Skills
Others
SQL and Linux
Strong SQL development and query optimization skills Experience supporting Linux-based production environments
Yes
Skills
Others
Apache experience
Experience developing applications using Apache Spark 2.x and/or Spark 3.x. Experience with Apache Hive 3.x for data processing and Hive table management.
Proficient (4-6 Years)
Yes
Skills
Others
Cloudera
Experience developing and maintaining applications running on Cloudera Data Platform (CDP) 7.x or equivalent enterprise Hadoop distributions