Senior Data Engineer – Encounter Processing & Data Platform - Remote

 Posted 2 days ago
     
5-10 years experience
Apply Now

Please mention DailyRemote when applying

AI Summary

The Senior Data Engineer will serve as the primary technical owner for Nebraska's Medicaid Encounter Processing and enterprise data ingestion framework. Responsibilities include maintaining the production code base, implementing regulatory changes, resolving application defects, and ensuring system reliability.

Senior Data Engineer – Encounter Processing & Data Platform


Job Summary
The Senior Data Engineer will serve as the primary technical owner responsible for the ongoing maintenance, enhancement, operational support, and reliability of Nebraska's enterprise data processing framework supporting Medicaid Encounter Processing and enterprise data ingestion.

This mission-critical platform processes provider, member, reference, and encounter data used by downstream Medicaid systems and reporting. The application framework is developed using Scala, Apache Spark, Hive, and Drools and operates on the Cloudera Data Platform (CDP).

Following the State's migration from an Apache open-source Hadoop environment to Cloudera Data Platform, this position will assume primary responsibility for maintaining the production code base, implementing business and regulatory changes, supporting production operations, resolving application defects, and performing routine operational administration of the data processing environment.

This position serves as the single primary technical owner of the application framework and is expected to maintain sufficient expertise across the complete technology stack—including Scala, Apache Spark, Hive, Drools, and the Cloudera Data Platform—to ensure the ongoing reliability, maintainability, and operational continuity of this mission-critical system.

Primary Ownership

·         Own the Scala/Spark application framework supporting Medicaid Encounter Processing.

·         Own the Drools business rule implementation and ongoing rule maintenance.

·         Own enterprise data ingestion processes for master data and source system files.

·         Own Spark and Hive batch processing workflows.

·         Own production issue investigation and application defect resolution.

·         Own production releases, version management, and deployment coordination.

·         Own application performance tuning and optimization.

·         Own technical documentation, operational procedures, and knowledge transfer.

·         Provide basic operational administration and monitoring of the Cloudera Data Platform while coordinating with infrastructure and cloud operations teams.

Required Qualifications

·         Minimum 7 years of experience developing enterprise-scale distributed data processing applications.

·         Strong hands-on development experience with Scala.

·         Experience developing applications using Apache Spark 2.x and/or Spark 3.x.

·         Experience developing and maintaining applications running on Cloudera Data Platform (CDP) 7.x or equivalent enterprise Hadoop distributions.

·         Experience with Apache Hive 3.x for data processing and Hive table management.

·         Experience implementing business rules using the Drools Rules Engine.

·         Strong SQL development and query optimization skills.

·         Experience supporting Linux-based production environments.

·         Experience troubleshooting distributed Spark applications in production.

·         Experience using Git and modern version control practices.

Preferred Qualifications

·         Medicaid or healthcare experience.

·         Experience with Medicaid Encounter Processing.

·         Experience with Cloudera Manager, HDFS, and YARN.

·         Experience integrating with IBM DataStage.

·         Familiarity with AWS infrastructure supporting Cloudera.

·         Experience with Agile, Jira, and Confluence.



Requirements


Type

Category

Qualification

Description

Competency

Required



Skills

Architecture

Cloudera Data

 

Advanced (7-9 Years)

Yes


Skills

Others

Distributed Data Processing

• Minimum 7 years of experience developing enterprise-scale distributed data processing applications.

Advanced (7-9 Years)

Yes


Skills

Others

Drools

Experience implementing business rules using the Drools Rules Engine

Proficient (4-6 Years)

Yes


Skills

Others

Scala

• Strong hands-on development experience with Scala.

Proficient (4-6 Years)

Yes


Skills

Others

SQL and Linux

• Strong SQL development and query optimization skills • Experience supporting Linux-based production environments

 

Yes


Skills

Others

Apache experience

•Experience developing applications using Apache Spark 2.x and/or Spark 3.x. •Experience with Apache Hive 3.x for data processing and Hive table management.

Proficient (4-6 Years)

Yes


Skills

Others

Cloudera

• Experience developing and maintaining applications running on Cloudera Data Platform (CDP) 7.x or equivalent enterprise Hadoop distributions

 

No


Similar Jobs

See all Remote Software Development jobs →

Personalize your Remote Job Search in 3 Easy Steps!

Discover remote opportunities in Data Engineer

Answer easy questions

Answer easy questions

200,000+ jobs across 15+ categories

Get your best job matches

Get your best job matches

Only hand-screened, legit jobs

Find a remote job faster

Find a remote job faster

No ads, scams, or junk

I was the first applicant for a remote marketing position that got listed on the company website the same day I applied. Had an interview within 48 hours!

Sarah J. — Sarah J. · Marketing Manager ★★★★★ Verified