# Senior Data Platform Engineer

**Company**: Komodo Health
**Location**: New York, NY
**Work arrangement**: hybrid
**Experience**: senior
**Job type**: full-time
**Salary**: $196,000-$230,000 USD
**Category**: Engineering
**Industry**: Healthcare

**Apply**: https://job-boards.greenhouse.io/komodohealth/jobs/8807352002?utm_source=yubhub.co&utm_medium=jobs_feed&utm_campaign=apply
**Canonical**: https://yubhub.co/jobs/job_33d529e6-246

## Description

At Komodo Health, our mission is to reduce the global burden of disease through smarter use of data.

Join Komodo Health's Data Foundations team as a Senior Data Platform Engineer and play a critical role in shaping the core data products that fuel our Healthcare Map.

The Opportunity at Komodo Health: As a Senior Data Engineer, you will design, build, operate, and improve large-scale data pipelines and foundational data products that power Komodo’s Healthcare Map, analytics products, and downstream AI/ML-enabled use cases.

Key Responsibilities:

- Build, operate, and optimize large-scale production data pipelines using Python, SQL, Airflow, cloud infrastructure, and distributed processing frameworks.

- Transform massive healthcare claims, EHR, and reference datasets into trusted, performant Healthcare Map data products and serving-ready data assets.

- Strengthen pipeline reliability through data quality checks, validation, lineage, observability, monitoring, and alerting.

- Debug complex data, system, and performance issues across computationally intensive workflows.

- Partner with Data Product Quality, Product, Platform, and Engineering teams to translate healthcare data needs into scalable technical solutions.

- Contribute to system design, architecture, code quality, testing, documentation, CI/CD, and rotational production support.

- Enable downstream analytics, product, and AI/ML use cases through high-quality, well-modeled, reliable data.

What you bring to Komodo Health:

- Healthcare data experience across claims, clinical, RWE, provider, patient, or life sciences datasets, including coding systems such as ICD-10, CPT, NDC, or NPI.

- Strong hands-on experience building, operating, and debugging production-grade data pipelines at scale.

- Advanced Python and SQL skills, with experience in Airflow or similar workflow orchestration tools.

- Experience with Spark or comparable distributed data processing frameworks.

- Proven experience designing and operating data solutions in AWS.

- Strong instincts for data quality, reliability, root-cause analysis, and production troubleshooting.

- Ability to communicate technical trade-offs clearly and collaborate with engineering, product, and data partners.

- Comfort using AI-assisted engineering tools for productivity, debugging, documentation, and technical exploration.

Additional skills and experience we’d prioritize (nice to have):

- Experience delivering external-facing data products through customers, APIs, serving layers, or production access patterns.

- Ability to optimize high-scale data architectures for performance, cost, versioning, and large-volume productization.

- Experience applying AI or agentic workflows to engineering, data quality, delivery, or operations.

- Success in high-growth or ambiguous environments that require balancing architecture, speed, and quality.

The pay range for this role is: San Francisco Bay Area and New York City: $196,000-$230,000 USD All Other US Locations: $170,000-$200,000 USD

## Skills

### Required
- Python
- SQL
- Airflow
- Spark
- AWS
- data quality
- reliability
- root-cause analysis
- production troubleshooting

### Nice to have
- Experience delivering external-facing data products
- Ability to optimize high-scale data architectures
- Experience applying AI or agentic workflows

---

Source: [Apply at job-boards.greenhouse.io](https://job-boards.greenhouse.io/komodohealth/jobs/8807352002?utm_source=yubhub.co&utm_medium=jobs_feed&utm_campaign=apply)
