See how much of this job your resume covers, and what’s missing.
Want a recruiter to go through it line by line?
Get professional reviewQuestions interviewers often ask for this role, with sample answers.
Upload your resume and we draft a letter for this exact role, tailored to what it asks for.
Set the technical direction and roadmap for Clarium’s AWS data platform, owning reliable batch and near-real-time pipelines, warehouse modeling, and performance across Postgres and Snowflake. Establish data engineering standards and governance, mentor engineers, and partner across teams to turn ambiguous needs into durable data models.
The healthcare industry overspends on its supply chain by over $25B each year, the result of fragmented data, inefficient workflows, and wasted supplies. Clarium is fixing that. Our AI-powered platform, Astra OS, gives hospitals end-to-end visibility into their supply chain operations, automating workflows and surfacing actionable insights so supply chain teams can focus on what matters most: patient care. We're trusted by some of the world's leading health systems, including Yale New Haven Health, Stanford, Geisinger, and Kaiser Permanente.
Founded in 2020, Clarium has raised $43M in total funding. Our Series A was led by Northzone, with participation from General Catalyst, AlleyCorp, Kaiser Permanente Ventures, Texas Medical Center Ventures, and 1984 Ventures.
This is a staff-level individual contributor role for an engineer who sets technical direction for our data platform rather than working ticket to ticket. You'll decide how data gets modeled, moved, and trusted across Clarium, and you'll leave both the systems and the engineers around you measurably better.
You'll work across our transactional databases and analytical warehouse, own the pipelines that connect them, and partner directly with the analytics, product, and engineering teams who depend on that data. Much of the work is supply chain data arriving from hospital ERP systems, where schemas and semantics vary by source and are rarely well documented, so a large part of the job is deciding what should be built, not only building it.
Own the architecture and technical roadmap for core data infrastructure on AWS, spanning ingestion, transformation, storage, and serving layers
Design, build, and operate reliable batch and near-real-time pipelines with clear SLAs and the observability to back them up
Model supply chain data from external ERP systems into a coherent, reusable warehouse model, and lead the migration of legacy assets toward it
Tune performance and cost across Postgres and Snowflake, including query plans, indexing, partitioning, warehouse sizing, and storage strategy
Establish engineering standards for data work (testing, code review, CI/CD, data quality checks, documentation) and mentor engineers through design reviews, pairing, and honest technical feedback
Partner with analytics, product, and engineering teams to turn ambiguous questions into durable data models instead of ad hoc extracts
Lead governance for sensitive data, including lineage, access controls, retention, auditability, and de-identification where required
8+ years of data or software engineering experience, including significant time owning systems end to end in production
Expert-level SQL: you write complex analytical queries as a matter of course and can read a query plan and explain why something is slow
Strong Python for production data work, including pipeline code, transformation logic, testing, and tooling
Deep experience with relational databases, including schema design, normalization tradeoffs, transactions, and performance tuning (Postgres and Snowflake strongly preferred)
Hands-on experience with data pipeline and orchestration tooling such as Airflow, Dagster, Prefect, dbt, Fivetran, Spark, or Kafka; we care more about depth and judgment than an exact stack match
Production experience running data workloads on AWS (e.g., S3, RDS, Lambda, ECS), with an understanding of the cost, security, and networking implications of how you build
A track record of leading multi-quarter, cross-team initiatives that depended on teams you don't manage, and comfort with the ambiguity of deciding what should be built
Experience with healthcare data (claims, EHR/EMR, HL7 or FHIR, ICD-10, CPT) or working under HIPAA with PHI, de-identification, and audit requirements
Familiarity with supply chain data from ERP systems such as Oracle, Workday, or Lawson, including where it tends to be unreliable
Infrastructure-as-code and deployment automation (Terraform, CDK, CI/CD for data infrastructure)
Streaming and event-driven architectures, or data quality and observability tooling
Need to Know: SQL · Python · Postgres · Snowflake · AWS (S3, RDS, Lambda, ECS)
Nice to Know: Terraform · Github Actions · EKS / Kubernetes · dbt · Prefect · Apache Kafka
Target Base Salary Range: $170K - $210K
The base salary Clarium offers may vary depending upon the ultimate scope and responsibilities of the position and on the candidate's job-related knowledge, skills, and experience. The total package will include equity, in addition to a full range of medical and/or other benefits, depending on the position offered. Pay and benefits are subject to change at any time, consistent with the terms of any applicable compensation or benefit plans.
Incentive Stock Options proportionate to your salary
Fully remote, with a NYC co-working space available; distributed team across multiple time zones with opportunities for in-person time
Unlimited PTO
Top-tier health, vision, and dental benefits
401K
The opportunity to build on a strong foundational team with deep data and engineering roots at a stage where your work genuinely shapes the product
A fast-paced, high-growth environment where we move quickly to solve critical healthcare challenges
The chance to empower frontline healthcare workers by automating the administrative friction in their day-to-day
Collaborative work alongside a high-caliber team of clinicians, engineers, and revenue leaders
Equal Opportunity Statement
Clarium is committed to promoting an inclusive work environment free of discrimination and harassment. We value a diverse and balanced team where everyone can belong.
Stop the endless job search. Our AI finds and applies to the best jobs for you.
Featuring 216,660+ Jobs in Data Engineer
Answer easy questions
216,660+ jobs across 15+ categories
Get your best job matches
Only hand-screened, legit jobs
Find a remote job faster
No ads, scams, or junk
“I was the first applicant for a remote marketing position that got listed on the company website the same day I applied. Had an interview within 48 hours!”