Backend Engineer — Data Platform
Savant Bio · 1 Pennsylvania Plaza, 54th floor, New York, NY 10119 · United States · Remote
Pay: USD 175,000 – 250,000 a year
Posted Aug 27, 2026
Sign up free: we match you to jobs like this, tailor your application and fill the form. 2 free applications every day.
About Us
Savant is transforming how healthcare and life sciences organizations unlock the value trapped in unstructured medical data. Our platform combines cutting-edge large language models (LLMs) with domain-specific quality controls to convert free-text clinical records into structured, analysis-ready data — efficiently, accurately, and at scale. We work with leading institutions across healthcare, life sciences, and research, supporting faster studies, sharper insights, and better care. Backed by Roivant (NASDAQ: ROIV), Savant is built for organizations that see structured data not just as an output, but as a foundation for innovation.
The Opportunity
We’re hiring a Backend Engineer to help mature and scale the data platform underlying Savant’s clinical data products. You’ll help design and implement the systems that move large, complex datasets through ingestion, processing, quality control, and delivery.
This is a backend engineering role with a strong data engineering and infrastructure orientation.
You’ll work across Python and Rust services, SQL-based data systems (DuckDB, Postgresql, and Ducklake), Dagster pipelines, and Kubernetes workloads. You’ll help establish the architectural patterns that allow Savant to process sensitive healthcare data efficiently while maintaining reproducibility, traceability, and quality.
Savant is based in NYC, and this role is remote-eligible. Preference will be given to candidates able to come to work in NYC.
Some Things You Might Work On
Design and build production pipelines for ingesting, transforming, validating, and delivering healthcare datasets
Develop backend services and APIs that support Savant’s clinical data products
Build reliable Dagster workflows with strong patterns for partitioning, retries, idempotency, backfills, and observability
Design data models and contracts that support schema evolution, lineage, reproducibility, and downstream analysis
Operate and improve containerized workloads running on Kubernetes
Build systems for monitoring pipeline health, data quality, performance, and infrastructure cost
Diagnose production issues across application, data, and infrastructure layers
Improve the performance of Python and SQL workloads operating over large datasets
Establish testing and deployment practices that make data systems safer and easier to change
Collaborate with AI engineers to productionize new model workflows and scale them across customer datasets
Help shape technical architecture as Savant’s platform, customer base, and processing volume grow
What We’re Looking For
5+ years of experience building production backend or data-intensive software systems
Excellent Python and SQL skills
Experience with AWS and Kubernetes
Sound technical judgment, taste, and the ability to make pragmatic architectural decisions
Experience designing and operating data pipelines at scale using Dagster, Airflow, Prefect, or a comparable orchestration…