Data Engineer
Remote
Full Time
#Engineering
#Healthcare
#Analytics
#Metaflow
#Spark
#AWS
#EMR
#Docker
#Kubernetes
#SQL
#NoSQL
#DynamoDB
#Elasticsearch
At Rad AI, we’re transforming healthcare through artificial intelligence. Founded by a radiologist, we’ve built one of the largest proprietary radiology report datasets in the world. Our technology has already helped uncover hundreds of new cancer diagnoses and reduced error rates in tens of millions of reports by nearly 50 percent. With more than $140 million raised-including a recent $68 million Series C that values the company at $528 million-we’re backed by investors such as Khosla Ventures and Transformation Capital. Our generative AI tools are now used daily by thousands of radiologists, supporting more than one-third of radiology groups and nearly half of all medical imaging in the United States.
What is this role?
We are hiring a Senior Data Engineer to join our engineering team on a full-time, remote basis. The position calls for someone with senior-level experience who can take ownership of large-scale data systems that power our AI products.
What will you do?
- Design and build scalable data architecture using pipeline tools such as Metaflow and distributed processing frameworks like Spark, ensuring the platform remains flexible and efficient as data volumes grow.
- Establish and evolve internal standards for code style, maintenance, and best practices that keep our high-scale data platform reliable and easy to manage.
- Partner with researchers and other teams to understand their data requirements for model training and production monitoring, then deliver solutions that meet those needs while maintaining data quality, integrity, and security.
What makes you a great fit?
You bring at least five years of hands-on data engineering experience and have designed, built, and maintained distributed data pipelines that handle large-scale datasets. You are comfortable working with big-data technologies on AWS, including Amazon EMR and AWS Batch, and you have production experience running Spark workloads. You have orchestrated complex workflows with Metaflow and worked with both SQL and NoSQL stores such as DynamoDB, Elasticsearch, and PostgreSQL. Containerization with Docker and Kubernetes is second nature to you, and prior software engineering experience is a plus. Strong communication skills in English help you collaborate effectively across teams. Experience in a HIPAA-compliant environment, early-stage startups, or healthcare and machine-learning projects is welcomed but not required.
What's in it for you?
We offer comprehensive medical, dental, vision, and life insurance, along with HSA (with employer match), FSA, and DCFSA options. You will also receive a 401(k) plan, 11 paid company holidays, a flexible PTO policy, an annual company-wide offsite, periodic team offsites, and an annual equipment stipend.




