Staff ML Platform Engineer @Datavant
Artificial Intelligence
Salary usd 224,000 - 2..
Remote Location
πŸ‡ΊπŸ‡Έ USA Only
Employment Type full-time
Posted 2d ago

[Hiring] Staff ML Platform Engineer @Datavant

2d ago - Datavant is hiring a remote Staff ML Platform Engineer. πŸ’Έ Salary: usd 224,000 - 280,000 per year πŸ“Location: USA

Role Description

We are looking for a Staff ML Platform Engineer to join our Data & ML Platform organization on the ML Platform team. We own the paved road that takes a model from notebook to production without rebuilding it each time: training and inference across SageMaker and Databricks, self-service LLM-endpoint serving, and the pipelines behind Datavant’s clinical-AI products. Data Science owns the models and their quality; we own the pipelines, serving infrastructure, and operational guardrails that let those models run safely against regulated health data.

As a Staff ML Platform Engineer, you will be a technical leader across the ML Platform:

  • Setting direction for the paved road
  • Owning the hardest architectural problems
  • Moving the team from bespoke plumbing toward a coherent platform that Data Science teams can self-serve

AI fluency is a baseline expectation. You should already be using Claude Code, Cursor, Copilot, or equivalent tools as a core part of your daily engineering workflow, have opinions about how they make a team faster, and know how to apply them responsibly when PHI and other sensitive data are in scope.

What You Will Do

  • Set technical direction across ML training, serving, and observability
  • Be the final escalation point for the most elusive infrastructure problems (GPU capacity, Spark tuning, production incidents)
  • Own and evolve our paved-road framework (the shared CI/CD spine, model-workflow scaffolding, and Databricks Asset Bundles)
  • Lead architecture for LLM-endpoint serving across managed providers (Databricks, AWS, Snowflake) and self-hosted deployments
  • Own the standards and tooling for MLflow, model registry, training image supply chain, and observability across training and inference
  • Partner closely with your Data & ML Platform teammates to present a cohesive ML platform to the Data Science, App Dev, and Operations teams at Datavant
  • Serve as a key technical input to vendor and platform selection decisions across model providers, ML tooling, and observability
  • Mentor senior engineers on the team, provide technical guidance to platform consumers, and stay hands-on writing high-leverage code and Infrastructure-as-Code alongside your teammates

Qualifications

  • 10+ years of software engineering experience
  • 3+ years designing, evolving, and operating enterprise-scale ML platforms in production
  • Strong technical judgment under ambiguity
  • Track record of setting standards, influencing peers, and raising the bar across teams
  • Hands-on production experience with Databricks and/or Amazon SageMaker
  • Experience with MLflow (or an equivalent tracking + registry system)
  • Fluency in Java (or a JVM equivalent) and Python
  • Real depth in Apache Spark for large-scale data and distributed compute
  • Real depth in AWS: networking, IAM, GPU compute, and the storage and messaging services this role touches
  • Fluency with Terraform, containers, Kubernetes, and GitHub-based CI/CD for ML workloads
  • Direct experience serving LLMs in production
  • AI-native working style: daily use of Claude Code, Cursor, Copilot, or equivalent
  • Clear written and verbal communication, especially in async, remote settings

What Helps You Stand Out

  • Prior technical leadership on a healthcare or regulated-industry ML platform (HIPAA, HITRUST, SOC 2, or equivalent)
  • Direct experience with Databricks Asset Bundles, Unity Catalog, and open table formats (Iceberg, Delta)
  • Production experience with specialized inference pipelines beyond generic model serving (e.g., document understanding, computer vision, or streaming inference)
  • GPU capacity planning at strategic, tactical, and operational horizons
  • Experience with real-time and event-driven inference and streaming platforms (Kafka, Kinesis)
  • Evaluation and red-teaming experience for clinical or safety-sensitive AI
  • Background contributing to open-source ML infrastructure or publishing on production ML systems

Benefits

  • Competitive salary range: $224,000 β€” $280,000 USD
  • Commitment to a diverse team and high-performance culture
  • Equal Employment Opportunity employer
Before You Apply
️
πŸ‡ΊπŸ‡Έ Be aware of the location restriction for this remote position: USA Only
β€Ό Beware of scams! When applying for jobs, you should NEVER have to pay anything. Learn more.
Staff ML Platform Engineer @Datavant
Artificial Intelligence
Salary usd 224,000 - 2..
Remote Location
πŸ‡ΊπŸ‡Έ USA Only
Employment Type full-time
Posted 2d ago
Apply for this position
Did not apply βœ“
Applied βœ“
Sent Follow-Up βœ“
Interview Scheduled βœ“
Interview Completed βœ“
Offer Accepted βœ“
Offer Declined βœ“
Application Denied βœ“
Unlock 130,000+ Remote Jobs
️
πŸ‡ΊπŸ‡Έ Be aware of the location restriction for this remote position: USA Only
β€Ό Beware of scams! When applying for jobs, you should NEVER have to pay anything. Learn more.
Apply for this position
Did not apply βœ“
Applied βœ“
Sent Follow-Up βœ“
Interview Scheduled βœ“
Interview Completed βœ“
Offer Accepted βœ“
Offer Declined βœ“
Application Denied βœ“
Unlock 130,000+ Remote Jobs
Γ—
Apply to the best remote jobs
before everyone else

Access 130,000+ vetted remote jobs and get daily alerts.

4.9 β˜…β˜…β˜…β˜…β˜… from 500+ reviews

⚑ 131,518+ remote jobs, refreshed hourly

πŸ”” Real-time alerts: Apply first

πŸ›‘οΈ Vetted companies, no scams, true remote only

Unlock All Jobs Now

Maybe later