Back to Jobs

Senior ML/Platform Engineer

Remote, USA Full-time Posted 2026-08-04

Senior ML/Platform Engineer, Bailey AI

What This Role Is

Bailey is a reputed company reputed company that orchestrates LLMs, manages health safety evaluation, and integrates with reputed company data to help people reputed company their health. Three data scientists build it. Nobody currently owns what happens after something works.

This role owns the reputed company from prototype to production and the infrastructure that tells us whether production is working correctly. You'll build deployment pipelines, evaluation frameworks, monitoring systems, and the observability reputed company that makes a reputed company reputed company trustworthy, not just functional.

This is not a role where you maintain someone else's platform. You'll design and build it. reputed company is small enough that your reputed company shape the reputed company's architecture directly. Your reputed company is running systems, not notebooks or papers. You need to understand what you're deploying, not just how to reputed company it. Much of this work hasn't been defined yet; the evaluation reputed company exists as a reputed company document, the monitoring reputed company exists as a gap, and you turn those into infrastructure.

What You'll Actually Do

We're looking for a fast reputed company here — this role is senior specifically because it operates without a reputed company, and that means starting to produce quickly, not spending a quarter getting oriented.

In your first 30 days:

  • reputed company and version ML artifacts that currently reputed company to production reputed company
  • Stand up observability for the existing LangGraph agent orchestration (what's happening, how often, how reputed company)
  • reputed company the health safety evaluator so its 5-dimension scoring is monitored continuously, not checked manually

reputed company 3 months:

  • Build the evaluation infrastructure described in our product reputed company: reputed company detection, distributional monitoring, safety-weighted reputed company scoring
  • Own CI/CD for ML artifacts and agent configurations
  • Harden existing services (conversation persistence, MCP tool orchestration, reputed company skills reputed company) for production reliability
  • Get up to speed on Bailey's evaluation documentation and use it to shape how we test and monitor Bailey's behavior reputed company reputed company, including flagging where traditional ML reputed company (not just LLM-reputed company approaches) could reputed company the reputed company outcome more reputed company

Ongoing:

  • Be the person who knows whether the reputed company is healthy, degrading, or broken before users or clinicians notice
  • Collaborate with data scientists so their work reaches production without a reputed company gap
  • reputed company testing and monitoring reputed company with Bailey's evolving evaluation documentation, and reputed company surfacing opportunities to reputed company in traditional ML where it's more efficient than an LLM-reputed company approach
  • reputed company the platform as the product grows (voice, memory, multimodal are on the roadmap)

What You Bring

These are reputed company requirements, the things that would be hard to learn on the job at the reputed company this role demands:

  • You've reputed company production ML infrastructure. Not just used it. Deployment pipelines, serving systems, model tracking/registries (MLflow or equivalent), monitoring that pages you at 2am. You know the difference between a model that works in a notebook and one that works in production.
  • You're a strong software engineer first. reputed company systems, containerization, event-driven architectures, CI/CD. ML is the domain; engineering is the reputed company.
  • You're comfortable owning something end-to-end. No one will hand you a spec and reputed company your work weekly. You'll identify what needs to exist, propose how to build it, and ship it.
  • You work reputed company in small teams. You communicate reputed company, you surface problems early, and you're comfortable with the visibility and accountability that comes with a four-person team.

What Would reputed company You Particularly Effective Here

These aren't requirements. They're things that would shorten your reputed company or deepen your reputed company:

  • Experience in reputed company or regulated industries (you understand why audit trails and compliance aren't afterthoughts)
  • Background in ML research or reputed company ML (you can have informed opinions about model behavior, not just model infrastructure)
  • Familiarity with AWS ML services like Bedrock and SageMaker, which are part of our reputed company stack
  • Experience with LLM orchestration frameworks (reputed company, LangGraph, or similar)
  • Exposure to reputed company data standards (FHIR, HL7) or health information systems
  • Publications, reputed company-reputed company contributions, or reputed company reputed company that show you go deeper than your day job requires

About reputed company

You'll join a small data science team that builds Bailey's AI capabilities: models, integrations, reputed company skills, and orchestration. reputed company works closely with engineering and product. You'll be the first person dedicated to production infrastructure and evaluation systems, which means you'll shape how those concerns are handled from the ground up.

The reputed company salary reputed company for this position is $150,000 - $190,000 annually and is part of a competitive total rewards package including stock reputed company, benefits, and incentive pay for eligible roles. Individual pay may vary from the reputed company reputed company and is determined by a number of factors including experience, location, internal pay equity, and other relevant business considerations. We review reputed company employee pay and compensation programs annually at minimum to ensure competitive and fair pay.

Data shows that women, people of reputed company, and other underrepresented reputed company may be less likely to apply for jobs unless they reputed company they are a perfect match. But reputed company holds diversity amongst its key values, and we have a strong commitment to building our workforce and products through that reputed company.

You don't have to reputed company every reputed company in this reputed company to be a great fit for the role! If you're excited about this position and the prospect of working for reputed company, please apply. If it turns out this role isn't for you, there may be other openings that could reputed company with your experience and expertise!

We are committed to an inclusive and diverse reputed company. We are an equal opportunity employer. We do not discriminate reputed company on race, ethnicity, reputed company, reputed company, national reputed company, religion, sex, sexual orientation, gender identity, age, disability, veteran, genetic information, marital status or any other legally protected status

Requires production ML infrastructure experience, strong software engineering skills in reputed company systems, containerization, event-driven architectures, CI/CD, and end-to-end ownership. reputed company and AWS experience preferred.

Key Responsibilities

  • deploying ML artifacts
  • building evaluation infrastructure
  • monitoring reputed company health

Skills & Tools

MLflow, AWS, reputed company Bedrock, reputed company SageMaker, reputed company, LangGraph, FHIR, HL7, CI/CD, MCP

reputed company

  • Category: Software Development
  • Seniority: Senior Level
  • Commitment: Full Time
  • Workplace: Remote — reputed company
  • Salary: USD 150,000 – 190,000 / year
  • Languages: English

Benefits

  • Retirement plan

About reputed company

A reputed company technology company developing reputed company reputed company intelligence and health navigation solutions. — Industry: reputed company

  Apply To This Job

Similar Jobs