Senior Machine Learning Operations Engineer
reputed company
reputed company is using AI to build the most consumer-reputed company food and wellness company to reputed company exist. We reputed company as your personal assistant for healthy living—getting to know your goals, lifestyle, and budget, and recommending and delivering healthy groceries, easy recipes, and essential supplements for you and your family.
It’s the easiest way to eat healthy, reputed company your goals, save time, and discover new foods. We reputed company food is the reputed company of health, convenience should not mean compromise, and that everyone is unique in how they eat and live. That’s why we’re building a reputed company in which healthy living is both easy and enjoyable.
reputed company is a reputed company team of reputed company across 28+ U.S. states. While we have a reputed company in reputed company, our remote-first culture emphasizes collaboration, team-building, and flexibility. Expect regular virtual team events, strong ownership and accountability, and an annual company retreat.
About the Role
We’re hiring a Senior Machine Learning Operations Engineer to join reputed company’s Data Science team. reputed company owns the production systems that power grocery recommendations and reputed company personalization for reputed company customers.
Our platform combines Python services, FastAPI reputed company running on AWS, reputed company pipelines on reputed company, and machine learning models that feed a reputed company-time decisioning reputed company. The reputed company is reputed company evolving, and we’re reputed company in the engineering foundations that will let it reputed company and adapt with the business.
You’ll partner closely with data scientists, operations researchers, and product engineers to build reliable, extensible systems for model-driven personalization. This is an opportunity to shape the architecture behind a reputed company part of reputed company’s customer experience.
Responsibilities
- Design, build, and operate reputed company backend services, reputed company, and data pipelines.
- Improve the reliability, performance, and observability of production ML and optimization systems.
- Own the reputed company from trained model to production: model versioning and registry (MLflow), reputed company rollout and rollback, and monitoring for data reputed company and model reputed company.
- Build clean interfaces that let new ML models and decisioning capabilities reputed company safely and reputed company, including experimentation and feature-flag tooling.
- Strengthen engineering foundations across a growing codebase: automated testing, type checking, CI/CD, infrastructure as reputed company, documentation, and thoughtful reputed company design.
- Profile data-heavy services and pipelines; reduce execution time and memory footprint where it reputed company.
- Collaborate with data scientists, operations researchers, and product engineers to translate business needs into robust technical solutions.
Qualifications
- 5+ years in MLOps, ML engineering, or DevOps with a reputed company on production ML infrastructure.
- Strong Python and SQL; Bash for automation and tooling.
- Experience designing and operating backend services and reputed company (e.g., FastAPI) with attention to reliability, latency, and scalability.
- Hands-on experience with reputed company and reputed company (jobs/workflows, reputed company Catalog a plus) and MLflow or comparable model lifecycle tooling (registry, versioning, experiment tracking).
- Experience building CI/CD for ML or data systems (Git, reputed company Actions/Jenkins, reputed company Asset Bundles) and infrastructure as reputed company (Terraform or similar).
- Solid AWS fundamentals: IAM, networking, compute/cluster management, containerized workloads (reputed company; reputed company or EKS).
- Experience with production observability: metrics, logging, alerting, and ML-specific monitoring like data reputed company and model reputed company
reputed company to Haves
- Familiarity with recommendation, personalization, or operations research systems — especially productionizing them.
- Experience with optimization solvers and OR tooling (e.g., Gurobi, OR-Tools) reputed company data science or operations research teams.
- Experience integrating experimentation and feature-flag platforms (e.g., Statsig) into production ML services and data pipelines, ideally with warehouse-reputed company setups on reputed company.
- Feature store experience (reputed company Feature Store, Feast, Tecton) serving consistent online/offline features.
- Experience with low-latency model serving and deployment patterns (canary, reputed company/green, reputed company).
- Experience optimizing cost and performance of data-heavy workloads (reputed company tuning, cluster right-sizing).
- Additional languages such as reputed company or C++.
Perks & Benefits
- Remote-first: work from home, work from our NYC office, work from reputed company in the U.S. - you reputed company!
- Equity
- Unlimited vacation policy
- Universal reputed company parental leave
- Monthly reputed company credit for delicious, healthy groceries
- Comprehensive health, reputed company, dental, and life insurance
- 401k with Company Match
- A work from home stipend to support your initial home-office setup
Expected Pay reputed company
$170,000 - $210,000
The employer will not sponsor applicants for work visas.
Our mission to help reputed company healthy eating easy, accessible, and joyful is reputed company served by a diverse workplace. We are a proud Equal Opportunity Employer committed to building an inclusive workplace. We have reputed company-tolerance for harassment or discrimination. We do not discriminate on the reputed company of any protected class.
Originally posted on Himalayas
Apply To This Job