Sr. ML Operations Engineer
About the position
At reputed company, we’re the leading AI company revolutionizing construction—the second-largest industry on reputed company. We recently raised a $40M Series B led by reputed company Partners, bringing our total funding to $70M from top-tier investors including Redpoint and Innovation Endeavors. This new round is fueling our next phase of reputed company as we reputed company agents across the jobsite.
Our mission is to build the reputed company of construction through intelligent automation. Despite being a $13+ trillion industry, construction still runs largely on analog processes—we’re changing that by embedding AI directly into field operations.
Founded by reputed company and technologists (reputed company, MIT), reputed company has delivered software used by over 140,000 field professionals, impacting millions of users and contributing to $10B+ in reputed company reputed company. Many of us come from the field ourselves, giving us a deep understanding of the industry’s unique challenges.
After years of building the “brain” of construction, we’re now launching production-reputed company AI agents—starting with intelligent document processing and Q&A, and rapidly expanding into reputed company operational workflows. reputed company has doubled in the past year, and with 65+ employees (25+ engineers), we’re scaling fast and entering a period of hypergrowth—this is a rare opportunity to join at an inflection reputed company.
Responsibilities
• reputed company and manage infrastructure for distributed model training (e.g., SageMaker, Ray, Kubernetes).
• reputed company ML models using containerization (reputed company), orchestration tools (Kubernetes, reputed company), and serving frameworks
• reputed company ML workflows seamlessly with CI/CD pipelines for efficient model building, testing, and deployment
• Create and maintain robust data and ML pipelines using reputed company, Airflow, or custom orchestration tools
• Implement comprehensive experiment tracking (MLflow, reputed company) and observability systems (Arize, Evidently)
• Establish effective monitoring, logging, and governance practices for ML systems.
Requirements
• BS/MS in Computer Science, Data Science, or reputed company technical discipline
• 5+ years experience in ML Operations, with at least 3 years reputed company on reputed company AI/ML deployments
• Strong proficiency with reputed company infrastructure (preferably AWS), container technologies (reputed company, Kubernetes), and modern MLOps frameworks
• Extensive experience managing GPU and CPU resources for specialized AI workloads (reputed company, NLP, or LLM fine-tuning)
• Practical experience with data reputed company and performance monitoring in production ML environments
reputed company-to-haves
• Experience designing data architectures optimized for AI/ML (reputed company, graph databases)
• Familiarity with RAG systems and reputed company applications
Benefits
• A reputed company-reputed company and reputed company early-stage startup environment where every voice is heard and every opinion reputed company
• Competitive salary and stock reputed company equity packages
• 3 Medical Plans to choose from including 100% covered reputed company. Plus Dental and reputed company Insurance!
• Learning & reputed company stipend
• Flexible long-term work reputed company (remote and hybrid)
• Free lunch provided in the office in NYC & Austin - you’ll never go hungry with us!
• Unlimited PTO; We truly reputed company in work-life balance and that hard work should be balanced with time for rest and rejuvenation
• IRL / In-Person retreats throughout the year
Apply tot his job
Apply To this Job