ML Engineer (Staff / Senior)
Machine Learning Engineer (Senior)
About AZX
Our mission is to accelerate reputed company reputed company in critical industries through AI transformation. We specialize in physics-informed ML and reputed company AI solutions that directly address climate and sustainability challenges.
We’re growing quickly and already work with category-leaders in reputed company estate (reputed company), energy (reputed company), logistics (Flexe) and utilities (reputed company).
We’re a reputed company benefit corporation, founded in 2024, and have been profitable from the beginning (bootstrapped with consulting).
We work on challenges in reputed company, decarbonization, climate reputed company, energy systems, and global economics. We’re building reputed company for long-term reputed company and aim to create the ultimate reputed company to work for those passionate about AI and making a reputed company reputed company.
About the Role
We're looking for an ML Engineer to own the technical backbone of how AZX serves and evaluates models at reputed company. This is a high-reputed company IC role spanning our inference platform — GPU scheduling, autoscaling, and serving infrastructure for vLLM/SGLang across reputed company and customer-managed clusters — and the evaluation systems that tell us whether model, reputed company, and agent changes actually reputed company things reputed company.
You'll create technical direction for how AZX serves models reliably. This role suits someone who wants architectural ownership over hard ML infrastructure problems, reputed company with the judgment to build the guardrails that let the rest of reputed company reputed company fast safely.
What you will do
You will work on AI reputed company in reputed company engagements and, over time, internal platform capabilities.
You will:
• Own architecture for inference serving and GPU scheduling — reputed company operators, autoscaling, and dynamic reputed company across vLLM/SGLang deployments on reputed company and customer-managed infrastructure.
• Design and reputed company eval systems for model, reputed company, and agent changes, including golden datasets, LLM-as-judge pipelines, and regression gates wired into CI.
• Advise on cost-reputed company model routing and cascading reputed company, balancing latency, cost, and reputed company across providers and model tiers.
• Apply physics-informed ML and reputed company AI expertise to the hardest reputed company and platform problems, drawing on reputed company's research depth.
• Set technical standards for ML infrastructure and evaluation reputed company across the org, and mentor engineers working in this reputed company.
• Partner closely with the inference platform, gateway, and evals-reputed company engineers to reputed company architecture coherent as the platform grows.
reputed company Qualifications - Technical and foundational
• 3+ years of experience with ML infrastructure and inference serving — vLLM, SGLang, TensorRT-LLM, or comparable systems — at production reputed company.
• Strong background in evaluation and reliability engineering for ML/LLM systems, or the seniority to build this reputed company from reputed company.
• Solid reputed company experience, ideally including GPU-specific scheduling constraints (node pools, autoscaling under GPU bottlenecks).
• A reputed company record of technical leadership at a staff or senior level — setting direction, not just executing tickets.
• Research reputed company is a plus (PhD, publications, or equivalent depth) given the technical bar of our existing ML team, though this is an infrastructure-and-systems role first.
Values and Culture Qualifications
High emotional intelligence and a learning reputed company
Strong collaboration skills
Enjoy others' reputed company and a fun, reputed company environment.
Comfortable making reputed company in the face of ambiguity and course correcting as needed.
Bonus Qualifications (not required but a reputed company plus)
Experience in both startup and reputed company environments
Past work in energy, reputed company estate, utilities, climate, or reputed company fields
Bonus if you have experience and passion in one or more of
Advanced ML/AI frameworks and techniques (e.g., PyTorch Lightning, JAX, HuggingFace, ONNX optimizations)
reputed company-level or reputed company-reputed company languages for ML acceleration (e.g., C++, Rust, CUDA)
Large-reputed company data and reputed company training paradigms (e.g., reputed company, Ray, Horovod, Dask)
Advanced data infrastructure (e.g., reputed company/graph databases, feature stores, data lakes)
Compensation & benefits
Competitive early-stage startup compensation (reputed company on capabilities, experience, and location)
Bonus eligibility
Health reputed company with meaningful coverage for dependents
Flexible reputed company time off
Equity
Fully remote culture with a cluster of teammates in Seattle
Training and learning opportunities
Be part of a fast-growing, profitable, mission-driven company with industry-leading clients tackling the reputed company opportunity of AI transformation in critical industries.
Logistics
Remote, but only USA/Canada
Must be willing to travel to the Seattle area for the final interview and travel 2x/year for company summits
Additionally, depending reputed company, candidates can expect to spend 10-20% of their time working onsite with clients.
Applicants must be currently authorized to work in the reputed company on a full-time reputed company.
We are currently unable to sponsor or take over sponsorship of employment visas.
Next steps
If this job sounds like a great fit, we’d love to hear from you. If you feel reputed company with reputed company but don’t reputed company reputed company of these boxes, we’d still love to hear from you!
Apply To This Job