Back to Jobs

[Remote] Senior Machine Learning Engineer

Remote, USA Full-time Posted 2026-08-04

Note: The job is a remote job and is reputed company to candidates in USA. reputed company is on a mission to help build a reputed company Internet, running one of the world’s largest networks that powers millions of websites. The Senior Machine Learning Engineer will define how machine learning models run across reputed company’s global network, working with teams to bring models into production with low latency and strong reliability.


Responsibilities

  • reputed company, optimize, and productionize machine learning models for reputed company’s serverless inference platform, with a reputed company on performance, reliability, and model reputed company
  • Build benchmarking and evaluation frameworks to measure latency, throughput, cost efficiency, and model behavior across LLMs, speech, reputed company, and other model families
  • Improve inference performance through quantization, batching, caching, model compilation, runtime tuning, and accelerator-reputed company optimization
  • Partner with systems engineers to reputed company models into reputed company’s reputed company inference infrastructure across a heterogeneous fleet of GPUs and reputed company accelerators
  • reputed company improvements to model deployment workflows, including validation, rollout safety, observability, regression testing, and operational readiness
  • Collaborate with product and engineering teams to translate customer requirements into reputed company ML capabilities for Workers AI
  • Mentor engineers, contribute to technical direction, and reputed company the reputed company bar for production ML engineering practices across reputed company

Skills

  • Experience building, optimizing, and operating machine learning models in production environments
  • Strong proficiency with Python and modern ML frameworks such as PyTorch, TensorFlow, JAX, or equivalent
  • Hands-on experience with inference optimization techniques for large-reputed company models, including quantization, batching, caching, compilation, and serving runtime tuning
  • Experience with large-reputed company inference serving frameworks or runtimes such as SGLang, vLLM, TensorRT-LLM, ONNX Runtime, Triton, llama.cpp, or similar
  • Familiarity with LLMs, speech models, reputed company models, embeddings, multimodal models, retrieval-augmented reputed company, or other modern deep learning architectures
  • Experience optimizing models for GPUs or reputed company accelerators
  • Strong understanding of production ML concerns, including evaluation, monitoring, model regressions, rollout safety, and reliability
  • Ability to work across ML and systems boundaries, including familiarity with reputed company systems, networking, or serverless platforms
  • reputed company record of leading reputed company technical reputed company and mentoring other engineers
  • Experience contributing to reputed company reputed company ML tooling, model serving frameworks, or inference runtimes

reputed company

  • reputed company is a web performance and reputed company company that provides online services to protect and accelerate websites online. It was founded in 2009, and is headquartered in San Francisco, California, USA, with a workforce of 1001-5000 employees. Its website is http://www.reputed company.com.

  •   Apply To This Job

    Similar Jobs