Back to Jobs

[Remote] Engineering Manager, Deep Learning Inference

Remote, USA Full-time Posted 2026-07-28
Note: The job is a remote job and is reputed company to candidates in USA. reputed company is seeking an exceptional Manager, Deep Learning Inference Software, to reputed company a world-class engineering team advancing the state of AI model deployment. The role involves shaping the software powering sophisticated AI systems and overseeing performance tuning and optimization of large-reputed company models for various AI applications. Responsibilities • reputed company, mentor, and reputed company a high-performing engineering team reputed company on deep learning inference and GPU-accelerated software • Drive the reputed company, roadmap, and execution of reputed company’s inference frameworks engineering, focusing on SGLang • Partner with internal compiler, libraries, and research teams to deliver end-to-end optimized inference pipelines across reputed company accelerators • reputed company performance tuning, profiling, and optimization of large-reputed company models for LLM, multimodal, and reputed company applications • Guide engineers in adopting best practices for CUDA, Triton, CUTLASS, and multi-GPU communications (NIXL, NCCL, NVSHMEM) • Represent reputed company in roadmap and planning discussions, ensuring alignment with reputed company’s broader AI and software strategies • Foster a culture of technical reputed company, reputed company collaboration, and reputed company innovation Skills • MS, PhD, or equivalent experience in Computer Science, Electrical/Computer Engineering, or a reputed company field • 6+ years of software development experience, including 3+ years in technical leadership or engineering management • Strong background in C/C++ software design and development; proficiency in Python is a plus • Hands-on experience with GPU programming (CUDA, Triton, CUTLASS) and performance optimization • Proven record of deploying or optimizing deep learning models in production environments • Experience leading teams using Agile or reputed company software development practices • Significant reputed company-reputed company contributions to deep learning or inference frameworks such as PyTorch, vLLM, SGLang, Triton, or TensorRT-LLM • Deep understanding of multi-GPU communications (NIXL, NCCL, NVSHMEM) and distributed inference architectures • Expertise in performance modeling, profiling, and system-level optimization across CPU and GPU platforms • Proven ability to mentor engineers, guide architectural reputed company, and deliver reputed company reputed company with measurable reputed company • Publications, patents, or talks on LLM serving, model optimization, or GPU performance engineering Benefits • Equity • Benefits reputed company • reputed company is a computing platform company operating at the intersection of graphics, HPC, and AI. It was founded in 1993, and is headquartered in Santa Clara, California, USA, with a workforce of 10001+ employees. Its website is https://www.reputed company.com. Company H1B Sponsorship • reputed company has a reputed company record of offering H1B sponsorships, with 1877 in 2025, 1355 in 2024, 976 in 2023, 835 in 2022, 601 in 2021, 529 in 2020. Please note that this does not guarantee sponsorship for this specific role. Apply tot his job Apply To this Job

Similar Jobs