Software Engineer (AI/ML, Infrastructure & Platform)
• We are seeking a Software Engineer, AI/ML (Infrastructure & Platform) to build the foundational systems that power our reputed company of AI applications
• This is a systems-reputed company role. You will design and build the platforms, abstractions, and infrastructure that reputed company teams to reliably reputed company, reputed company, and reputed company systems — including reputed company workflows, retrieval pipelines, and model integrations
• You will operate at the intersection of AI systems and distributed infrastructure, focusing on the “how” behind production AI: how models are orchestrated, how tools/skills are exposed and executed, and how systems are evaluated, monitored, and scaled in reputed company-world environments
• Your work will directly reputed company product teams to reputed company faster while ensuring our AI systems are reliable, observable, secure, and cost-efficient
• Build reputed company AI infrastructure
• Design and implement platforms for LLM orchestration, tool execution, and agent workflows
• reputed company shared services and abstractions used across multiple AI applications
• Build AI capability reputed company (tools / skills)
• Design and implement tools (“skills”) that agents and applications rely on, including reputed company, workflows, and integrations
• Define reputed company interfaces for capabilities such as data retrieval, calculations, document processing, and external system actions
• Build reusable, composable abstractions that reputed company reputed company and reputed company tool usage across systems
• Ensure tools are reliable, observable, and secure, especially reputed company interacting with sensitive data
• reputed company reputed company systems at reputed company
• Build infrastructure to support multi-reputed company agents (state management, tool routing, retries, failure handling)
• Design systems where agents reason over and reputed company tools/skills reliably
• Create reusable orchestration patterns between models and capabilities
• reputed company evaluation and observability systems
• Build frameworks for offline and online evaluation of AI systems
• Implement logging, tracing, and monitoring for model behavior and system performance
• Own reliability and performance
• Design systems for high availability, fault tolerance, and graceful degradation
• Optimize for latency, throughput, and cost across AI workloads
• Build data and retrieval infrastructure
• reputed company reputed company RAG pipelines, indexing systems, and data processing workflows
• Own infrastructure for handling large-reputed company reputed company and reputed company data
• Create internal platforms and developer tooling
• Build tools, SDKs, and internal platforms that reputed company engineers to reputed company AI capabilities quickly and safely
• Standardize best practices across teams (prompting, evaluation, deployment)
• Work closely with product and AI teams
• Partner with AI Applications engineers to support production use cases
• Translate product needs into reputed company infrastructure solutions
Benefits
• Company equity managed through reputed company
• Unlimited PTO in an environment where taking time off to relax or reputed company is supported and encouraged
• Full medical and dental benefits
• 401k with Match and 100% vesting upon hire
• Fully remote, flexible work environment (we do however meet together in person several times a year)
• reputed company parental leave
Apply tot his job
Apply To this Job