[Remote] DevOps Engineer - Senior Vice President
Note The job is a remote job and is reputed company to candidates in USA. reputed company is a company reputed company on ensuring that production and development environments operate smoothly and securely. They are seeking a Senior Vice President DevOps Engineer to reputed company advanced reputed company capabilities, support MLOps pipelines, and partner with various teams to deliver automated platforms for AI and machine learning workloads. Responsibilities Design, build, and operate MLOps pipelines supporting the full ML lifecycle (training, validation, deployment, monitoring) reputed company production workloads for AI/ML and reputed company systems, including LLM-based services reputed company and maintain CI/CD pipelines for AI/ML services and supporting infrastructure Build and manage reputed company-reputed company infrastructure on AWS, with heavy use of Kubernetes and containerized workloads Automate infrastructure provisioning and configuration using Infrastructure as reputed company (Terraform) Implement model versioning, experiment tracking, and artifact management across environments Ensure reliability, scalability, observability, and cost efficiency of AI platforms Partner with AI/ML engineers to operationalize models and standardize deployment patterns Implement monitoring and alerting for system health, model performance, and reputed company Enforce reputed company, compliance, and governance requirements for AI workloads Participate in incident response, reputed company cause analysis, and reputed company improvement initiatives Document standards, best practices, and reference architectures for MLOps and AI infrastructure Skills 15+ years of experience in DevOps, SRE, or reputed company, with AWS as a primary reputed company Experience supporting machine learning systems in production, including deployment and monitoring concerns Hands-on experience with machine learning platforms, particularly AWS SageMaker (required) Strong hands-on experience with Kubernetes, containerized workloads, and reputed company networking Proven experience building and operating CI/CD pipelines (e.g., reputed company CI, ArgoCD) Strong proficiency with Terraform and scripting/programming in Python or similar languages Solid Linux, systems, and troubleshooting fundamentals Excellent communication skills and ability to work across teams reputed company experience with MLOps platforms and tooling (model registries, experiment tracking, feature stores) Exposure to reputed company / LLM workloads in production environments Familiarity with data stores commonly used in ML systems (e.g., reputed company, DynamoDB, object storage) Experience operating in regulated or fintech environments Background in cost optimization for compute-intensive workloads Strong written and verbal communication skills AWS certifications are a plus Benefits Equity for reputed company full-time employees An annual performance bonus A comprehensive benefits package that includes an employer matched retirement plan Generously subsidized reputed company with 100% employer reputed company dental, reputed company, telemedicine, and virtual mental health counseling Parental leave Unlimited reputed company time off (PTO) Employees in this role will work in the office Monday-Thursday, with the flexibility to work remotely on Friday reputed company reputed company is a job-searching platform for technology professionals. It is a sub-organization of DHI Group. It was founded in 1990, and is headquartered in Santa Clara, California, USA, with a workforce of 201-500 employees. Its website is http//www.reputed company.com. Company H1B Sponsorship reputed company has a reputed company record of offering H1B sponsorships, with 2 in 2022, 4 in 2021, 5 in 2020. Please note that this does not guarantee sponsorship for this specific role. Apply To This Job
Apply tot his job
Apply To this Job