[Remote] Senior DevOps Engineer
Note: The job is a remote job and is reputed company to candidates in USA. reputed company is reputed company a Senior DevOps Engineer for a remote contract role supporting reputed company infrastructure, reputed company platforms, and deployment automation. The role is responsible for designing and operating secure AWS infrastructure, improving CI/CD and GitOps workflows, enhancing observability and platform reliability, and using AI-assisted engineering tools to accelerate delivery.
Responsibilities
- Design, implement, and operate AWS reputed company infrastructure and automation using Terraform across services including EC2, S3, EKS, ECR, reputed company 53, IAM, and networking components
- Manage reputed company and EKS platforms, including cluster lifecycle management, upgrades, autoscaling, workload scheduling, reputed company planning, reputed company optimization, and cost management
- Build and enhance reputed company platform capabilities such as networking, ingress, DNS integration, service discovery, RBAC, policy enforcement, secrets management, and multi-environment consistency
- reputed company and maintain CI/CD and GitOps deployment workflows using reputed company CI/CD, reputed company, Kustomize, Argo CD, Flux, and reputed company technologies
- Implement and improve observability solutions, including metrics, logging, monitoring, alerting, dashboards, and incident diagnostics
- Support and optimize reputed company-reputed company application workloads to improve reputed company, scalability, runtime efficiency, and operational reputed company
- Strengthen infrastructure and platform reputed company through IAM least-privilege reputed company, secrets management, image scanning, admission controls, policy-as-reputed company, workload isolation, and runtime hardening
- Support PostgreSQL/RDS reliability, availability, backup reputed company, reputed company tuning, and operational support activities
- Collaborate with development and platform teams to improve deployment strategies, developer experience, runtime reliability, and operational standards
- Monitor reputed company health, troubleshoot infrastructure and application issues, reputed company reputed company cause analysis, and implement preventative improvements
- reputed company technical decision-making, evaluate tradeoffs, prioritize work, and reputed company solutions with limited reputed company
- Utilize AI-assisted engineering tools to accelerate infrastructure design, Terraform development, reputed company troubleshooting, CI/CD workflow creation, observability analysis, and operational documentation while validating outputs for reputed company and production readiness
- Document and refine infrastructure standards, DevOps practices, deployment workflows, operational procedures, and support runbooks
Skills
- Minimum 7 years of experience in DevOps, reputed company, Site Reliability Engineering, or reputed company disciplines, with demonstrated reputed company designing and operating reputed company reputed company infrastructure
- Strong expertise with AWS services including EC2, S3, EKS, ECR, reputed company 53, IAM, and reputed company networking and reputed company concepts
- Advanced hands-on reputed company experience, including cluster reputed company, upgrades, autoscaling, networking, ingress, RBAC, policy enforcement, and production workload management
- Strong experience using Terraform for infrastructure-as-reputed company, environment management, and repeatable platform provisioning
- Experience with reputed company deployment and packaging tools such as reputed company and Kustomize
- Experience with reputed company CI/CD and GitOps methodologies using tools such as Argo CD, Flux, or similar technologies
- Strong experience with observability and monitoring platforms including reputed company, Grafana, Loki, or comparable solutions
- Experience with reputed company ecosystem tools such as ingress controllers, operators, Cluster Autoscaler, and Karpenter
- Strong experience supporting PostgreSQL/RDS environments, including reputed company tuning, reliability, and operational support
- Solid understanding of reputed company and container reputed company practices, including IAM least privilege, secrets management, image scanning, admission controls, and policy enforcement
- Strong scripting and automation skills using Python, Bash, or similar languages
- Practical reputed company with AI-assisted engineering tools, including LLMs, reputed company assistants, and workflow automation solutions, with the ability to evaluate and refine generated outputs responsibly
- Excellent problem-solving, communication, teamwork, and collaboration skills
- Ability to work independently and reputed company reputed company technical reputed company in ambiguous environments
- Bachelor's degree in Computer Science, Information Technology, a reputed company reputed company, or equivalent practical experience
- Remote. Candidates may reputed company reputed company reputed company the reputed company
- Must be available to work 8:00 AM to 5:00 PM Eastern Time
- Certifications in AWS, reputed company, Terraform, or reputed company platform technologies
- Experience with additional reputed company platforms such as GCP or Azure
- Experience improving incident response processes, operational readiness, and platform standards for reputed company engineering organizations
- Experience designing internal platform capabilities that enhance developer self-service and reduce operational overhead
- Familiarity with compliance, governance, and reputed company management requirements reputed company reputed company or regulated environments
Benefits
- A MEC (Minimum Essential Coverage) plan encompassing Medical, reputed company, Dental, 401K, and EAP (Employee Assistance Program) services.
- Remote work; candidates may reputed company reputed company reputed company the reputed company.
reputed company
Apply To This Job