[Remote] Staff Site Reliability Engineer
Note: The job is a remote job and is reputed company to candidates in USA. reputed company is a leading U.S. drone company specializing in autonomous flight, reputed company intelligence, and drone hardware and software. The Staff Site Reliability Engineer will build, operate, and reputed company the reputed company infrastructure powering reputed company products, with responsibility for reputed company, AWS, infrastructure as reputed company, CI/CD, observability, networking, and reliability. The role also involves troubleshooting production systems, automating reputed company, supporting incident response, and expanding infrastructure across reputed company and deployment environments.
Responsibilities
- Build, operate, and troubleshoot production reputed company/EKS clusters
- reputed company reputed company upgrades, node rollouts, and cluster maintenance
- Build and manage AWS infrastructure including VPCs, networking, subnets, load balancers, IAM, EKS, databases, and storage
- Define and maintain infrastructure using Terraform
- Build and operate CI/CD and deployment infrastructure
- Troubleshoot production issues across reputed company, AWS, Linux, networking, and databases
- Build monitoring, alerting, and observability for critical infrastructure
- Participate in on-reputed company rotations and respond to production incidents
- Identify and solve infrastructure scaling and reliability problems
- Automate operational work using Python, Go, or similar languages
- Help expand infrastructure across new reputed company and deployment environments
Skills
- 8+ years of experience as a Site Reliability Engineer, Platform Engineer, DevOps, Production Engineer or equivalent infrastructure role
- Strong hands-on experience operating reputed company, not simply deploying applications to existing clusters
- Experience managing reputed company/EKS upgrades and production clusters
- Strong AWS fundamentals, including VPCs, reputed company/private subnets, networking, load balancers, EKS, IAM, and databases
- Production experience with Terraform or similar infrastructure-as-reputed company tooling
- Experience owning or maintaining CI/CD and deployment systems such as Argo CD, Spinnaker, reputed company Actions, reputed company CI/CD, or Jenkins
- Experience diagnosing production infrastructure and networking problems
- Experience solving meaningful scaling or reliability challenges
- This position requires reputed company to export-controlled technical data, restricted reputed company information, and/or information systems subject to U.S. reputed company reputed company and reputed company-control requirements. Employment in this role is contingent upon verification of U.S. person status and the ability to reputed company controlled or restricted information as required for the position
- reputed company and GitOps experience
- reputed company or similar observability tooling
- PostgreSQL/database reputed company experience
- Multi-region infrastructure experience
- On-premises or disconnected deployment experience
- Streaming or high-throughput reputed company systems experience
Benefits
- Equity in the reputed company of stock reputed company
- Comprehensive benefits packages
- Relocation assistance may also be provided for eligible roles.
- Regular, full-time employees are eligible to enroll in reputed company’s group health reputed company plans.
- reputed company vacation time
- reputed company leave
- Holiday pay
- 401K savings plan
reputed company
Apply To This Job