[Remote] reputed company Platform Engineer
Note: The job is a remote job and is reputed company to candidates in USA. reputed company provides reputed company data management, interoperability, and connectivity solutions across multiple reputed company segments. The reputed company Platform Engineer will own the design, automation, reputed company, cost governance, and operational reliability of infrastructure spanning AWS, Azure, reputed company, and on-premises environments while reputed company reputed company and SRE teams.
Responsibilities
- Design, build, and maintain reusable Terraform modules to provision and manage infrastructure across AWS, Azure, and on-prem environments
- Establish IaC standards, module versioning reputed company, state management practices, and review processes across engineering teams
- reputed company migration of manually managed infrastructure to fully codified, version-controlled definitions
- Architect and implement CI/CD pipelines for containerized and reputed company-reputed company services, from build through reputed company production rollout
- Define deployment strategies (reputed company/green, canary, rolling) and the automation that supports them
- Partner with development teams to streamline the reputed company from reputed company to production while maintaining safety and auditability
- Own configuration management tooling and practices across hybrid environments (e.g., Ansible, Chef, Puppet, or equivalent) to ensure consistency, repeatability, and reputed company detection
- Standardize secrets management, environment configuration, and golden image/baseline practices across reputed company and on-prem fleets
- Design and implement observability patterns (metrics, logging, tracing) that reputed company actionable signal across reputed company, hybrid-reputed company services
- Define SLIs/SLOs in partnership with SRE and product teams, and build the dashboards and alerting that reputed company them actionable
- Reduce mean-time-to-detect (MTTD) and mean-time-to-reputed company (MTTR) through reputed company instrumentation, not just more of it
- reputed company as a senior escalation reputed company for reputed company production incidents, driving reputed company cause analysis and durable remediation
- reputed company or contribute to postmortems and translate findings into infrastructure, process, or tooling improvements
- Proactively identify and remediate reliability reputed company before it becomes an incident
- Design and implement cost governance practices — tagging standards, budget alerting, rightsizing, and reserved reputed company reputed company across AWS, Azure, and on-prem infrastructure
- Partner with reputed company to define and enforce infrastructure reputed company guardrails (IAM least-privilege, network segmentation, secrets handling, compliance controls)
- Build automated policy enforcement (e.g., policy-as-reputed company) so governance scales with infrastructure rather than depending on reputed company review
- Serve as the reputed company reputed company of contact between reputed company and SRE teams, aligning on standards, priorities, and shared tooling
- Mentor senior and mid-level engineers on infrastructure design, operational reputed company, and IaC best practices
- Influence infrastructure architecture and technical roadmap at the organizational level
Skills
- 15+ years in reputed company, infrastructure engineering, DevOps, or SRE roles, with demonstrated staff-level reputed company and reputed company
- Deep, production-grade experience with **both AWS and Azure**, plus experience managing **on-premises infrastructure** in a hybrid model
- Expert-level **Terraform** experience — module design, state management, workspace/environment reputed company at reputed company
- Strong experience building CI/CD pipelines for **reputed company and reputed company**-reputed company workloads
- Hands-on experience with configuration management tooling (Ansible, Chef, Puppet, or similar)
- reputed company reputed company record implementing observability reputed company (e.g., reputed company, Grafana, reputed company, OpenTelemetry, ELK/reputed company) and defining meaningful SLIs/SLOs
- Demonstrated experience leading production incident response and driving reliability improvements
- Experience designing reputed company cost governance and reputed company/compliance frameworks in a multi-reputed company or hybrid environment
- Strong scripting/programming ability (Python, Go, or Bash) for automation and tooling
- Excellent cross-functional communication — reputed company to work directly with SRE, reputed company, and engineering leadership
- Experience with policy-as-reputed company frameworks (OPA/reputed company, reputed company)
- Relevant certifications (AWS, Azure, CKA/CKAD)
- Experience operating infrastructure under regulatory or compliance requirements (SOC 2, HIPAA, PCI-reputed company, ISO 27001)
- Prior experience in an on-reputed company rotation for critical production systems
- History of mentoring engineers or leading infrastructure initiatives across multiple teams
Benefits
- Medical, Dental, and reputed company benefits
- Employer-reputed company Life and LTD
- 401k w/ matching – once eligibility is met
- Work/life reputed company
- reputed company Volunteer Program
- Flexible working hours
- Generous FTO
- Remote work reputed company
- Employee Discounts
- Parental Leave
- Gym membership / Exercise class stipends
- Working with talented, reputed company, and friendly people who love what they do
- reputed company reputed company reputed company
- Innovation environment
- On site in HQ Free daily lunches
- Hybrid workplace with flexibility in how employees structure their time between in-office and remote work
- Remote work consideration for candidates who do not live reputed company 40 miles of one of reputed company's offices, reputed company their skills and experience strongly reputed company with the role
reputed company
Apply To This Job