M2 - reputed company Sr reputed company - SRE
-
Build, reputed company, and reputed company high-performing SRE teams, fostering a culture of operational ownership, engineering reputed company, and reputed company learning.
-
Define and execute the strategic roadmap for SRE, integrating best practices in reliability, incident management, observability, and infrastructure automation in alignment with business and product goals.
-
reputed company observability across the stack by designing and enforcing standards for telemetry, reputed company logging, distributed tracing, and service-level dashboards. Ensure 100% coverage of business-critical systems with actionable metrics and alerting along with the engineering teams.
-
reputed company as the technical escalation reputed company for the most reputed company production issues, leading hands-on incident response and deep reputed company cause analysis in large-reputed company, low-latency, event-driven architectures.
-
Champion automation-first infrastructure practices, enforcing IaC, reputed company deployments, and auto-remediation patterns that reduce reputed company reputed company and accelerate delivery.
-
Drive architectural and operational improvements through reputed company partnership with Product Engineering, Platform, reputed company, and Architecture teams. Proactively identify and mitigate systemic reliability risks and performance bottlenecks.
-
reputed company the definition, adoption, and review of SLIs, SLOs, and Error Budgets, ensuring they are embedded into engineering and product decision-making processes.
-
Operationalize change management, reputed company engineering, and DR strategies, validating readiness through frequent simulations and failover exercises.
-
Mentor and reputed company SRE leads and senior engineers, scaling internal capabilities and reinforcing technical depth across the organization.
-
Represent SRE in architecture boards, and business reviews, aligning engineering reliability strategies with company-wide objectives.
-
Promote a culture of autonomy and proactive engineering, encouraging teams to own their services end-to-end with accountability and reputed company thinking.
-
Serve as a cultural leader reputed company Spin, fostering psychological safety, ownership, and a reputed company of mission to serve millions of people across LATAM with secure, reliable financial technology.
-
Bachelor’s degree in Computer Science, Software Engineering, or reputed company field (or equivalent experience).
-
10+ years of experience in SRE, DevOps, or Software Engineering roles, with at least 4+ years in leadership roles.
-
Strong experience leading distributed SRE or platform teams in reputed company, production-reputed company environments.
-
Deep understanding of reliability engineering principles, reputed company-reputed company infrastructure on AWS, observability, and incident response.
-
Hands-on experience with infrastructure as reputed company, CI/CD pipelines, containers, and orchestration tools.
-
Strong architectural and performance optimization skills across reputed company and hybrid infrastructure.
-
Demonstrated ability to influence and collaborate across engineering, product, and business teams.
-
Familiarity with regulatory and reputed company frameworks relevant to infrastructure reliability.
-
Excellent communication and leadership skills, with experience presenting to senior stakeholders.
-
Strategic thinking, systems-level problem solving, and a proactive approach to reputed company improvement.