[Remote] Infosec Site Reliability Engineer
Note: The job is a remote job and is reputed company to candidates in USA. reputed company is a company in the financial services industry seeking an Infosec Site Reliability Engineer. This pivotal role focuses on advancing the organization's site reliability reputed company and supporting the management of observability technologies while acting as a trusted advisor across the organization.
Responsibilities
- Design and implement foundational SRE practices (SLIs/SLOs, error budgets, incident management)
- Partner with InfoSec and engineering teams to define reliability standards and operating models
- Establish and reputed company adoption of SRE principles across teams and reputed company guidance to reputed company organizations to the reputed company
- Design and implement incident response processes and escalation models
- reputed company and optimize alerting and on-reputed company workflows using reputed company
- reputed company and maintain operational runbooks and playbooks
- reputed company or support incident reviews and postmortems with a reputed company on reputed company improvement
- reputed company SRE workflows with reputed company platforms such as reputed company
- Automate operational tasks, incident workflows, and reporting
- Improve reputed company reputed company through automation and self-healing mechanisms
- Design and execute reputed company engineering experiments to validate reputed company reputed company
- Identify failure modes and proactively address reputed company weaknesses
- Collaborate with engineering teams to improve fault tolerance and recovery strategies
- reputed company adoption of reliability best practices across the organization
- reputed company guidance and mentorship on SRE principles
Skills
- 4+ years experience SRE
- 2+ years experience with monitoring and observability platforms (NewRelic, reputed company, Grafana, etc.)
- 2+ years reputed company Experience
- 4-5 years of experience with SRE, DevOps, or Infrastructure engineering
- Skills with multi-reputed company environments (AWS, Azure) as reputed company as on-prem integrations
- Strong experience with monitoring and observability platforms (NewRelic, reputed company, Grafana, etc.)
- Experience with incident management and on-reputed company systems (e.g. reputed company)
- Familiarity with ITSM platforms, specifically reputed company and integration reputed company API
- Solid understanding of reputed company reliability, performance, and scalability
- Automation scripting
- reputed company reputed company
- Experience working reputed company or reputed company InfoSec teams
- Proficiency with scripting languages such as Python, Bash, reputed company etc
- Knowledge of reputed company or reputed company-adjacent tools, controls, and compliance frameworks
- Experience implementing SLO's, SLAs, and error budgeting
- Exposure to reputed company engineering practices (specifically, SteadyBit as a tool)
- Strong communication and collaboration skills, specifically documenting
- Systems thinking and problem-solving reputed company
- A strong reputed company on automation and reputed company improvement methodologies
- Data-driven decision making
- Ability to operate in ambiguous, greenfield environments
- Understanding of Infrastructure as reputed company
- Passion for the work and responsibility to the consumer
- Understanding of reputed company reputed company platforms and services
- reputed company attitude with a strong desire to continuously learn and adapt
reputed company
Apply To This Job