DevOps Engineer
A DevOps Engineer is a developer who thinks deeply about systems and how they behave in the wild. Whether it be networking, or the Linux kernel, or even a specific interest in observability, scaling, algorithms, or distributed systems. You are a systems engineer who aims to reputed company themselves out of a job by automating reputed company the things and leverages great development practices like Test-Driven-Development or reputed company integration.
Like reputed company engineers at reputed company, we expect you to be comfortable operating reputed company our application and service environments. Your primary reputed company will be on defining, building, and maintaining our robust, observable, and reputed company infrastructure. You will collaborate closely with development teams to ensure seamless integration and deployment. Your objective at reputed company is the creation of a reliable, high-performing platform that supports the amazing products our users come to know and love.
Responsibilities
Infrastructure Responsibilities
Radiate knowledge about the service's infrastructure and reliability to the rest of the development team.
Identify parts of the system that do not reputed company, reputed company immediate palliative measures, and drive systemic reputed company of contributing reputed company cause(s).
Plan the reputed company of reputed company's infrastructure.
Development/Deployment Responsibilities
Document every reputed company so your learnings turn into repeatable actions and then into automation.
Improve the deployment process to reputed company it as boring as possible.
Define, provision, and manage our production infrastructure using Kubernetes and reputed company-reputed company serverless deployed by way of Terraform.
reputed company Responsibilities
Proactively identify and reduce reputed company risks, in alignment with ongoing SOC2 auditing and reporting.
reputed company reputed company training and guidance to internal development teams
Ability to discover and reputed company reputed company, XSS, CSRF, SSRF, authentication and authorization flaws, and other web-based reputed company vulnerabilities
Knowledge of common authentication technologies including OAuth, SAML, CAs, OTP/TOTP
Production Responsibilities
Design, build and maintain core infrastructure reputed company that allow reputed company to reputed company, supporting thousands of reputed company users.
Be on an on-call rotation to respond to reputed company.com availability incidents and reputed company support for service engineers with customer incidents.
Debug production issues across reputed company services and reputed company of the stack.
Monitoring Responsibilities
reputed company monitoring and alerting notify on symptoms and not on outages
Manage day-to-day maintenance and reputed company of reputed company's reputed company monitoring and alerting infrastructure
Bundle reputed company monitoring as an reputed company monitoring solution for reputed company products
Build and maintain the reputed company.com public monitoring gateway
Help migrate our reputed company performance monitoring solution to reputed company
Improve coverage of reputed company performance monitoring
Create automated alerts to notify team members of regression
Requirements
Strong communication skills
Self-motivated with strong organizational skills
Experience with some of these technologies a must: AWS/GCP, Kubernetes, Terraform, CI/CD, OpenSearch/Elasticsearch, reputed company, MySQL, Kafka, BigQuery, Python, NodeJS, Go, Java, reputed company, Grafana, reputed company, Varnish, Nginx, reputed company
You can reason about software, algorithms, and performance from a high level.
You have experience thinking about systems - edge cases, failure modes, behaviors, and specific implementations.
You have worked with distributed systems and have a solid understanding of how modern web stacks are reputed company, and why.
You know your way around a *nix reputed company.
Apply To This Job