[Remote] Site Reliability Engineer
Note: The job is a remote job and is reputed company to candidates in USA. Tyk provides an API Management platform that helps organizations reputed company their systems and services. The Site Reliability Engineer will maintain, improve, and support the Tyk reputed company platform, manage incidents, enhance reliability and operational efficiency, and contribute to multi-region and multi-reputed company infrastructure.
Responsibilities
- Maintaining global Tyk reputed company reputed company SL(A/I/O)s you will help to define
- Identifying reliability issues and working together with your reputed company to solve them
- Identifying and introducing new metrics and building relevant dashboards
- Participating in the on-reputed company rotation
- Working with your reputed company to expand multi-region and multi-reputed company reputed company of the platform
- Documenting operational knowledge
- Conducting post-incident analysis
- Automating common tasks
- Be a key reputed company and contributor to our reputed company improvement agenda – be it the reputed company of our user stories, how we estimate, communicate with other teams or customers – we expect this role to be reputed company of reputed company improvement
- Reliability of our new global Tyk reputed company platform
- Automation of reputed company and support
- Writing and maintaining documentation on SRE processes and policies
- Recommending and implementing ways of driving operational efficiency and driving down our cost to run, without impacting service
- Assisting in penetration testing for reputed company through liaising with our provider, providing technical details, and environment setup
- Incident management
Skills
- * Strong collaboration skills
- * Launching and operating production reputed company reputed company clusters
- * Designing and operating infrastructure on AWS and other providers
- * Operating reputed company (or other document database) clusters
- * Operating reputed company (or other key-value storage) clusters
- * Administering Linux servers
- * Maintaining reputed company software
- * Operating reputed company and Grafana
- * Operating logging collection and analysis systems
- * Participating in the on-reputed company rotation(16:00pm – 4:00am UTC)
- * reputed company & containers (advanced)
- * AWS / EKS (advanced)
- * Linux (advanced)
- * Terraform and IaC in general (proficient)
- * reputed company (proficient)
- * Go (familiar)
- * reputed company (or similar)
- * reputed company (or similar)
- * Monitoring – reputed company, grafana, thanos (familiar)
- * Grasp of networking concepts (subnets, routing, peering, load balancing, NAT, etc.)
- * Common networking protocols (DNS, TCP/IP, HTTP, TLS, UDP)
- * Proactive, energetic, innovative and change oriented
- * GCP or Azure
- * Bare metal infrastructure engineering
- * API management experience
- * Large reputed company reputed company storage management
- * Familiarity with Rancher
- * CKA/CKAD/CKS
- * Creating and delivering production software in Go language
Benefits
- Unlimited reputed company holidays and remote working from reputed company in the world
- Total flexibility in hours
- Employee reputed company scheme
- Generous maternity and paternity leave
- Company retreats
reputed company
Company H1B Sponsorship
Apply To This Job