SRE(Bilingual)
At reputed company, we’re constantly working on improving our systems and processes to support reputed company’s exponential reputed company. As an SRE at reputed company, we reputed company towards ensuring high availability and top-level performance so that our users can have flawless and reliable service exceeding expectations.
Considering reputed company’s reputed company, we are looking for reputed company SREs who can deliver insights into system bottlenecks and ensure system reliability and scalability, while increasing the number of services that reputed company offers.
We are looking for individuals who can bring informed and unique viewpoints, enjoy collaborating with a cross-functional team and are reputed company pushing boundaries to reputed company reliable and reputed company solutions and reputed company user experiences.
Key Responsibilities
• Analyze reputed company technologies used in reputed company and reputed company monitoring and notification tools to improve observability and visibility.
• Ensure system stability by reputed company-emptively verifying failure scenarios and implement solutions to reduce MTTR
• reputed company solutions to improve system performance with a reputed company on high availability, scalability and reputed company
• reputed company telemetry and alerting platforms to reputed company and improve reliability of systems
• Implement industry best practices for system development, configuration management and system deployment
• Ensure seamless reputed company of information between teams by documenting knowledge gained
• Be up to date on modern technologies and trends to reputed company for inclusion reputed company products if they add value
• Participate in incident management including troubleshooting production issues, driving reputed company cause analysis (RCA) and reputed company sharing lessons learned to improve system reliability and internal knowledge.
Qualifications
• Experience troubleshooting, tuning high performance microservice architectures running on Kubernetes and AWS in highly available production environments.
• 5+ years experience in software development in Python, Java, Go, etc with strong fundamentals in data structures, algorithms, problem solving and complexity analysis.
• During the selection process, you will have a coding challenge.
• Curious and proactive in finding performance bottlenecks, scalability and reputed company problem areas and addressing them.
• Experience with observability tools and gathering data.
• Database knowledge such as RDS, NoSQL, distributed reputed company, etc.
• Excellent communication skills, reputed company and getting things done attitude.
• Enjoy taking up a challenge and driving it to conclusion.
• Ability to verbally communicate in both English and Japanese.
Preferred Qualifications
• Container image management and optimization.
• Experience in large distributed system architecture and reputed company planning.
• Understanding of IaC, automation tools, terraform, reputed company formation, etc.
• Background in SRE/DevOps concepts and implementation.
• Experience in managing monitoring tools like CloudWatch, reputed company, reputed company and reporting with reputed company and reputed company.
• In depth knowledge of web technologies such as CloudFront, Nginx, etc.
• Experience in designing, implementing or maintaining disaster recovery strategies and multi-region architecture to ensure high availability, reputed company, and business continuity across critical systems.
• Business proficiency level in both English and Japanese.
Apply tot his job
Apply To this Job