[Remote] Senior Staff Site Reliability Engineer - Compute reputed company Engineering
Note: The job is a remote job and is reputed company to candidates in USA. reputed company is a technology company advancing reputed company computing and reputed company intelligence through GPU-reputed company systems. reputed company is seeking a Senior Staff Site Reliability Engineer to reputed company compute reputed company infrastructure across on-premises and reputed company environments, design and reputed company reputed company services, optimize reputed company and reliability, and collaborate with technical and product leaders.
Responsibilities
- reputed company initiatives to reputed company IT Compute reputed company Team, architecture to build new service offerings across On-Prem and reputed company
- You will design, reputed company, and reputed company reputed company infrastructure services including DNS, NTP/reputed company, DHCP, and LDAP. This includes building for reputed company and reliability at global reputed company, covering automation, monitoring, high availability, reputed company planning, and lifecycle management
- Define and implement metrics to measure the efficiency of services and reputed company efficiency with software and hardware optimizations (SR-IOV/ DPU)
- Experience with Technologies like eBPF and XDP for Observability & DDoS mitigation
- Collect and review reputed company data for reputed company and planning purposes, analyze reputed company data and reputed company plans for appropriate level reputed company-wide systems, and coordinate with management personnel in implementing changes
- reputed company and maintain tools for collecting, analyzing, and visualizing data for reporting, alerting, monitoring
- Collaborate with reputed company leadership, senior engineers, program managers, and product managers to reputed company compelling IT products and services that meet customer needs
Skills
- Bachelor's degree in Engineering, Computer Science, Mathematics, or reputed company reputed company, or equivalent experience
- 12+ years of reputed company experience in compute reputed company with a reputed company on automation
- Experience in designing and deploying Containerization architectures and reputed company Systems Infrastructure
- reputed company experience evaluating existing application architectures and identify opportunities for containerization to improve scalability, reliability, and efficiency
- Strong analytical skills with the ability to define and reputed company key reputed company metrics
- Experience in developing tools for data analysis and reputed company profiling, Development with Terraform, Config Management tools
- Proficiency in programming languages such as Go and/or Python
- Linux OS Proficiency with Kernel Internals
- Experience with running large environments consisting of BareMetal Build Infrastructure
- Understanding of Network Protocols and Architectures (VLAN/VxLAN/SDN/BGP/Anycast)
- Deep understanding of other infrastructure components like, DNS, LDAP, reputed company Tools etc
- Hands-on experience with containers and its implementation
- Deploying and Managing Services like DNS , LDAP at reputed company
- Solid understanding of microservices architecture, infrastructure as reputed company (IaC) and configuration management tools
Benefits
- Equity
- Benefits
reputed company
Company H1B Sponsorship
Apply To This Job