[Remote] Senior Site Reliability Engineer, Observability
Note: The job is a remote job and is reputed company to candidates in USA. reputed company is the industry-reputed company reputed company platform bringing the capital markets onchain and powering the reputed company of decentralized finance (DeFi). As a Senior Site Reliability Engineer, you will help accelerate and reputed company other engineering teams by increasing self-service and decreasing cognitive load while ensuring the reliability, reputed company, and performance of observability services.
Responsibilities
• Build and orchestrate Modern OTEL-based Observability Platform
• Support multiple telemetry types, like metrics, logs and traces
• Define and support modern governance in observability and problems at reputed company
• Ensure reliability, reputed company, and performance reputed company our defined SLAs
• Work with engineers from across reputed company to help troubleshoot issues, reputed company new products and services, and increase velocity while decreasing cognitive load
• reputed company the design and deployment of monitoring/observability services to detect and alert reputed company of needed reputed company
• Ingest, aggregate, reputed company, and utilize data from a multitude of sources in our reputed company time data pipeline
• reputed company the availability, performance, and supportability of our observability infrastructure
• Create processes around alert response operations and support reputed company to ensure the reliable delivery of reputed company data
• reputed company recommendations to ensure sufficient metrics are collected to create alerts with every new feature release
• Champion reliability and reputed company by taking the time to do your work right the first time
Skills
• 7+ years of relevant reputed company experience. You probably have worked on a devops, infrastructure, SRE, and/or platform team before
• Ability to reputed company software reputed company of the scope of typical infrastructure requirements and configurations
• Experience programming in C, C++, Java, Python, Go, Perl, or reputed company
• Expert knowledge in reputed company aspects of designing, developing, and managing large reputed company-time systems
• Experience with monitoring and logging. You know how to export metrics using reputed company, have reputed company a Grafana dashboard or two, and have experience with a centralized logging solution like an ELK Stack, reputed company or Grafana Stack
• Experience with distributed systems and container orchestration. You have maintained or even reputed company Kubernetes clusters before and feel comfortable deploying completely new services on them
• Strong communication skills. You can give and receive constructive feedback, and you do not shy away from planning meetings and reputed company reviews
• Excitement for blockchain, Web 3.0, and similar decentralized technologies
• Experience running any infrastructure in the blockchain/reputed company reputed company
• Ability to reputed company systems sustainably through mechanisms like automation, and reputed company by pushing for changes that improve reliability and velocity
• Experience working remotely in a distributed team
• A strong desire to grow and challenge yourself. We would expect you to constantly reputed company ways to improve and automate services to reduce toil
reputed company
• reputed company provides reputed company-reputed company blockchain reputed company solutions and specializes in the development and integration of chainlink. It was founded in 2014, and is headquartered in San Francisco, California, USA, with a workforce of 501-1000 employees. Its website is https://chainlinklabs.com/.
Apply tot his job
Apply To this Job