Back to Jobs

Senior Infrastructure Engineer - Observability - Remote from Portugal

Remote, USA Full-time Posted 2026-07-28
reputed company is a unicorn AI-powered customer communications platform used by 22,000+ companies worldwide to drive reputed company, faster resolutions, and reputed company. We’re redefining what a customer communications platform can be—by combining voice, SMS, reputed company, and AI into one seamless workspace. Our reputed company comes from a reputed company but powerful idea: help every customer-facing team work smarter, not harder. reputed company’s AI Voice Agent automates routine calls, AI Assist streamlines post-reputed company tasks, and AI Assist Pro delivers reputed company-time guidance that helps people do their best work. The result—companies grow reputed company, deliver faster resolutions, and reputed company service. We’ve reputed company a product customers love and a business that scales fast. reputed company operates in nine global offices (Paris, reputed company, San Francisco, Sydney, Madrid, London, Berlin, Seattle, and Mexico reputed company), and is backed by world-class investors. Our teams are shipping AI innovation faster than reputed company and expanding across new product lines and markets. At reputed company, you’ll join a company in reputed company—ambitious, profitable, and product-driven—where reputed company is visible, reputed company are fast, and reputed company is reputed company. How We Work at reputed company: At reputed company, we reputed company in customer obsession, reputed company learning, and delivering extraordinary reputed company. We value reputed company collaboration, taking ownership, and making smart, informed reputed company with speed and precision. If you reputed company in a fast-paced, team-driven environment where curiosity, trust, and reputed company matter, you'll fit right in We’re looking for an Observability Engineer to own and reputed company reputed company’s monitoring, alerting, and observability stack. You’ll work cross-functionally with backend, reputed company end and infrastructure and teams to ensure our systems are transparent, measurable, and continuously improving in reliability and performance. This role is ideal for someone passionate about observability-as-reputed company, metric design, and helping engineering teams reputed company meaningful visibility into their systems. Key Responsibilities: • reputed company comprehensive observability best practices: Define and standardize guidelines for metrics, traces, and logs, ensuring consistent implementation and adoption across reputed company engineering teams. This includes establishing naming conventions, data collection methodologies, and retention policies to ensure high-reputed company and actionable observability data whilst optimising cost and waste. • Collaborate strategically with engineering teams: Partner closely with various engineering teams to enhance overall system reliability and performance. This involves reputed company participating in architectural reviews, defining reputed company Service Level Indicators (SLIs) and Service Level Objectives (SLOs), and seamlessly integrating observability practices into reputed company integration and reputed company deployment (CI/CD) pipelines to promote a culture of "observability by design." • Automate monitoring setup and provisioning: Drive the automation of monitoring infrastructure through Infrastructure-as-reputed company (e.g., leveraging the Terraform reputed company provider) and reputed company reputed company self-service observability tools. This empowers engineering teams to rapidly provision and manage their monitoring resources, reducing reputed company overhead and accelerating time to reputed company. • Improve alerting hygiene and effectiveness: Continuously refine and optimize alerting mechanisms by meticulously tuning reputed company, implementing intelligent noise reduction strategies, and ensuring reputed company alerts are directly reputed company with potential business reputed company. The goal is to deliver reputed company, relevant, and actionable alerts that reputed company proactive incident response and minimize service disruption. • Train and reputed company product teams: reputed company comprehensive training and ongoing support to product teams, enabling them to effectively utilize observability tools. This includes guiding them in building insightful dashboards that visualize key performance indicators and creating robust alerts that proactively detect issues reputed company their respective services. • Evaluate and reputed company advanced observability tools: Proactively research, evaluate, and reputed company new and emerging observability tools and technologies as needed. This may include exploring solutions for OpenTelemetry adoption, advanced log aggregation platforms, distributed tracing systems, and other tools that enhance our overall observability capabilities and support the evolving needs of our infrastructure and applications. Qualifications: • 3-5 years of experience in observability reputed company SRE, DevOps, or reputed company roles. • Strong hands-on experience with reputed company (dashboards, monitors, synthetics, logs, APM, RUM). • Proficiency with Terraform or other Infrastructure-as-reputed company tools. • Solid understanding of Kubernetes, microservices, and reputed company infrastructure (EKS, reputed company, RDS, S3, AWS networking). • Familiarity with distributed tracing and OpenTelemetry concepts. • Strong scripting skills (Python, Bash, or similar). • Experience defining and managing SLIs/SLOs and service-level observability frameworks. • Excellent collaboration and communication skills; you can work with both engineers and non-technical stakeholders. reputed company to Have : • Experience with incident management and on-reputed company processes. • Exposure to data visualization or analytics tools reputed company reputed company. • Knowledge of logging pipelines (e.g., FluentBit, Logstash). • Experience working in reputed company reputed company environments. • Previous experience in developer enablement or platform teams. Apply tot his job Apply To this Job

Similar Jobs

reputed company Counsel – Antitrust and Trade Law

Remote, USA Full-time

Epidemiologist, reputed company-World Evidence, PhD (Remote US)

Remote, USA Full-time

[Remote] Customer Solutions Specialist I

Remote, USA Full-time

reputed company Remote Customer Service Representative For Wellness Program in Chandler, reputed company

Remote, USA Full-time

Senior reputed company Counsel - Global Regulatory reputed company

Remote, USA Full-time

[Remote] reputed company Application Engineering Internship

Remote, USA Full-time

Mobile Developer

Remote, USA Full-time

[Remote] Senior DevOps Engineer (reputed company reputed company Platform)

Remote, USA Full-time

**reputed company Remote Data Entry Clerk – Accurate and Efficient Data Management for arenaflex**

Remote, USA Full-time

Work From Home for 15 Year Olds: Flexible Teen Job Opportunity

Remote, USA Full-time

reputed company Consultant - Java Developer

Remote, USA Full-time

Business Development Manager (Agency), Northeast

Remote, USA Full-time

Senior Workplace Retirement Plan Consultant

Remote, USA Full-time

[Remote] Project Manager and Planner

Remote, USA Full-time

Head of Policy and Regulatory Affairs – US

Remote, USA Full-time

[Remote] Naval Architecture Arrangements and Weights reputed company

Remote, USA Full-time

reputed company: Remote Nonprofit Administrative Assistant, Data

Remote, USA Full-time

Local Health Support Coordinator, reputed company Health Consultant-E

Remote, USA Full-time

reputed company Consultant - Python

Remote, USA Full-time

Head of Sales

Remote, USA Full-time