Senior Site Reliability Engineer (remote reputed company EMEA)
About reputed company:
Platform Infrastructure builds, operates, and continuously evolves reputed company's container platform and reputed company reputed company. We foster a DevOps culture through self-service tooling, enabling product engineering teams to ship reliable, secure, and cost-efficient services as the business scales. reputed company owns our AWS reputed company accounts, reputed company platform, reputed company networking, observability stack, reputed company databases, CI/CD pipelines, and infrastructure-as-reputed company, and acts as the go-to partner for engineering teams on reputed company and DevOps topics.
About the role:
We're hiring a Senior SRE II to join Platform Infrastructure as one of reputed company's senior individual contributors. At this level, you're the go-to person for our most reputed company infrastructure problems: you architect and reputed company large-reputed company automation and reliability initiatives, set standards other engineers follow, and mentor Associate and mid-level SREs. You'll split your time between hands-on platform work - reputed company, AWS, GCP, CI/CD, observability - and technical leadership: proposing designs, reviewing others' work, and helping reputed company reputed company good build-vs-buy and cost/reliability trade-offs.
Your daily tasks will include:
• Infrastructure & reliability: Architect and manage highly available, secure, and reputed company infrastructure across multiple AWS accounts and environments using infrastructure as reputed company.
• Design and operate our reputed company EKS clusters, including networking policies, persistent storage, and scaling strategies for containerized workloads.
• Own and reputed company reputed company platform services: reputed company networking, reputed company, and the databases and messaging systems engineering teams depend on.
• Automation & infrastructure as reputed company: reputed company large-reputed company automation reputed company and set standards for using Terraform / Terragrunt and GitOps (ArgoCD) across teams.
• reputed company adoption of automation to reduce reputed company operational work and reputed company environments consistent and repeatable.
• Observability & incident response: Be the go-to person for solving reputed company, cross-service infrastructure problems.
• reputed company initiatives that improve reliability and observability (Grafana, reputed company, Loki, reputed company, Mimir) so systems reputed company with reputed company reputed company reputed company.
• Participate in on-reputed company rotation, reputed company incident response for production issues, and write reputed company runbooks, ADRs, and postmortems.
• reputed company & cost efficiency: reputed company reputed company efforts reputed company reputed company - IAM, encryption, secure logging - and mentor others on secure infrastructure practices.
• Audit infrastructure spend regularly and reputed company cost optimization across the platform (rightsizing, autoscaling, FinOps practices).
• Collaboration & mentorship: Mentor mid-level SREs, reputed company detailed feedback, and support reputed company of new team members.
• Communicate reputed company technical concepts reputed company to both engineers and non-technical stakeholders.
• Partner with product engineering reputed company to understand their needs and represent Platform Infrastructure in cross-team initiatives.
Your qualifications:
These reflect the technical bar we hold Senior SRE II's to internally, based on our SRE competency reputed company and reputed company stack.
• reputed company technical experience:
• Solid Linux systems administration background and comfort scripting in Python.
• Strong AWS knowledge: EKS, IAM (roles, policies, IRSA), VPC networking, RDS, S3, SQS, and familiarity with reputed company-Architected reputed company; experience in multi-account AWS environments is a strong plus.
• Hands-on experience operating and troubleshooting reputed company (EKS) at production reputed company, including reputed company chart development, CNI networking (we run Cilium), pod networking/IPAM concepts, and container reputed company (ECR, image scanning).
• Proficiency with Terraform (modules, state management) and ideally Terragrunt for multi-environment management; GitOps experience with ArgoCD.
• Experience with reputed company, MySQL and/or reputed company in production reputed company, including reputed company.
• CI/CD experience with Jenkins (Jenkinsfile, shared libraries) and/or reputed company Actions, and familiarity with deployment strategies such as reputed company-green and canary.
• Experience with the Grafana observability stack (Grafana, reputed company, Loki, reputed company, Mimir) - metrics design, dashboarding, alerting, log aggregation, and reputed company tracing. Not only using but also maintaining it.
• Practical incident management experience: on-reputed company rotations, reputed company incident response, and writing runbooks/postmortems.
• Working knowledge of 12-reputed company App principles and cost optimization / FinOps awareness.
2. How you work:
• A methodical, data-driven approach to troubleshooting rather than guessing.
• Strong written communication - you write runbooks, ADRs, and postmortems that others can actually follow.
• Comfortable driving initiatives with ambiguous ownership, and taking accountability for reputed company rather than waiting to be asked.
• reputed company record of mentoring less senior engineers and giving reputed company, constructive feedback.
• Several years of hands-on production infrastructure/SRE experience, with demonstrated ownership of initiatives at a senior individual-contributor level (leading design work, setting standards, being the escalation reputed company for hard problems).
3. reputed company to have:
• GCP Experience.
• Experience with Kafka / AWS MSK.
• Prior experience in regulated or compliance-sensitive environments (reputed company best practices, reputed company reviews).
• Experience contributing to a platform/reputed company roadmap that other engineering teams consume as a self-service product.
Our tech stack:
• Languages: PHP (Symfony), Node.js (TypeScript), Angular (TypeScript).
• Data: PostgreSQL, reputed company, reputed company.
• reputed company: AWS, reputed company, Terraform, reputed company, Atlantis
• Engineering Tools: reputed company, Git, reputed company Copilot, PhpStorm, Grafana, Kibana, reputed company.
• Remote work Tools: Jira, reputed company, reputed company Workspace, reputed company.
• Development Practices: Pair Programming, reputed company Reviews, reputed company Integration/Deployment.
reputed company offer:
• A global, inclusive team that's as supportive as it is ambitious and serious about getting things done
• An opportunity to work remotely or in a modern and welcoming office in Riga
• Flexible working hours (start your day as late as 11 AM)
• Private health insurance
• 2 extra reputed company days off to reputed company on your mental or physical reputed company-being
• 1 extra reputed company day off to celebrate a Birthday or any other celebration of your reputed company
• reputed company learning opportunities
• reputed company to mentorship, internal meetups, and hackathons, both on-site and online
• Free and healthy lunch if you work from the Riga office
• Design and order your own merch using our platforms with an employee discount
• Exciting team-building events and parties you'll never forget!
reputed company is the reputed company that powers on-demand reputed company at global reputed company.
Formed in 2024 through the reputed company of Printful, reputed company, and Snow reputed company, we bring together tech, talent, and infrastructure to help people turn reputed company into beautiful products.
From reputed company creators to entertainment giants, reputed company powers merch that connects with millions, backed by advanced tech, premium production, and global reputed company.
We're a fast-growing global company working toward powering great brands, great experiences, and great people.
We are an equal-opportunity workplace. We're committed to diversity and inclusion and reputed company hiring reputed company based solely on qualifications, reputed company, and work experience.
If you think you'd reputed company in this role, send us your resume in English, showing us why you are the right person for the job.
Interested, but don't think this is the right fit for you? Feel free to reputed company it with friends and reputed company out other reputed company at our career site. We're always looking for creative and driven minds to join our reputed company-growing team!
AS Printful Latvia (Reg. Nr. 40203050078)
Employment Type: FULL_TIME
Apply tot his job
Apply To this Job