Partner companyPartner company
Partner companyPartner company
Staff Site Reliability Engineer is sought by a partner company to own end-to-end infrastructure for large-scale, customer-facing technology products in a fully remote role based in the United Kingdom. The role combines cloud architecture, distributed systems, reliability engineering, observability, performance optimization, and developer tooling. Responsibilities include end-to-end ownership of core infrastructure products or subsystems from design through deployment and production operations; defining goals and success metrics; translating product requirements into robust designs; building secure, reliable, cost-efficient infrastructure; developing production software and developer-facing tooling; managing infrastructure as code (Terraform); collaborating with product engineering, security, and DevOps; participating in incident response; improving monitoring and observability; applying a security-minded approach; mentoring engineers; and guiding infrastructure strategy. Requirements include 6–10 years of experience in infrastructure, platform, or backend engineering in cloud environments (AWS preferred); proven ownership of large infrastructure products; Terraform proficiency; deep cloud infrastructure knowledge (networking, load balancing, containers, Kubernetes/EKS, distributed systems); programming skills in Go or Python; Redis/ElastiCache experience; observability tools (Prometheus, Grafana, OpenTelemetry); strong performance tuning and incident management; excellent English communication; ability to collaborate in a distributed environment. Benefits include fully remo
Track similar jobs
Get email alerts when new roles like this are posted.
Based on: Staff Site Reliability Engineer
See how this role matches your resume
Sign in to get an AI match score and personalized feed.
Staff Site Reliability Engineer is sought by a partner company to own end-to-end infrastructure for large-scale, customer-facing technology products in a fully remote role based in the United Kingdom. The role combines cloud architecture, distributed systems, reliability engineering, observability, performance optimization, and developer tooling. Responsibilities include end-to-end ownership of core infrastructure products or subsystems from design through deployment and production operations; defining goals and success metrics; translating product requirements into robust designs; building secure, reliable, cost-efficient infrastructure; developing production software and developer-facing tooling; managing infrastructure as code (Terraform); collaborating with product engineering, security, and DevOps; participating in incident response; improving monitoring and observability; applying a security-minded approach; mentoring engineers; and guiding infrastructure strategy. Requirements include 6–10 years of experience in infrastructure, platform, or backend engineering in cloud environments (AWS preferred); proven ownership of large infrastructure products; Terraform proficiency; deep cloud infrastructure knowledge (networking, load balancing, containers, Kubernetes/EKS, distributed systems); programming skills in Go or Python; Redis/ElastiCache experience; observability tools (Prometheus, Grafana, OpenTelemetry); strong performance tuning and incident management; excellent English communication; ability to collaborate in a distributed environment. Benefits include fully remo
Track similar jobs
Get email alerts when new roles like this are posted.
Based on: Staff Site Reliability Engineer
See how this role matches your resume
Sign in to get an AI match score and personalized feed.
42 active roles from this employer in the JobMatcher catalog.
Software Engineer
Reltio
Senior Software Developer - Platform & Infrastructure Engine
Contabo
Senior Platform Engineering Manager
Prolific
Senior JavaScript Engineer
Ciklum
Digital PR Manager
Single Grain
Technical Delivery Manager
Stevens
B2B Account Executive
Fastest Labs of Boise & Meridian
Staff Software Engineer
Valon
Project Manager
The Brandon Green Management Group (BGMG)
Contract Technical Recruiter
Blue Acorn iCi
42 active roles from this employer in the JobMatcher catalog.
Software Engineer
Reltio
Senior Software Developer - Platform & Infrastructure Engine
Contabo
Senior Platform Engineering Manager
Prolific
Senior JavaScript Engineer
Ciklum
Digital PR Manager
Single Grain
Technical Delivery Manager
Stevens
B2B Account Executive
Fastest Labs of Boise & Meridian
Staff Software Engineer
Valon
Project Manager
The Brandon Green Management Group (BGMG)
Contract Technical Recruiter
Blue Acorn iCi