← All verified jobs

Site Reliability Engineering (SRE) Leader

patsnap Remote, UKRemote

See all open roles at patsnap

Ghost-risk verdict

Likely real

  • 38 open roles at this company in 30 days (mass-hiring blitz)
  • no salary disclosed (correlates with ghost postings)

How we score ghost risk →

See your fit for this role and apply with a truthfully tailored résumé.

See my fit, free

About the role

What You'll be Doing:

Build, lead and develop the UK SRE team, establishing operational standards, best practices, and reliability goals.

Ensure the high availability, stability, security, and performance of business-critical platforms and services.

Define and drive the operational strategy for our global SaaS platform, ensuring exceptional reliability, availability and performance.

Lead major incident management, acting as the senior escalation point during critical production events.

Establish and monitor reliability metrics, including SLIs, SLOs and operational KPIs.

Drive automation across infrastructure, deployments, monitoring and operational workflows to improve efficiency and reduce manual effort.

Champion the adoption of AI-powered operations, leveraging modern AI technologies to enhance engineering productivity and operational excellence.

Partner with Engineering, Product, Security and Infrastructure teams to improve platform architecture, scalability and operational readiness.

Lead disaster recovery planning, operational resilience initiatives and risk management across the platform.

Continuously evaluate emerging cloud, AI and platform technologies to keep PatSnap at the forefront of engineering excellence.

Stay current with emerging cloud, AI, and SRE technologies, driving continuous improvement across the organization.

What We'd Love From You:

Bachelor’s degree in Computer Science or a related field, with at least 8 years of experience in DevOps, SRE, or infrastructure operations.

Proven experience leading technical teams and managing production environments at scale.

Strong expertise in cloud platforms (AWS preferred), Kubernetes, Docker, CI/CD pipelines, Infrastructure as Code, and observability platforms.

Deep understanding of distributed systems, high-availability architectures, and large-scale SaaS environments.

Experience driving automation and operational excellence initiatives.

Hands-on experience using AI tools such as ChatGPT, Claude, GitHub Copilot, Codex, or similar technologies to improve engineering productivity.

Strong problem-solving, leadership, communication, and stakeholder management skills.

Fluent in English; Mandarin is highly desirable to facilitate collaboration with teams across multiple regions.

Stop applying to ghosts.

OyaPilot surfaces only verified, real jobs, scores your fit, and tailors your application truthfully.