Developer Publish

  • Jobs
Developer Publish > Senior Software Engineer, Site Reliability Engineering, Cloud IRT

Senior Software Engineer, Site Reliability Engineering, Cloud IRT

June 29, 2026 by admin

  • Full Time
  • London
  • Posted 2 weeks ago
Google

Google

Senior Software Engineer, Site Reliability Engineering, Cloud IRT Overview

Company Name Google
Job Role Senior Software Engineer, Site Reliability Engineering, Cloud IRT
Qualifications Bachelor’s
Category IT Jobs
Job Type Full Time
Location London

As a Senior Software Engineer in Site Reliability Engineering (SRE) for the Cloud Incident Response Team (IRT), you will play a crucial role in ensuring the reliability and performance of Google Cloud’s services. This position involves responding to and coordinating the resolution of major incidents across the Google Cloud Platform, working alongside some of the most experienced engineers in the field.

Your responsibilities will include engaging in the entire lifecycle of services, from inception and design through deployment and refinement. You will support services before they go live by providing system design consulting, developing software platforms, and conducting capacity planning and launch reviews. Once services are live, you will maintain them by measuring and monitoring their availability, latency, and overall health.

In addition, you will build systems and tooling to support the Cloud IRT team, improving visibility into the state of the cloud and enhancing the detection of large-scale issues. Effective communication with customers and stakeholders will be essential, especially during critical incident responses, for which you will participate in an on-call rotation.

Minimum Qualifications

  • Bachelor’s degree in Computer Science, a related field, or equivalent practical experience.
  • 5 years of experience with software development in one or more programming languages.
  • 3 years of experience in designing, analyzing, and troubleshooting large-scale distributed systems.
  • 2 years of experience leading projects and providing technical leadership.
  • Experience troubleshooting production incidents as part of an on-call rotation.

Preferred Qualifications

  • Master’s degree in Computer Science or Engineering.
  • Experience in telemetry systems, incident, and risk management.
  • Ability to work across organizational boundaries.
  • Excellent systematic problem-solving skills and effective communication abilities.

This position is based in London, UK, and is not eligible for visa sponsorship.


Degree Requirement: Bachelor’s

Visa Sponsorship Promising

To apply for this job please visit www.google.com.

Related

Recent Jobs

  • HR Payroll Transformation- Managing Consultant

    • Newcastle
    • Capgemini
    • Full Time
  • Veterinary Surgeon

    • UK
    • Medivet
    • Full Time
  • Senior Multinational Property & Casualty Underwriter

    • London
    • Allianz Insurance
    • Full Time
  • ERP Program Senior Manager

    • Manchester
    • Capgemini
    • Full Time
  • Band 7 Therapy Practice Placement Manager

    • Birmingham
    • NHS
    • Full Time
  • Consultant/Senior Consultant – Data Governance

    • Glasgow
    • Capgemini
    • Full Time
  • Salesforce Architect

    • London
    • Capgemini
    • Full Time
  • Consultant in Palliative Medicine

    • UK
    • NHS
    • Full Time
  • Senior Manager – AI Product Innovation

    • Manchester
    • Capgemini
    • Full Time
  • Lead Data Scientist

    • UK
    • Allianz Insurance
    • Full Time