Shaerose R.
Principal Site Reliability Engineer (Acting) @Lloyds Banking Group
Signup · Get unlimited contacts
WORK HISTORY
Principal Site Reliability Engineer (Acting) @Lloyds Banking Group
Leeds, GB
Sole Site Reliability Engineer responsible for the availability, reliability, and operational stability of multiple business-critical banking platforms, effectively “keeping the lights on” following a significant reduction in SRE capacity.Primary reliability and platform escalation point for 8 parallel data migration teams delivering on-prem → GCP transitions within a core insurance data platform.Key responsibilities & impact- Sole SRE owner across multiple production platforms after team reduction from 5 SREs to 1, holding end-to-end responsibility for reliability, incident response, and operational resilience- Primary escalation and problem-solving authority for complex production issues, engaged directly by senior engineering and delivery leadership during high-severity incidents- Designed and implemented observability, monitoring, and alerting strategies to improve incident detection, reduce MTTR, and prevent repeat failures- Delivered systemic reliability improvements, prioritising long-term fixes over reactive firefighting under sustained production pressure- Led and successfully delivered a large-scale data migration, managing a cross-functional team of ~10 engineers despite limited prior domain exposure- Embedded SRE principles (resilience, automation, error reduction) across engineering teams operating without dedicated SRE support- Acted as a trusted subject-matter expert for cloud infrastructure, reliability engineering, and operational risk, regularly unblocking delivery where multiple teams were stalled- Provided technical leadership and mentoring to engineers across the organisation, influencing standards and operational practices without formal line managementTechnologies & domains:GCP • Kubernetes • Terraform • CI/CD • Monitoring & Observability • Incident Management • Cloud Infrastructure • Reliability Engineering
EDUCATION
Rotherham College
HND In Computing, IT
ABOUT SHAEROSE R.
I am a Principal Site Reliability Engineer with 12+ years of experience operating and stabilising complex, business-critical systems in large, regulated environments.I specialise in end-to-end reliability ownership, observability, and incident leadership focusing on fixing systemic problems rather than repeatedly firefighting symptoms. I am trusted to operate independently, make sound technical decisions under pressure, and deliver outcomes without close supervision.Currently acting as the sole SRE following significant organisational downsizing, I own reliability, availability, and operational resilience across multiple production platforms. My role involves being the primary escalation point for high-severity incidents, identifying root causes, and implementing long-term fixes that reduce risk and recurrence.I have led cross-functional delivery outside my core domain including successfully running a large-scale data migration with a team of ~10 engineers demonstrating strong technical leadership, pragmatism, and the ability to deliver under uncertainty.My approach to SRE is outcome-driven and pragmatic- Reliability over heroics- Automation over toil- Calm decision-making over noise- Ownership over titles- Core strengths:• Reliability & resilience engineering• Observability & incident management• Cloud & platform engineering (GCP, Kubernetes)• Root cause analysis & systemic problem solving• Technical leadership without bureaucracyI thrive in environments that value trust, autonomy, and engineering maturity, and where reliability is treated as a first-class concern rather than an afterthought.
This profile is compiled from publicly available professional sources. Unifers is not affiliated with or endorsed by LinkedIn. Request removal of this profile.