Site Reliability Engineer-GCP
Job Summary
Were looking for a Google Product Site Reliability Engineer to join our Public Cloud Platform. Youll have a unique opportunity to be part of an ambitious team to strengthen observability reliability and operation excellence across our GCP platform with the purpose of driving our tech modernisation agenda and enable us to become the biggest Fintech in the UK.
The ideal candidate will have demonstrable experience in Cloud engineering Observability platforms and a passion for technology. Commitment to delivering high-quality scalable solutions is a must.
Responsibilities:
Define and evolve observability standards across metrics logs traces and events
Partner with teams to ensure services are observable by design
Use Dynatrace as the primary observability tool to ensure effective instrumentation and coverage meaningful dashboards and SLO based alerting aligned to user impact
Be hands-on engineering maintaining our Infrastructure as Code and CI/CD pipeline-based product and services by responding to change implementing enhancements & improving reliability and customer experience
Observing investigating & fixing service issues with an engineering attitude resolving via code changes and implementing improvements to prevent repeat issues
Implementing further automation and reducing toil by utilising existing Cloud tooling or implementing new technologies
Primary Skills (Must-Have):
- GCP (Hands-on)
- Kubernetes (Production experience)
- CI/CD Terraform (IaC)
- Observability (Dynatrace preferred / equivalent acceptable with alignment)
Good-to-Have Skills:
- SRE practices (SLO/SLI incident management RCA)
- Python / Bash / scripting
- Prometheus Grafana ELK Splunk
- Cloud networking & security
- Multi-cloud exposure (AWS / Azure)
- Banking / financial services domain
Required Experience:
IC
About Company
As a global leader, Wipro blends consulting and AI expertise across design, engineering and operations to accelerate business transformation and deliver future-ready technology.