Google

Site Reliability Engineer, ML Infrastructure, Large Models SRE

₹ Check with seller / month
📍 London, England, United Kingdom 💼 IT Engineer & Developer ✓ Active
✓ Actively Hiring 📍 London 💼 Full Time
Advertisement
🛡️ Verified Listing
⚡ Direct Apply — No Agent
🔒 Your Data is Safe
⭐ Trusted by 5 Lakh+ Jobseekers

Job at a Glance

Category
IT Engineer & Developer
Location
London, England, United Kingdom
Salary
Check with seller
Job Type
Full Time
Company
Google
Status
Open & Active

Job Description

Minimum qualifications:
Bachelor's degree in Computer Science or a related technical field or equivalent practical experience.
5 years of experience with software development in one or more programming languages.
3 years of experience in designing, analyzing, and troubleshooting distributed systems.
2 years of experience leading projects and providing technical leadership.

Preferred qualifications:
Experience in Large Language Models/Machine Learning tooling and infrastructure.
Experience in automation, monitoring, and incident response.
Experience in C++, Java, Python, or Go.
Understanding of Site Reliability Engineering (SRE) principles and best practices.
Excellent communication, project and stakeholder management skills.
About the job
Site Reliability Engineering (SRE) combines software and systems engineering to build and run large-scale, massively distributed, fault-tolerant systems. SRE ensures that Google's services—both our internally critical and our externally-visible systems—have reliability, uptime appropriate to users' needs and a fast rate of improvement. Additionally SRE’s will keep an ever-watchful eye on our systems capacity and performance.
Ready to take the next step?

Don't wait — new applications are being reviewed daily.

Apply Now →
Job Safety Alert Real jobs on Jobsiya are always free. Never pay for an interview and never share bank or OTP details. Report this job →