Google
Site Reliability Engineer, ML Infrastructure, Large Models SRE
₹ Check with seller / month
✓ Actively Hiring
📍 London
💼 Full Time
🛡️ Verified Listing
⚡ Direct Apply — No Agent
🔒 Your Data is Safe
⭐ Trusted by 5 Lakh+ Jobseekers
Job at a Glance
- Category
- IT Engineer & Developer
- Location
- London, England, United Kingdom
- Salary
- Check with seller
- Job Type
- Full Time
- Company
- Status
- Open & Active
Job Description
Minimum qualifications:
Bachelor's degree in Computer Science or a related technical field or equivalent practical experience.
5 years of experience with software development in one or more programming languages.
3 years of experience in designing, analyzing, and troubleshooting distributed systems.
2 years of experience leading projects and providing technical leadership.
Preferred qualifications:
Experience in Large Language Models/Machine Learning tooling and infrastructure.
Experience in automation, monitoring, and incident response.
Experience in C++, Java, Python, or Go.
Understanding of Site Reliability Engineering (SRE) principles and best practices.
Excellent communication, project and stakeholder management skills.
About the job
Site Reliability Engineering (SRE) combines software and systems engineering to build and run large-scale, massively distributed, fault-tolerant systems. SRE ensures that Google's services—both our internally critical and our externally-visible systems—have reliability, uptime appropriate to users' needs and a fast rate of improvement. Additionally SRE’s will keep an ever-watchful eye on our systems capacity and performance.
Bachelor's degree in Computer Science or a related technical field or equivalent practical experience.
5 years of experience with software development in one or more programming languages.
3 years of experience in designing, analyzing, and troubleshooting distributed systems.
2 years of experience leading projects and providing technical leadership.
Preferred qualifications:
Experience in Large Language Models/Machine Learning tooling and infrastructure.
Experience in automation, monitoring, and incident response.
Experience in C++, Java, Python, or Go.
Understanding of Site Reliability Engineering (SRE) principles and best practices.
Excellent communication, project and stakeholder management skills.
About the job
Site Reliability Engineering (SRE) combines software and systems engineering to build and run large-scale, massively distributed, fault-tolerant systems. SRE ensures that Google's services—both our internally critical and our externally-visible systems—have reliability, uptime appropriate to users' needs and a fast rate of improvement. Additionally SRE’s will keep an ever-watchful eye on our systems capacity and performance.
Job Safety Alert
Real jobs on Jobsiya are always free. Never pay for an interview and never share bank or OTP details.
Report this job →
Similar Jobs:
Site Reliability Jobs in London
—
IT Engineer & Developer Jobs Near You