Site Reliability Engineer
Skills
- Software Engineering
- Systems Engineering
- Distributed Systems
- Unix
- Linux
- IP Networking
- Performance Tuning
- Reliability Engineering
About the role
About the job:
Site Reliability Engineering (SRE) combines software and systems engineering to build and run large-scale, massively distributed, fault-tolerant systems. SRE ensures that Google Cloud's services—both our internally critical and our externally-visible systems—have reliability, uptime appropriate to customer's needs and a fast rate of improvement. Additionally SRE’s will keep an ever-watchful eye on our systems capacity and performance.
Responsibilities:
Design, code and execute on projects to improve the reliability posture of critical enterprise applications.
Minimum qualifications:
Bachelor’s degree in Computer Science, a related field, or equivalent practical experience.
1 year of experience with software development in one or more programming languages.
1 year of experience with data structures and algorithms.
Preferred qualifications:
Experience in an engineering or operations role in large-scale enterprise space.
Expertise in Unix/Linux systems, IP networking, performance and application issues.
Expertise in problem solving and analyzing complex enterprise systems.
Proficiency in navigating enterprise software, deployment and management of workloads.
Ability to work with multiple global stakeholders.
Ready to apply?
Sign up first — takes a minute — and you get Hiro’s take on this role, a resume tailored to it, and (if available) a referral from a real employee at Google. All free with your Pro gift.
Sign up to apply