JobFlexy

Software Development Manager II, Site Reliability – Google – Waterloo, ON

Location: Ontario | Company: Google

Google’s Site Reliability Development team in Waterloo, Ontario is looking for a Software Development Manager II to lead engineers building and running some of the most complex distributed systems in the world. This is a senior leadership role at the intersection of software engineering and systems reliability — where your decisions directly affect the uptime and performance of services used by billions of people.

Sponsored Links

This isn’t a hands-off management role. You’ll be expected to lead by example, write and deliver software, mentor your team, and take direct ownership of service availability. The work spans automation, infrastructure optimization, and large-scale distributed system design — all within Google’s unique engineering culture that values intellectual curiosity, openness, and bold thinking.

About the Role: Software Development Manager II, Site Reliability

As a Software Development Manager II on the Site Reliability Development team, you’ll lead a group of software and systems developers working on projects that keep Google’s internal and external services running reliably at scale. You’ll own end-to-end availability and performance of key services, driving automation to prevent recurring issues and designing systems that are faster, more scalable, and more efficient.

The team operates with a follow-the-sun on-call model, meaning you’ll manage rotations across continents and help build a culture of shared responsibility. Google’s SRD teams are known for their collaborative, blame-free environments where people are encouraged to take risks, think big, and learn from outcomes. You’ll have both the autonomy to drive meaningful technical work and the support structure to grow as a leader.

Sponsored Links

Benefits and Salary

Google offers a competitive compensation package for this role in Canada. The salary range is $216,000 – $221,000 CAD, plus a 20% bonus target, equity (GSU grants through Alphabet Inc.), and a comprehensive benefits package. Google is known for offering strong employee benefits — details are available on Google’s official careers site.

Job Details

🏢 Company: Google

📍 Location: Waterloo, ON, Canada

📌 Job Type: Advanced / Senior Leadership

💰 Pay: $216,000 – $221,000 CAD + 20% bonus target + equity + benefits

Responsibilities

In this role, you’ll be accountable for both the people and the systems. Day-to-day work includes technical leadership, hands-on software development, and cross-continental coordination — all aimed at keeping Google’s services performing at the highest standards. Here’s what you can expect to be doing:

  • Lead a team of software and systems developers on reliability-focused projects, with direct responsibility for service uptime
  • Own availability and performance end-to-end for key services, building automation to prevent problem recurrence and respond to non-exceptional service conditions
  • Design, write, and deliver software to improve the availability, scalability, latency, and efficiency of Google’s services
  • Mentor and develop team members, establishing credibility through quality technical execution and leading by example
  • Manage on-call rotations across continents using a follow-the-sun model to ensure continuous coverage and response

Requirements / Skills

Google is looking for someone who brings both deep technical expertise and proven leadership experience. The ideal candidate is comfortable operating at the intersection of software development and systems reliability, and can inspire a team while still rolling up their sleeves when it counts.

  • Bachelor’s degree in Computer Science, a related field, or equivalent practical experience
  • 8+ years of software development experience in one or more programming languages
  • 3+ years managing people or teams in a technical engineering environment
  • 3+ years leading projects, with demonstrated ability to drive outcomes at scale
  • 3+ years of experience designing, analyzing, and troubleshooting distributed systems
  • Master’s degree in Computer Science or Engineering is preferred but not required

How to Apply

To apply, use the link below to visit the official Google Careers posting. Make sure your resume is current and reflects your experience with distributed systems, team leadership, and software development before submitting.

Share This Opportunity

Know someone who might be interested? Share this job posting and help them join Google in Waterloo.

Job Summary & Tips for Applying

AI-generated summary and tips to help you highlight your strengths effectively.

Quick Summary & What to Highlight: This Software Development Manager II, Site Reliability role at Google in Waterloo is well-suited for candidates who excel in distributed systems design, technical team leadership, and large-scale software development. On your resume, emphasize your experience managing engineering teams, driving reliability initiatives, and writing production-quality software. If you’ve previously worked in site reliability engineering, platform infrastructure, or systems software, make sure to highlight specific outcomes — uptime metrics, automation projects, or system improvements you led.

Resume & Application Tips: Before applying, tailor your resume to match the job description. Include keywords like site reliability, distributed systems, and on-call management that appear throughout the posting. Quantify your achievements where possible (e.g., “reduced incident response time by 40%” or “led a team of 8 engineers across 3 time zones”). Write a brief cover letter expressing your interest in Google’s SRD culture and why the challenges of reliability at Google’s scale appeal to you. Double-check your application for accuracy and ensure your contact information is current.

Interview Preparation: If selected for an interview, research Google‘s Site Reliability Development philosophy, including their published books and engineering blog posts on the subject. Prepare specific examples using the STAR method (Situation, Task, Action, Result) to demonstrate your experience with system reliability, automation, and people management. Common questions may include scenarios about handling major incidents, building team culture, and making technical trade-offs under pressure. Dress professionally, arrive (or log in) early, and bring copies of your resume if interviewing in person. Prepare thoughtful questions about the team’s current reliability challenges and growth opportunities within Google’s SRD organization. Follow up with a thank-you email within 24 hours of your interview.