Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Software Engineering Manager II, Site Reliability Engineering, Data Intelligence

$207k - $300k

Google

Hire, develop, and mentor a high-performing, SRE team to support career growth and team health.Set team goals, prioritize resources, and define technical roadmaps aligned with partner teams and key stakeholders.Guide the architecture and review of resilient, high-performance systems powering core AI infrastructure.Drive incident response, maintain high reliability standards, and actively automate operational toil.Partner across development teams to align technical direction while leveraging AI to accelerate team productivity.Minimum qualifications:Bachelor’s degree in Computer Science, a related field, or equivalent practical experience.8 years of experience with software development in one or more programming languages (e.g., C++, Java, Python), or with data structures/algorithms.3 years of experience in designing, analyzing, and troubleshooting large-scale distributed systems.3 years of experience managing and growing engineering teams, including performance management and career development.Preferred qualifications:Experience managing distributed teams across multiple sites or timezones.Experience developing long-term technical roadmaps, driving organizational change, and influencing cross-functional stakeholders (Dev, PM, Leadership).Proven track record of hiring, mentoring, and leading high-performing Site Reliability Engineering or Software Engineering teams.Systematic problem-solving and troubleshooting skills in complex, ambiguous software systems.Passion for AI infrastructure and driving AI transformation within engineering workflows.Site Reliability Engineering (SRE) combines software and systems engineering to build and run large-scale, massively distributed, fault-tolerant systems. SRE ensures that Google's services—both our internally critical and our externally-visible systems—have reliability, uptime appropriate to users' needs and a fast rate of improvement. Additionally SRE’s will keep an ever-watchful eye on our systems capacity and performance.Much of our software development focuses on optimizing existing systems, building infrastructure and eliminating work through automation. On the SRE team, you’ll have the opportunity to manage the complex challenges of scale which are unique to Google, while using your expertise in coding, algorithms, complexity analysis and large-scale system design. SRE's culture of intellectual curiosity, problem solving and openness is key to its success. Our organization brings together people with a wide variety of backgrounds, experiences and perspectives. We encourage them to collaborate, think big and take risks in a blame-free environment. We promote self-direction to work on meaningful projects, while we also strive to create an environment that provides the support and mentorship needed to learn and grow.To learn more: check out our books on Site Reliability Engineering or read a career profile about why a Software Engineer chose to join SRE.Join Google’s Core AI Foundations SRE team! We build and scale the critical infrastructure—authorization, ML training storage, and RPC scheduling—powering products like Gemini, Workspace, NotebookLM, and Cloud.You will manage and grow a North American SRE team. You will partner closely with development teams and our Sydney counterpart to run a follow-the-sun rotation and ensure exceptional reliability for Google’s frontier AI capabilities.The Core team builds the technical foundation behind Google’s flagship products. We are owners and advocates for the underlying design elements, developer platforms, product components, and infrastructure at Google. These are the essential building blocks for excellent, safe, and coherent experiences for our users and drive the pace of innovation for every developer. We look across Google’s products to build central solutions, break down technical barriers and strengthen existing systems. As the Core team, we have a mandate and a unique opportunity to impact important technical decisions across the company.Individual pay is determined by factors including job-related skills, experience, and relevant education or training. US: $207000 - $300000 (USD) + 20% bonus target + equity + benefitsLearn more about benefits at Google.Bachelor’s degree in Computer Science, a related field, or equivalent practical experience.8 years of experience with software development in one or more programming languages (e.g., C++, Java, Python), or with data structures/algorithms.3 years of experience in designing, analyzing, and troubleshooting large-scale distributed systems.3 years of experience managing and growing engineering teams, including performance management and career development.

Vacancy posted 7 days ago
Similar jobs that could be interesting for youBased on the Software Engineering Manager II, Site Reliability Engineering, Data Intelligence in San Jose, CA vacancy
  • $207k - $300k

    Lead a team of Software/Systems Engineers on projects for users and...  ...technical execution.Manage on-call rotations across...  ...in Artificial Intelligence or Machine Learning....  ...Large Language Model.Site Reliability Engineering (SRE) combines...  ...and maintaining our data centers to building... 
    Intelligence

    Google

    Sunnyvale, CA
    4 days ago
  • $207k - $300k

    Manage a team of Software/Systems Engineers on projects for users and remain directly responsible...  ...sustainable multi-site on-call rotations across...  ...practical expertise in Site Reliability Engineering practices, including...  ...of Google's foundational data pipeline and will manage... 
    Suggested

    Google

    San Jose, CA
    7 days ago
  • $207k - $300k

    Lead a team of Software/Systems Engineers on projects for users and be directly...  ...quality technical execution.Manage on-call rotations across...  ...years of experience with data structures or algorithms.5...  ...people management experience. Site Reliability Engineering (SRE) combines... 
    Suggested

    Google

    Mountain View, CA
    14 days ago
  •  ...Software Engineer II TENEX is an AI-native, automation-first, built-for-scale Managed Detection and Response (MDR) provider...  ...deploy scalable and reliable backend services and...  ...or Professional Data Engineer are a plus...  ...expertise with artificial intelligence and have experience... 
    Intelligence
    Remote work
    Monday to Friday

    TenEx

    San Jose, CA
    2 days ago
  • $207k - $300k

     ...code developed by engineers to ensure...  ...excellence and reliability, ensuring strict financial data integrity, high...  ...orchestration engines that manage YouTube's end-...  ...experience in software development.3...  ..., artificial intelligence, natural...  ...across multiple sites internationally... 
    Intelligence

    Google

    Mountain View, CA
    4 days ago
  • $207k - $300k

     ...of experience in software development. ~3...  ...experience in a people management or team leadership...  ...of a Software Engineer goes beyond just Search...  ..., artificial intelligence, natural language...  ...networking, security, data compression, user...  ...across multiple sites internationally.... 
    Intelligence

    Jobleads-US

    Sunnyvale, CA
    1 day ago
  •  ...procurement, transport logistics, data center design and construction, equipment management, and daily operations. Bitdeer...  ...with high demand for artificial intelligence. Headquartered in Singapore,...  ...remediation-actuator and workflow engine land here — you make the control... 
    Intelligence
    Local area

    Bitdeer Technologies Group

    San Jose, CA
    2 days ago
  •  ...Job Description Job Description Site Reliability Engineer II Bay Area, offices in San Jose · Hybrid · 24/7 FedRAMP Operations · Rotational Shift · Initial Contract till March 27. KEY REQUIREMENT This role requires US citizenship and residence on US soil... 
    Hourly pay
    Contract work
    For contractors
    Shift work
    Night shift
    Weekend work

    C-Serv

    San Jose, CA
    19 days ago
  • $248k - $396.75k

    Site Reliability Engineering (SRE) at NVIDIA is an engineering discipline focused...  ...availability. It combines software and systems engineering practices...  ..., databases, capacity management, continuous delivery, and...  ...AI agents, AI skills, and intelligent automation that accelerate... 
    Intelligence
    Full time

    NVIDIA

    Santa Clara, CA
    1 day ago
  • $122.5k - $175k

     ...from cyberattacks and data loss by securely...  ...amplified by machine intelligence to solve the world’s...  ...looking for a Staff Site Reliability Engineer to join our team. This...  ...infrastructure and managing platforms like Kubernetes...  ...deploy systems and software in diverse... 
    Intelligence
    Full time
    Work at office
    Local area
    3 days per week

    Zscaler

    San Jose, CA
    1 day ago
  • $137.5k - $200.5k

     ...customers by building reliable dashboards and providing...  ...Cloud Operations, and Engineering to guide the product roadmap...  ...to help us build our data strategy?Your impactAs...  ...Data Science Engineer II, you will lead the development...  ...see the Cisco careers site to discover more... 
    Full time
    Temporary work
    Work experience placement
    Local area
    Flexible hours

    CISCO Systems

    San Jose, CA
    7 days ago
  • $168k - $270.25k

     ...developments in Artificial Intelligence, High-Performance...  ...team at NVIDIA as a Senior Site reliability engineer focused on HPC storage and...  ...Storage solutions tailored for data-intensive applications, optimizing...  ...automate deployment and management of large-scale... 
    Intelligence
    Full time

    Nvidia

    Santa Clara, CA
    25 days ago
  •  ...Enterprise Technologies Inc. is a recognized provider of professional IT Consulting services in the US. We are actively seeking Data Engineer II for one of our client. Role: Data Engineer II Location: Santa Clara,CA Duration: Long Term The project has two modules: Data... 

    Rootshell Enterprise Technologies

    Santa Clara, CA
    6 days ago
  • $207k - $300k

     ...optimization and data processing strategies...  ...experience with software development in...  ...experience in a people management or team...  ...degree or PhD in Engineering, Computer Science...  ...retrieval, artificial intelligence, natural language...  ...across multiple sites internationally.The... 
    Intelligence

    Google

    Sunnyvale, CA
    1 day ago
  • $207.4k - $259.2k

     ...systems that combine software intelligence with physical...  ...seeking exceptional engineers, operators and builders...  ...passionate Sr. Staff Site Reliability Engineer (SRE) to join...  ...Architect and optimize data pipelines to ensure...  ..., and release management.Champion cloud-first... 
    Intelligence
    Permanent employment
    Local area
    Visa sponsorship
    Night shift

    Archer Aviation

    San Jose, CA
    7 days ago
  •  ...Introduction At IBM Software, we transform client...  ...degree. As a Site Reliability Engineer, you will work in an...  ...day operations, alert management, incident support, migration...  ...SQL, NoSQL, and data streaming...  ...business operations with intelligence-from machine learning... 
    Intelligence
    Full time
    Contract work
    Part time
    Fixed term contract
    Internship
    Worldwide
    Flexible hours
    Shift work

    IBM

    San Jose, CA
    4 days ago
  • ScaleFlux is seeking a Senior Product Manager to drive strategy, roadmap, and execution...  ...NVMe SSD products aimed at AI, cloud, and data center markets. You will shape storage...  ...SSD controller architectures, firmware intelligence, and computational storage to meet AI training... 
    Intelligence

    ScaleFlux

    Milpitas, CA
    3 days ago
  • $207k - $300k

     ...code developed by other engineers and provide feedback to...  ...years of experience with software development in one or more...  ...experience in a people management or team leadership role....  ...design, networking and data storage, security, artificial intelligence, natural language processing... 
    Intelligence

    Google

    Sunnyvale, CA
    2 days ago
  • $168k - $270.25k

    Site Reliability Engineering (SRE) at NVIDIA is an engineering discipline to design...  ...using the combination of software and systems engineering...  ...coding, database, capacity management, continuous delivery and deployment...  ...in Artificial Intelligence, High-Performance Computing... 
    Intelligence
    Full time

    Nvidia

    Santa Clara, CA
    11 days ago
  • $224k - $356.5k

     ...midst of a revolution where artificial intelligence is helping us accelerate all aspects...  ...intelligence to robots.NVIDIA is seeking an Engineering Manager to lead our Robotics Neural...  ...lead a group of world-class robotics software and applied research engineers focused... 
    Intelligence
    Full time

    Nvidia

    Santa Clara, CA
    a month ago
  • $207k - $300k

     ...Software Engineer Manager II, Spanner Enterprise Google Sunnyvale...  ...retrieval, artificial intelligence, natural language...  ...networking, security, data compression, user interface...  ...across multiple sites internationally....  ...their scalability, reliability, and availability. You... 
    Intelligence

    Jobleads-US

    Sunnyvale, CA
    5 days ago
  • $272k - $431.25k

    NVIDIA is looking for a Cloud Site Reliability Engineering Architect to work in IPP's...  ...Deep Learning, Artificial Intelligence and Autonomous Vehicles to...  ...of thousands of NVIDIA's software engineers worldwide. The...  ...solutions, mine through data to uncover real problems and... 
    Intelligence
    Full time
    Work experience placement
    Worldwide

    Nvidia

    Santa Clara, CA
    a month ago
  • $105.5k - $213.5k

     ...Software Engineer II This role has been designed as 'Onsite' with an expectation...  ..., analyze, and act on their data and applications wherever...  ...We have the flexibility to manage our work and personal needs....  ...to deliver high-quality, reliable, and cost-effective software... 
    Work experience placement
    Work at office
    Local area
    Immediate start

    Hewlett Packard Enterprise

    Cupertino, CA
    3 days ago
  • $226.14k

     ...time to building systems that operate reliably on a global scale. When you work here,...  ...we'd love to meet you. Job Title: Software Engineer II Location: 50 West San Fernando Street...  ...integrations following secure coding and data governance practices. Build and operate... 
    Full time
    Temporary work
    Work at office
    Remote work
    Worldwide

    The Trade Desk

    San Jose, CA
    2 days ago
  • $115.2k - $172.8k

     ...Software Development Engineer II At F5, we strive to bring a better digital world to life. Our teams empower organizations across the globe to create, secure, and run applications that enhance how we experience our evolving digital world. We are passionate about cybersecurity... 
    Work at office
    Local area
    Remote work
    Home office

    F5

    San Jose, CA
    7 hours ago
  • $180k - $220k

     ...is building an operational intelligence platform for digital infrastructure...  ...experience, and automated data engineering pipelines. Our...  ...Job Overview  Product Manager is responsible for working...  ...observability, APM, or IT operations software.  ~ Strong understanding... 
    Intelligence
    Night shift

    Selector Software

    Santa Clara, CA
    26 days ago
  •  ...procurement, transport logistics, data center design and construction, equipment management, and daily operations. Bitdeer...  ...high demand for artificial intelligence. Headquartered in Singapore,...  ...capabilities. Partner closely with engineering, infrastructure, operations,... 
    Intelligence
    Full time
    Local area

    Bitdeer Technologies Group

    San Jose, CA
    5 days ago
  • $100k

     ...evolve to unify innovations in software models, compilers, platforms...  ...connection to Product, Engineering, Field Applications, and Marketing...  ...training, inference, data and analytics, HPC, databases...  ...opportunity data, using market intelligence and AI-enabled tools to... 
    Intelligence
    Permanent employment

    Tenstorrent

    Santa Clara, CA
    7 days ago
  • $250k - $300k

     ...Position reports to:  VP, Engineering Application...  ...and operating highly reliable, scalable, and secure...  ...connectivity, traffic management, load balancing, and network...  ...critical middleware and data services, including...  ...We may use artificial intelligence (AI) tools to support... 
    Intelligence
    Full time
    H1b
    Work at office
    Local area
    Work from home
    Work visa
    Flexible hours
    3 days per week

    Extreme Networks

    San Jose, CA
    24 days ago
  • $171.5k - $245k

     ...customers from cyberattacks and data loss by securely connecting...  ...is amplified by machine intelligence to solve the world’s hardest...  ...looking for a Principal Product Manager - Sovereign Cloud to join our...  ...functional collaborations across engineering, sales, and marketing to... 
    Intelligence
    Full time
    Work at office
    Local area

    Zscaler

    San Jose, CA
    14 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Software Engineering Manager II, Site Reliability Engineering, Data Intelligence. Be the first to apply!