Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Software Engineering Manager II, Site Reliability Engineering, Data Intelligence

$207k - $300k

Google

Hire, develop, and mentor a high-performing, SRE team to support career growth and team health.Set team goals, prioritize resources, and define technical roadmaps aligned with partner teams and key stakeholders.Guide the architecture and review of resilient, high-performance systems powering core AI infrastructure.Drive incident response, maintain high reliability standards, and actively automate operational toil.Partner across development teams to align technical direction while leveraging AI to accelerate team productivity.Minimum qualifications:Bachelor’s degree in Computer Science, a related field, or equivalent practical experience.8 years of experience with software development in one or more programming languages (e.g., C++, Java, Python), or with data structures/algorithms.3 years of experience in designing, analyzing, and troubleshooting large-scale distributed systems.3 years of experience managing and growing engineering teams, including performance management and career development.Preferred qualifications:Experience managing distributed teams across multiple sites or timezones.Experience developing long-term technical roadmaps, driving organizational change, and influencing cross-functional stakeholders (Dev, PM, Leadership).Proven track record of hiring, mentoring, and leading high-performing Site Reliability Engineering or Software Engineering teams.Systematic problem-solving and troubleshooting skills in complex, ambiguous software systems.Passion for AI infrastructure and driving AI transformation within engineering workflows.Site Reliability Engineering (SRE) combines software and systems engineering to build and run large-scale, massively distributed, fault-tolerant systems. SRE ensures that Google's services—both our internally critical and our externally-visible systems—have reliability, uptime appropriate to users' needs and a fast rate of improvement. Additionally SRE’s will keep an ever-watchful eye on our systems capacity and performance.Much of our software development focuses on optimizing existing systems, building infrastructure and eliminating work through automation. On the SRE team, you’ll have the opportunity to manage the complex challenges of scale which are unique to Google, while using your expertise in coding, algorithms, complexity analysis and large-scale system design. SRE's culture of intellectual curiosity, problem solving and openness is key to its success. Our organization brings together people with a wide variety of backgrounds, experiences and perspectives. We encourage them to collaborate, think big and take risks in a blame-free environment. We promote self-direction to work on meaningful projects, while we also strive to create an environment that provides the support and mentorship needed to learn and grow.To learn more: check out our books on Site Reliability Engineering or read a career profile about why a Software Engineer chose to join SRE.Join Google’s Core AI Foundations SRE team! We build and scale the critical infrastructure—authorization, ML training storage, and RPC scheduling—powering products like Gemini, Workspace, NotebookLM, and Cloud.You will manage and grow a North American SRE team. You will partner closely with development teams and our Sydney counterpart to run a follow-the-sun rotation and ensure exceptional reliability for Google’s frontier AI capabilities.The Core team builds the technical foundation behind Google’s flagship products. We are owners and advocates for the underlying design elements, developer platforms, product components, and infrastructure at Google. These are the essential building blocks for excellent, safe, and coherent experiences for our users and drive the pace of innovation for every developer. We look across Google’s products to build central solutions, break down technical barriers and strengthen existing systems. As the Core team, we have a mandate and a unique opportunity to impact important technical decisions across the company.Individual pay is determined by factors including job-related skills, experience, and relevant education or training. US: $207000 - $300000 (USD) + 20% bonus target + equity + benefitsLearn more about benefits at Google.Bachelor’s degree in Computer Science, a related field, or equivalent practical experience.8 years of experience with software development in one or more programming languages (e.g., C++, Java, Python), or with data structures/algorithms.3 years of experience in designing, analyzing, and troubleshooting large-scale distributed systems.3 years of experience managing and growing engineering teams, including performance management and career development.

Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Software Engineering Manager II, Site Reliability Engineering, Data Intelligence in San Jose, CA vacancy
  • $207k - $300k

    Manage a team of Software/Systems Engineers on projects for users and remain directly responsible...  ...sustainable multi-site on-call rotations across...  ...practical expertise in Site Reliability Engineering practices, including...  ...of Google's foundational data pipeline and will manage... 
    Suggested

    Google

    San Jose, CA
    3 days ago
  •  ...Job Description Job Description Site Reliability Engineer II Bay Area, offices in San Jose · Hybrid · 24/7 FedRAMP Operations · Rotational Shift · Initial Contract till March 27. KEY REQUIREMENT This role requires US citizenship and residence on US soil... 
    Suggested
    Hourly pay
    Contract work
    For contractors
    Shift work
    Night shift
    Weekend work

    C-Serv

    San Jose, CA
    26 days ago
  • Software Engineer II TENEX is an AI-native, automation-first, built-for-scale Managed Detection and Response (MDR) provider....  ...deploy scalable and reliable backend services...  ...Developer or Professional Data Engineer are a plus...  ...with artificial intelligence and have experience... 
    Intelligence
    Remote work
    Monday to Friday

    TenEx

    San Jose, CA
    2 days ago
  •  ...Location: 5 On-Site Days a Week in...  ...Headquarters Our Engineering team is driven...  ...SRE Engineer II, you will be responsible for managing our multi-cloud...  ...system reliability and scalability...  ...for automated software delivery and deployment...  ...environments to protect data, applications,... 
    Suggested
    Work experience placement
    Immediate start

    Illumio

    Sunnyvale, CA
    2 days ago
  • $104.9k - $174.7k

     ...Site Reliability Engineer The Site Reliability Engineer role is responsible for improving the reliability...  ...gaps. Respond to system-management alerts and operational exceptions within...  ...troubleshoot, and support hardware, software, storage, network, cloud, Kubernetes,... 
    Suggested
    Temporary work
    Local area

    Lexus Nexus

    San Jose, CA
    11 minutes ago
  • $170k - $220k

     ...We are seeking a Senior Software Engineer in Test (SET II) to play a pivotal role in...  ...existing tools to streamline data collection trackers....  ...cameras in Android apps. Managing complex image manipulation...  ...! We may use artificial intelligence (AI) tools to support parts... 
    Intelligence
    Full time

    Autoroboto

    Mountain View, CA
    more than 2 months ago
  •  ...Data Engineer IIRootshell Enterprise Technologies Inc. is a recognized provider of professional IT Consulting services in the US. We are actively seeking Data Engineer II for one of our client.Role: Data Engineer II Location: Santa Clara, CA Duration: Long TermThe project... 

    Rootshell Inc

    Santa Clara, CA
    2 days ago
  • $168k - $270.25k

     ...developments in Artificial Intelligence, High-Performance...  ...team at NVIDIA as a Senior Site reliability engineer focused on HPC storage and...  ...Storage solutions tailored for data-intensive applications, optimizing...  ...automate deployment and management of large-scale... 
    Intelligence
    Full time

    Nvidia

    Santa Clara, CA
    a month ago
  • $65 - $85 per hour

     ...Site Reliability Engineer Sustainable Talent is partnering with...  ...groups within NVIDIA Software such as Graphics...  ...Learning, Artificial Intelligence and Driverless Cars...  ...solutions, mine through data to uncover real...  ...latest Configuration Management & Infrastructure Automation... 
    Intelligence
    Full time
    Contract work
    Worldwide

    Sustainable Talent

    Santa Clara, CA
    3 days ago
  • $248k - $396.75k

     ...Full time JR2023973 Site Reliability Engineering (SRE) at NVIDIA is an engineering...  ...availability. It combines software and systems engineering...  ..., databases, capacity management, continuous delivery, and...  ...AI agents, AI skills, and intelligent automation that accelerate... 
    Intelligence
    Full time

    NVIDIA

    Santa Clara, CA
    4 days ago
  • $122.5k - $175k

     ...from cyberattacks and data loss by securely...  ...amplified by machine intelligence to solve the world’s...  ...looking for a Staff Site Reliability Engineer to join our team. This...  ...infrastructure and managing platforms like Kubernetes...  ...deploy systems and software in diverse... 
    Intelligence
    Full time
    Work at office
    Local area
    3 days per week

    Zscaler

    San Jose, CA
    12 hours ago
  • $122.5k - $175k

     ...from cyberattacks and data loss by securely...  ...amplified by machine intelligence to solve the world’s...  ...looking for a Staff Site Reliability Engineer to join our team. This...  ...infrastructure and managing platforms like Kubernetes...  ...deploy systems and software in diverse... 
    Intelligence
    Full time
    Work at office
    Local area
    3 days per week

    Zscaler

    San Jose, CA
    2 days ago
  • $143.7k - $194.4k

     ...companies worldwide to manage day-to-day...  ...power of Artificial Intelligence and the large...  ...solving highly complex engineering and algorithmic...  ...and talented Software Development Engineers...  ..., AI, ML, Big Data and more. This team...  ...design patterns, reliability and scaling) of... 
    Intelligence
    Internship
    Local area
    Worldwide
    Flexible hours

    Amazon

    Santa Clara, CA
    a month ago
  • ScaleFlux is seeking a Senior Product Manager to drive strategy, roadmap, and execution...  ...NVMe SSD products aimed at AI, cloud, and data center markets. You will shape storage...  ...SSD controller architectures, firmware intelligence, and computational storage to meet AI training... 
    Intelligence

    ScaleFlux

    Milpitas, CA
    1 day ago
  •  ...Solve complex reliability challenges at scale...  ...Influence architecture and engineering culture at a company level...  ...Architect, implement, and manage highly available and...  ...Database services or shared data platforms for broad...  ...We may use artificial intelligence (AI) tools to support... 
    Intelligence

    Saviynt

    Milpitas, CA
    13 days ago
  • $207.4k - $259.2k

     ...physical artificial intelligence (“AI”) solutions, and...  ...Sr. Staff Site Reliability Engineer (SRE) to join our growing...  ...Architect and optimize data pipelines to ensure...  ...deployment, and release management.Champion cloud-first...  ...is built into the software development lifecycle... 
    Intelligence
    Permanent employment
    Local area
    Worldwide
    Visa sponsorship

    Archer Aviation

    San Jose, CA
    9 days ago
  • $133.92k

     ...seeking a highly motivated and skilled Software Development Engineer II to join our team. This role offers...  ...the improvement of the performance, reliability, and success metrics of product...  ...ownership of tasks while effectively managing time and priorities. Collaboration Skills... 
    Full time
    Internship
    Summer internship
    Local area
    Remote work

    F5

    San Jose, CA
    16 hours ago
  •  ...procurement, transport logistics, data center design and construction, equipment management, and daily operations. Bitdeer...  ...high demand for artificial intelligence. Headquartered in Singapore,...  ...capabilities. Partner closely with engineering, infrastructure, operations,... 
    Intelligence
    Full time
    Local area

    Bitdeer Technologies Group

    San Jose, CA
    13 days ago
  • $224k - $356.5k

     ...midst of a revolution where artificial intelligence is helping us accelerate all aspects...  ...intelligence to robots.NVIDIA is seeking an Engineering Manager to lead our Robotics Neural...  ...lead a group of world-class robotics software and applied research engineers focused... 
    Intelligence
    Full time

    Nvidia

    Santa Clara, CA
    a month ago
  • $168k - $270.25k

    Site Reliability Engineering (SRE) at NVIDIA is an engineering discipline to design...  ...using the combination of software and systems engineering...  ...coding, database, capacity management, continuous delivery and deployment...  ...in Artificial Intelligence, High-Performance Computing... 
    Intelligence
    Full time

    Nvidia

    Santa Clara, CA
    5 days ago
  •  ...Introduction At IBM Software, we transform client...  ...As a Site Reliability Engineer, you will work in an...  ...day operations, alert management, incident support, migration...  ...maintaining SQL, NoSQL, and data streaming...  ...business operations with intelligence-from machine learning... 
    Intelligence
    Full time
    Contract work
    Part time
    Fixed term contract
    Internship
    Worldwide
    Flexible hours
    Shift work

    IBM

    San Jose, CA
    3 days ago
  •  ...You Will Contribute:Software Engineer II (Full Stack)Are you excited...  ...decisions with their data? Join our growing...  ...that power water management and conservation around...  ...while ensuring quality, reliability, and...  ...three days per week on-site in Los Gatos, CAReceive... 
    Full time
    Temporary work
    Local area
    Flexible hours
    3 days per week

    Badger Meter

    Los Gatos, CA
    3 days ago
  • $180k - $220k

     ...is building an operational intelligence platform for digital infrastructure...  ...experience, and automated data engineering pipelines. Our...  ...Job Overview  Product Manager is responsible for working...  ...observability, APM, or IT operations software.  ~ Strong understanding... 
    Intelligence
    Night shift

    Selector Software

    Santa Clara, CA
    a month ago
  •  ...design, coding, testing, and debugging of platform and system software tasks of small to medium complexity. Working on a variety of technical problems of varying scope, the Software Development Engineer II implements high-quality code and deploys features or APIs under... 
    Full time
    Work at office
    Local area
    Remote work
    Home office

    F5 Networks

    San Jose, CA
    13 days ago
  • $165.2k - $223.6k

     ...of companies worldwide to manage day-to-day operations. We will...  ...and help build the secure data foundation that powers AWS'...  ...their information.As a Software Development Engineer II, you'll take ownership of production...  ...root causes to maintain reliable data deletion and opt-out... 
    Permanent employment
    Internship
    Local area
    Worldwide
    Flexible hours

    AmazonWebServices

    Santa Clara, CA
    a month ago
  • $272k - $431.25k

    NVIDIA is looking for a Cloud Site Reliability Engineering Architect to work in IPP's...  ...Deep Learning, Artificial Intelligence and Autonomous Vehicles to...  ...of thousands of NVIDIA's software engineers worldwide. The...  ...solutions, mine through data to uncover real problems and... 
    Intelligence
    Full time
    Work experience placement
    Worldwide

    Nvidia

    Santa Clara, CA
    a month ago
  • $226.14k

     ...time to building systems that operate reliably on a global scale. When you work here,...  ...we’d love to meet you. Job Title: Software Engineer II Location: 50 West San Fernando Street...  ...integrations following secure coding and data governance practices. Build and operate... 
    Full time
    Temporary work
    Work at office
    Remote work
    Worldwide

    The Trade Desk

    San Jose, CA
    5 days ago
  • $125.7k - $203.1k

    Software Engineer Embedded Systems II Join a vibrant community of passionate...  ...platforms and intelligent connected...  ...ensure product reliability. Apply systems...  ...concurrency, memory management, and low-level...  ...how data and infrastructure...  ...Cisco careers site to discover more... 
    Full time
    Temporary work
    Apprenticeship
    Work experience placement
    Local area
    Flexible hours

    Webex Events (formerly Socio)

    Milpitas, CA
    2 days ago
  • $75k - $150k

     ...a Salesforce Developer II to design, develop, and...  ...experiences, automation, and data solutions that support...  ...activities, release management, and production support...  ..., Information Systems, Engineering, or a related field, or...  ..., or resumes to this site or to any Columbia Bank... 

    Columbia Bank

    San Jose, CA
    16 hours ago
  • $100k

     ...evolve to unify innovations in software models, compilers, platforms...  ...connection to Product, Engineering, Field Applications, and Marketing...  ...training, inference, data and analytics, HPC, databases...  ...opportunity data, using market intelligence and AI-enabled tools to... 
    Intelligence
    Permanent employment

    Tenstorrent

    Santa Clara, CA
    15 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Software Engineering Manager II, Site Reliability Engineering, Data Intelligence. Be the first to apply!