Software Engineering Manager II, Site Reliability Engineering, Data Intelligence
$207k - $300kHire, develop, and mentor a high-performing, SRE team to support career growth and team health.Set team goals, prioritize resources, and define technical roadmaps aligned with partner teams and key stakeholders.Guide the architecture and review of resilient, high-performance systems powering core AI infrastructure.Drive incident response, maintain high reliability standards, and actively automate operational toil.Partner across development teams to align technical direction while leveraging AI to accelerate team productivity.Minimum qualifications:Bachelor’s degree in Computer Science, a related field, or equivalent practical experience.8 years of experience with software development in one or more programming languages (e.g., C++, Java, Python), or with data structures/algorithms.3 years of experience in designing, analyzing, and troubleshooting large-scale distributed systems.3 years of experience managing and growing engineering teams, including performance management and career development.Preferred qualifications:Experience managing distributed teams across multiple sites or timezones.Experience developing long-term technical roadmaps, driving organizational change, and influencing cross-functional stakeholders (Dev, PM, Leadership).Proven track record of hiring, mentoring, and leading high-performing Site Reliability Engineering or Software Engineering teams.Systematic problem-solving and troubleshooting skills in complex, ambiguous software systems.Passion for AI infrastructure and driving AI transformation within engineering workflows.Site Reliability Engineering (SRE) combines software and systems engineering to build and run large-scale, massively distributed, fault-tolerant systems. SRE ensures that Google's services—both our internally critical and our externally-visible systems—have reliability, uptime appropriate to users' needs and a fast rate of improvement. Additionally SRE’s will keep an ever-watchful eye on our systems capacity and performance.Much of our software development focuses on optimizing existing systems, building infrastructure and eliminating work through automation. On the SRE team, you’ll have the opportunity to manage the complex challenges of scale which are unique to Google, while using your expertise in coding, algorithms, complexity analysis and large-scale system design. SRE's culture of intellectual curiosity, problem solving and openness is key to its success. Our organization brings together people with a wide variety of backgrounds, experiences and perspectives. We encourage them to collaborate, think big and take risks in a blame-free environment. We promote self-direction to work on meaningful projects, while we also strive to create an environment that provides the support and mentorship needed to learn and grow.To learn more: check out our books on Site Reliability Engineering or read a career profile about why a Software Engineer chose to join SRE.Join Google’s Core AI Foundations SRE team! We build and scale the critical infrastructure—authorization, ML training storage, and RPC scheduling—powering products like Gemini, Workspace, NotebookLM, and Cloud.You will manage and grow a North American SRE team. You will partner closely with development teams and our Sydney counterpart to run a follow-the-sun rotation and ensure exceptional reliability for Google’s frontier AI capabilities.The Core team builds the technical foundation behind Google’s flagship products. We are owners and advocates for the underlying design elements, developer platforms, product components, and infrastructure at Google. These are the essential building blocks for excellent, safe, and coherent experiences for our users and drive the pace of innovation for every developer. We look across Google’s products to build central solutions, break down technical barriers and strengthen existing systems. As the Core team, we have a mandate and a unique opportunity to impact important technical decisions across the company.Individual pay is determined by factors including job-related skills, experience, and relevant education or training. US: $207000 - $300000 (USD) + 20% bonus target + equity + benefitsLearn more about benefits at Google.Bachelor’s degree in Computer Science, a related field, or equivalent practical experience.8 years of experience with software development in one or more programming languages (e.g., C++, Java, Python), or with data structures/algorithms.3 years of experience in designing, analyzing, and troubleshooting large-scale distributed systems.3 years of experience managing and growing engineering teams, including performance management and career development.
$207k - $300k
Manage a team of Software/Systems Engineers on projects for users and remain directly responsible... ...sustainable multi-site on-call rotations across... ...practical expertise in Site Reliability Engineering practices, including... ...of Google's foundational data pipeline and will manage...Suggested- ...Job Description Job Description Site Reliability Engineer II Bay Area, offices in San Jose · Hybrid · 24/7 FedRAMP Operations · Rotational Shift · Initial Contract till March 27. KEY REQUIREMENT This role requires US citizenship and residence on US soil...SuggestedHourly payContract workFor contractorsShift workNight shiftWeekend work
- Software Engineer II TENEX is an AI-native, automation-first, built-for-scale Managed Detection and Response (MDR) provider.... ...deploy scalable and reliable backend services... ...Developer or Professional Data Engineer are a plus... ...with artificial intelligence and have experience...IntelligenceRemote workMonday to Friday
- ...Location: 5 On-Site Days a Week in... ...Headquarters Our Engineering team is driven... ...SRE Engineer II, you will be responsible for managing our multi-cloud... ...system reliability and scalability... ...for automated software delivery and deployment... ...environments to protect data, applications,...SuggestedWork experience placementImmediate start
$104.9k - $174.7k
...Site Reliability Engineer The Site Reliability Engineer role is responsible for improving the reliability... ...gaps. Respond to system-management alerts and operational exceptions within... ...troubleshoot, and support hardware, software, storage, network, cloud, Kubernetes,...SuggestedTemporary workLocal area$170k - $220k
...We are seeking a Senior Software Engineer in Test (SET II) to play a pivotal role in... ...existing tools to streamline data collection trackers.... ...cameras in Android apps. Managing complex image manipulation... ...! We may use artificial intelligence (AI) tools to support parts...IntelligenceFull time- ...Data Engineer IIRootshell Enterprise Technologies Inc. is a recognized provider of professional IT Consulting services in the US. We are actively seeking Data Engineer II for one of our client.Role: Data Engineer II Location: Santa Clara, CA Duration: Long TermThe project...
$168k - $270.25k
...developments in Artificial Intelligence, High-Performance... ...team at NVIDIA as a Senior Site reliability engineer focused on HPC storage and... ...Storage solutions tailored for data-intensive applications, optimizing... ...automate deployment and management of large-scale...IntelligenceFull time$65 - $85 per hour
...Site Reliability Engineer Sustainable Talent is partnering with... ...groups within NVIDIA Software such as Graphics... ...Learning, Artificial Intelligence and Driverless Cars... ...solutions, mine through data to uncover real... ...latest Configuration Management & Infrastructure Automation...IntelligenceFull timeContract workWorldwide$248k - $396.75k
...Full time JR2023973 Site Reliability Engineering (SRE) at NVIDIA is an engineering... ...availability. It combines software and systems engineering... ..., databases, capacity management, continuous delivery, and... ...AI agents, AI skills, and intelligent automation that accelerate...IntelligenceFull time$122.5k - $175k
...from cyberattacks and data loss by securely... ...amplified by machine intelligence to solve the world’s... ...looking for a Staff Site Reliability Engineer to join our team. This... ...infrastructure and managing platforms like Kubernetes... ...deploy systems and software in diverse...IntelligenceFull timeWork at officeLocal area3 days per week$122.5k - $175k
...from cyberattacks and data loss by securely... ...amplified by machine intelligence to solve the world’s... ...looking for a Staff Site Reliability Engineer to join our team. This... ...infrastructure and managing platforms like Kubernetes... ...deploy systems and software in diverse...IntelligenceFull timeWork at officeLocal area3 days per week$143.7k - $194.4k
...companies worldwide to manage day-to-day... ...power of Artificial Intelligence and the large... ...solving highly complex engineering and algorithmic... ...and talented Software Development Engineers... ..., AI, ML, Big Data and more. This team... ...design patterns, reliability and scaling) of...IntelligenceInternshipLocal areaWorldwideFlexible hours- ScaleFlux is seeking a Senior Product Manager to drive strategy, roadmap, and execution... ...NVMe SSD products aimed at AI, cloud, and data center markets. You will shape storage... ...SSD controller architectures, firmware intelligence, and computational storage to meet AI training...Intelligence
- ...Solve complex reliability challenges at scale... ...Influence architecture and engineering culture at a company level... ...Architect, implement, and manage highly available and... ...Database services or shared data platforms for broad... ...We may use artificial intelligence (AI) tools to support...Intelligence
$207.4k - $259.2k
...physical artificial intelligence (“AI”) solutions, and... ...Sr. Staff Site Reliability Engineer (SRE) to join our growing... ...Architect and optimize data pipelines to ensure... ...deployment, and release management.Champion cloud-first... ...is built into the software development lifecycle...IntelligencePermanent employmentLocal areaWorldwideVisa sponsorship$133.92k
...seeking a highly motivated and skilled Software Development Engineer II to join our team. This role offers... ...the improvement of the performance, reliability, and success metrics of product... ...ownership of tasks while effectively managing time and priorities. Collaboration Skills...Full timeInternshipSummer internshipLocal areaRemote work- ...procurement, transport logistics, data center design and construction, equipment management, and daily operations. Bitdeer... ...high demand for artificial intelligence. Headquartered in Singapore,... ...capabilities. Partner closely with engineering, infrastructure, operations,...IntelligenceFull timeLocal area
$224k - $356.5k
...midst of a revolution where artificial intelligence is helping us accelerate all aspects... ...intelligence to robots.NVIDIA is seeking an Engineering Manager to lead our Robotics Neural... ...lead a group of world-class robotics software and applied research engineers focused...IntelligenceFull time$168k - $270.25k
Site Reliability Engineering (SRE) at NVIDIA is an engineering discipline to design... ...using the combination of software and systems engineering... ...coding, database, capacity management, continuous delivery and deployment... ...in Artificial Intelligence, High-Performance Computing...IntelligenceFull time- ...Introduction At IBM Software, we transform client... ...As a Site Reliability Engineer, you will work in an... ...day operations, alert management, incident support, migration... ...maintaining SQL, NoSQL, and data streaming... ...business operations with intelligence-from machine learning...IntelligenceFull timeContract workPart timeFixed term contractInternshipWorldwideFlexible hoursShift work
- ...You Will Contribute:Software Engineer II (Full Stack)Are you excited... ...decisions with their data? Join our growing... ...that power water management and conservation around... ...while ensuring quality, reliability, and... ...three days per week on-site in Los Gatos, CAReceive...Full timeTemporary workLocal areaFlexible hours3 days per week
$180k - $220k
...is building an operational intelligence platform for digital infrastructure... ...experience, and automated data engineering pipelines. Our... ...Job Overview Product Manager is responsible for working... ...observability, APM, or IT operations software. ~ Strong understanding...IntelligenceNight shift- ...design, coding, testing, and debugging of platform and system software tasks of small to medium complexity. Working on a variety of technical problems of varying scope, the Software Development Engineer II implements high-quality code and deploys features or APIs under...Full timeWork at officeLocal areaRemote workHome office
$165.2k - $223.6k
...of companies worldwide to manage day-to-day operations. We will... ...and help build the secure data foundation that powers AWS'... ...their information.As a Software Development Engineer II, you'll take ownership of production... ...root causes to maintain reliable data deletion and opt-out...Permanent employmentInternshipLocal areaWorldwideFlexible hours$272k - $431.25k
NVIDIA is looking for a Cloud Site Reliability Engineering Architect to work in IPP's... ...Deep Learning, Artificial Intelligence and Autonomous Vehicles to... ...of thousands of NVIDIA's software engineers worldwide. The... ...solutions, mine through data to uncover real problems and...IntelligenceFull timeWork experience placementWorldwide$226.14k
...time to building systems that operate reliably on a global scale. When you work here,... ...we’d love to meet you. Job Title: Software Engineer II Location: 50 West San Fernando Street... ...integrations following secure coding and data governance practices. Build and operate...Full timeTemporary workWork at officeRemote workWorldwide$125.7k - $203.1k
Software Engineer Embedded Systems II Join a vibrant community of passionate... ...platforms and intelligent connected... ...ensure product reliability. Apply systems... ...concurrency, memory management, and low-level... ...how data and infrastructure... ...Cisco careers site to discover more...Full timeTemporary workApprenticeshipWork experience placementLocal areaFlexible hours$75k - $150k
...a Salesforce Developer II to design, develop, and... ...experiences, automation, and data solutions that support... ...activities, release management, and production support... ..., Information Systems, Engineering, or a related field, or... ..., or resumes to this site or to any Columbia Bank...$100k
...evolve to unify innovations in software models, compilers, platforms... ...connection to Product, Engineering, Field Applications, and Marketing... ...training, inference, data and analytics, HPC, databases... ...opportunity data, using market intelligence and AI-enabled tools to...IntelligencePermanent employment
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Software Engineering Manager II, Site Reliability Engineering, Data Intelligence. Be the first to apply!
- site reliability engineer San Jose, CA
- site reliability engineer sre San Jose, CA
- senior data center engineer San Jose, CA
- senior cloud data engineer San Jose, CA
- senior data integration developer San Jose, CA
- data developer San Jose, CA
- data engineer San Jose, CA
- big data cloud engineer San Jose, CA
- data platform engineer San Jose, CA
- data engineering intern summer San Jose, CA




