Engineering Manager, DSC AI Inference Platform
$207k - $301kPeople Management and Talent Development: Lead, mentor, and grow a high-performing team of systems and ML engineers. Drive a culture of excellence, psychological safety, and continuous learning. Guide career paths, define OKRs, and conduct performance evaluations.Strategic and Technical Roadmap: Define the technical goal and strategy for enhancing the LLM serving stack, focusing on performance, scalability, and resource efficiency.Architectural Leadership: Drive the design and implementation of advanced serving architectures, including disaggregated serving, to optimize resource utilization and latency.*Infrastructure Oversight:* Oversee the building and maintenance of critical infrastructure and tooling for in-depth performance analysis, profiling, and benchmarking of LLM models on GPU accelerators.Cross-Functional Collaboration: Partner closely with Research, SRE, Product, and Core GPU library teams to optimize and deploy LLMs in production globally. Align team efforts with broader organizational AI priorities.Minimum qualifications:Bachelor's degree or equivalent practical experience.8 years of experience programming in C++ or Python.5 years of experience optimizing, profiling, and scaling production-grade systems on GPU accelerators or specialized AI hardware.5 years of experience directly managing and leading engineering teams focused on machine learning infrastructure, AI platforms, or high-performance distributed computing systems.5 years of experience in a people management or team leadership role.3 years of experience managing engineering organizations across multi-team infrastructure dependencies.Preferred qualifications:Master's degree or PhD degree in Computer Science or a related technical field.5 years of experience working in a complex, matrixed organization.4 years of experience implementing advanced LLM serving architectures and optimization techniques, such as disaggregated serving, continuous batching, or specialized compiler technologies (e.g., XLA).3 years of experience utilizing deep-dive ML profiling tools (e.g., Nsight, xprof) to troubleshoot and resolve low-level bottlenecks within major frameworks like JAX, PyTorch, or TensorFlow.Like Google's own ambitions, the work of a Software Engineer goes beyond just Search. Software Engineering Managers have not only the technical expertise to take on and provide technical leadership to major projects, but also manage a team of Engineers. You not only optimize your own code but make sure Engineers are able to optimize theirs. As a Software Engineering Manager you manage your project goals, contribute to product strategy and help develop your team. Teams work all across the company, in areas such as information retrieval, artificial intelligence, natural language processing, distributed computing, large-scale system design, networking, security, data compression, user interface design; the list goes on and is growing every day. Operating with scale and speed, our exceptional software engineers are just getting started -- and as a manager, you guide the way.With technical and leadership expertise, you manage engineers across multiple teams and locations, a large product budget and oversee the deployment of large-scale projects across multiple sites internationally.The Distributed Cloud (DSC) AI Inference Platform team operates at the critical intersection of Large Language Models (LLMs) and high-performance computing. Our mission is to engineer the future of AI serving infrastructure, driving foundational improvements in efficiency, latency, and throughput. We develop innovative solutions, including disaggregated serving architectures, and build the essential tools to analyze and optimize LLM performance on cutting-edge GPU platforms. Our work directly enables Google to deploy and scale state-of-the-art AI models (like Gemini) effectively and efficiently across Google's global infrastructure, products, and Cloud.The Google Cloud AI Research team addresses AI challenges motivated by Google Cloud’s mission of bringing AI to tech, healthcare, finance, retail and many other industries. We work on a range of unique problems focused on research topics that maximize scientific and real-world impact, aiming to push the state-of-the-art in AI and share findings with the broader research community. We also collaborate with product teams to bring innovations to real-world impact that benefits our customers. Individual pay is determined by factors including job-related skills, experience, and relevant education or training. US: $207000 - $301000 (USD) + 20% bonus target + equity + benefitsLearn more about benefits at Google.Bachelor's degree or equivalent practical experience.8 years of experience programming in C++ or Python.5 years of experience optimizing, profiling, and scaling production-grade systems on GPU accelerators or specialized AI hardware.5 years of experience directly managing and leading engineering teams focused on machine learning infrastructure, AI platforms, or high-performance distributed computing systems.5 years of experience in a people management or team leadership role.3 years of experience managing engineering organizations across multi-team infrastructure dependencies.
$224k - $356.5k
...the unlimited potential of AI to define the next era of... ...open-source benchmarking platform, AIPerf, is the growing standard... ...across various inference frameworks. Hyperscalers,... ...scaling. As Technical Lead Manager, you will lead the engineering team within NVIDIA’s Dynamo...PlatformFull timeLocal areaRemote workWorldwide$278.1k - $347.6k
...building the next generation of AI-driven game experiences,... ...runtime. As our Principal Engineer for On-Device AI Inference & Systems, you will be the... ...profile with browser and platform tools (Chrome/Dawn GPU... ...render path, and frame-budget management alongside the renderer....PlatformWork at officeWorldwideRelocation package- ...Systems builds the world's largest AI chip, 56 times larger than GPUs. This... ...industry-leading training and inference speeds; over 10 times faster than GPU... ...About The RoleWe're hiring a Principal Engineer for our Inference Cloud Platform. This team owns the cloud layer...Platform
$224k - $356.5k
NVIDIA is seeking an exceptional Manager, Deep Learning Inference Software, to lead a world-class engineering team advancing the state of AI model deployment. You will shape the software... ...level optimization across CPU and GPU platforms.Proven ability to mentor engineers,...PlatformFull timeRemote workWorldwide$168k - $270.25k
...computing, and modern AI. Today, NVIDIA high-performance processing platforms power breakthroughs across... ...for training and inference. As AI models, GPU architectures... ...large-scale systems engineering. To address these... ...seeking an Engineering Manager to spearhead our strategy...PlatformFull time$201.3k - $352.3k
...DescriptionIt all started when engineer Fred Luddy wrote code... ..., ServiceNow is the AI control tower for... ...reinvention. Our ServiceNow AI platform brings together any AI,... ...engineers, and product managers with a dual mission. We... ...window efficiency, and inference costs.Lead a High-...PlatformWork experience placementWork at officeImmediate startRemote workFlexible hoursShift work$224k - $356.5k
...NVIDIA is the platform upon which every new AI-powered application is built. We are seeking a deeply technical software manager to lead production AI inference for NVIDIA Inference Microservices (NIM), the... ...optimized inference engines, model profiles/recipes, validated...Platform- ...generation computing experiences—from AI and data centers, to PCs, gaming and... ...ROLEWe are seeking a Principal GenAI Inference Optimization Engineer to join our Models and Applications team... ...AI inference workloads on AMD GPU platforms. You will contribute to optimizing latency...Platform
$272k - $431.25k
NVIDIA's Object Storage Platform team builds and... ...infrastructure that stores, manages, and serves exabytes of... ...backbone for NVIDIA's AI infrastructure, enabling researchers and engineers to reliably store massive... ...accelerating training and inference pipelines.We are...PlatformFull time$224k - $356.5k
NVIDIA is seeking an Engineering Manager to lead the development of an agentic platform for observing, debugging, and optimizing... ...visibility into model behavior, inference performance, reliability, and cost... ...will work across the NVIDIA AI software stack with teams focused...PlatformFull time- ...Principal Machine Learning Engineer, you will drive the... ...teams to build AI functionalities into Atlassian... ...training and online inference at scaleOversee end-to... ...closely with product managers, designers, and engineering... ...larger products and platforms Compensation At...PlatformLocal areaRemote work
$272k - $431.25k
...for a Machine Learning (ML) Engineer to join the GPU accelerated Apache... ...and ML/DL model training and inference pipelines, spanning many... ...You will apply the latest ML/AI methods to empower enterprises... ...large-scale data processing platforms, such as Apache Spark.Proven...PlatformFull time$207k - $301k
Lead a team of software engineers to design, build, deploy, and in some... ...systems that directly manage the global data center networking... ...disaster recovery protocols.The AI and Infrastructure team is redefining... ..., and providing the essential platforms that enable developers to...PlatformWorldwide$207k - $340k
...a Principal Staff Software Engineer to lead LinkedIn’s GPU-Based Retrieval Platform, a foundational AI infrastructure stack that powers... ...with NCCL, distributed inference, multi-GPU communication, or... ...precision, batching, memory management, and throughput or latency tuning...PlatformFor contractorsWork at officeRemote workWork from homeFlexible hours$296.3k
...at minimum.The Role:We are seeking a Principal AI Engineer to lead the design and advancement of our AI platform. You will play a key role in shaping the infrastructure... ...that powers large-scale training and cloud inference. This includes accelerating training throughput,...PlatformFull timeLocal areaRemote workWork from homeFlexible hours$224k - $356.5k
...then seamlessly runs them on NVIDIA DRIVE Platforms inside the vehicle.The NvSci team... ...processing pipeline. We are seeking a seasoned Engineering Manager to lead our rapidly growing NvStreams... ...for an existing vacancy. NVIDIA uses AI tools in its recruiting processes....PlatformFull time$175k - $275k
...builds the world's largest AI chip, 56 times larger... ...-leading training and inference speeds; over 10 times... ...complex electrical engineering boards and full systems... ...constraints. People management is not required (mentoring... ...a breakthrough AI platform beyond the constraints...Platform$139.7k - $216.5k
...California in 2004 when a visionary engineer, Fred Luddy, saw the potential... ...leader, bringing innovative AI-enhanced technology to over 8,... .... Our intelligent cloud-based platform seamlessly connects people,... ...Learning Engineering Manager, GAI Search Relevance As the leader...PlatformWork at officeRemote workFlexible hours$272k - $431.25k
As a Senior Engineering Manager for Agentic Systems & Platform Architecture, you will lead the strategy and execution for NVIDIA’s agentic... ...databases, and GPU-optimized training and inference workflows.Lead integration of the AI Data Platform into NVIDIA’s on-prem AI Factory...PlatformFull time$320k
...is a world leader in physical AI, powering self-driving cars, humanoid... ..., and more. Our software platforms are at the core of this... ...globe. This team is the execution engine behind NVIDIA’s Vision AI strategy... ...robust, low-latency inference at scale. You have led teams that...PlatformFull time$207k - $300k
...scale a high-performing team of 8 engineers, setting the technical... ...the integration of generative AI into the software development... ...building or operating large-scale, managed services.3 years of experience... ...the strategy for Google Cloud Platform (GCP’s) data location governance...PlatformShift work$278.1k - $417.1k
...building the next generation of AI-driven game experiences,... ...runtime. As our Principal Engineer for On-Device AI Inference & Systems, you will be the... ...profile with browser and platform tools (Chrome/Dawn GPU... ...render path, and frame-budget management alongside the renderer....PlatformWork at officeWorldwideRelocation package$170k - $277k
...DescriptionTeam Overview:The Notification AI team builds large-scale, high... ...value members get from the platform, driving over 50% of total... ...closely with Product, Engineering, and Data Science to deliver... ...We’re looking for a technical manager with strong leadership,...PlatformFor contractorsWork at officeFlexible hours$204k - $337k
...New York, NY. Team Overview: The Video AI team sits at the heart of our LinkedIn’... ...for all.Responsibilities: As a Senior Engineering Manager you will lead a team of 15-20 engineers... ...and efficiency: Partner with infra and platform teams to optimize retrieval and serving...PlatformFor contractorsWork at officeFlexible hours$194k
.... About the TeamThe Search AI Product & Mobile team sits... ...closely with ML and backend platform teams, and ship features that... ...RoleWe're looking for a Manager - Mobile Enginerring, who will lead the Search API... ...of ML models (on-device inference, server-side ranking signal...PlatformTemporary work$224k - $356.5k
We are now looking for an AI Developer Technology Engineering Manager:Join our global Developer Technology (DevTech) team at NVIDIA, where we drive innovation and enhance the value of our platforms for developers. As a key member of our AI DevTech team, you'll lead a team...PlatformFull timeTemporary work$272k - $431.25k
...choice for the most ambitious AI research, and the... ...them. NVIDIA seeks a Senior Engineering Manager to define and drive NVIDIA's... ...peak performance across the platform, from single-accelerator workloads... ...across training, post-training, inference, and robotics, bridging new...PlatformFull time$224k - $356.5k
...Today, we are increasingly known as “the AI computing company.” We're looking to... ...world.We are looking for an excellent engineering manager to own and deliver an end to end manageability... ...working on server firmware (BMC) and platform software development5+ years of...PlatformFull timeWork at office$272k - $431.25k
..., we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which... ...(IPP) Team is seeking a Principal Software Engineer to lead the next generation of AI-powered engineering platforms. In this role, you will define and build agentic...PlatformFull time$168k - $258.75k
Inference is the fastest growing and most competitive area in Generative AI today. It is where AI models impact our daily life,... ...techniques. As a Senior Product Manager for AI Platform Inference you will be... ...Computer Science, Computer Engineering, or similar experience (or...PlatformFull time
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Engineering Manager, DSC AI Inference Platform. Be the first to apply!


