ML Systems Engineer
$150k - $350kChipAgents
About ChipAgents ChipAgents is redefining the future of chip design and verification with agentic AI workflows. Our platform leverages cutting‑edge generative AI to assist engineers in RTL design, simulation, and verification, dramatically accelerating chip development. Founded by experts in AI and semiconductor engineering, we partner with top semiconductor firms, cloud providers, and innovative startups to build intelligent AI agents. The company is a Series A company backed by tier‑1 VC firms. ChipAgents is deployed in production to companies that have shipped 16B chips. Position Overview We are seeking an ML Systems Engineer to optimize the performance and efficiency of large language model inference powering our agentic AI platform. This is a technical role focused on low‑level systems optimization. You will implement performance optimizations, build evaluation harnesses, and architect multi‑node clusters for training and inference that push the limits of LLM throughput and latency. Your work will directly impact the responsiveness and cost‑efficiency of AI agents used by leading semiconductor companies to design chips. Key Responsibilities Design, deploy, and optimize LLM inference systems across multi‑node clusters, maximizing throughput and minimizing latency for production workloads. Implement and benchmark concrete inference optimizations. Profile and analyze inference bottlenecks at the systems level—from GPU kernel execution to memory bandwidth constraints. Build robust evaluation harnesses and benchmarking frameworks that measure accuracy, throughput, latency, and resource utilization across various parallelism strategies. Collaborate with research scientists to integrate new model architectures and optimizations into production inference infrastructure. Investigate and apply emerging techniques from research papers and open‑source projects to continuously improve inference performance. Qualifications B.S., M.S., or PhD in Computer Science, Electrical Engineering, or related field (or equivalent experience). Experience with large‑scale ML systems, GPU computing, or high‑performance inference optimization. Strong proficiency in Python and C++/CUDA; hands‑on experience with SGLang, vLLM, PyTorch, or similar inference frameworks. Deep understanding of GPU architecture, memory hierarchies, and parallel computing paradigms. Experience deploying and optimizing LLMs in production: model serving, batching strategies, distributed inference, or quantization. Strong systems‑level debugging and profiling skills; comfort working at multiple layers of the stack from CUDA kernels to application logic. Familiarity with distributed computing frameworks (Ray, multi‑node training/inference) is a plus. Self‑directed problem solver who is interested in working on ambitious optimization challenges. Why Join Us Work on cutting‑edge LLM inference optimization problems with real‑world production impact. Access to substantial GPU compute resources for experimentation and benchmarking. Collaborate with a world‑class team spanning AI research, systems engineering, and EDA. Shape the performance characteristics of AI systems used by leading semiconductor companies. What we offer $150K/yr – $350K/yr + Offers Equity. We are open to discuss above‑scale compensation with exceptional candidates on a case‑by‑case basis. Unlimited PTO and full benefits (medical, vision, dental, 401k). Two engineering‑centric offices with free parking, private gym, and free lunch, drinks and snacks. #J-18808-Ljbffr ChipAgents
$224k - $356.5k
...the next phase, we are building agentic systems that can reason about, build, evaluate, and... ...about creating the meta-layer of modern ML: the agents, tooling, pipelines, and feedback... .... We are looking for exceptional engineers who are passionate about the idea of AI-native...SuggestedFull time- ...hiring for a role in the Cosmos team to design and build agentic ML systems that accelerate model development. You will own large Python/... ...spanning data generation, evaluation, and deployment. We seek engineers with deep ML system experience, strong software fundamentals,...Suggested
- NVIDIA is seeking an ML and Agentic Systems Engineer (Finance) to help build the meta-layer of modern ML: agents, tooling, pipelines, and feedback loops that accelerate model development. You will own large Python/PyTorch codebases and design end-to-end agentic workflows...Suggested
- NVIDIA Gruppe is seeking a Senior Engineer in Santa Clara, CA, to join the Cosmos team. This role focuses on creating AI-native systems that enhance the efficiency of machine learning workflows. Candidates should have extensive Python and PyTorch experience, along with...Suggested
$174.9k - $261.3k
...and understand the world!The Data Labeling Engineering team designs, builds, and operates hybrid... ...engineering, data engineering, and AI/ML, defining the strategies, tooling, and quality... ...leadership, and direct impact on systems that unblock the next generation of AV capabilities...SuggestedFull timeLocal areaRemote workWork from homeRelocation packageFlexible hours$90.1k - $191.8k
...and understand the world!The Data Labeling Engineering team designs, builds, and operates high‑... ...engineering, data engineering, and ML, defining labeling strategies, tooling, and... ...technical leadership, and work directly on systems that unblock the next generation of AV models...Full timeWork experience placementLocal areaRemote workWork from homeRelocation packageFlexible hours- NVIDIA is seeking a talented engineer in Santa Clara, California, to enhance augmented and virtual environments using advanced techniques... ...cross-platform development and work on innovative simulation systems. With over 8 years of experience required, candidates must excel...
$142.2k - $213.2k
...Technologies, Inc.Job AreaEngineering Group, Engineering Group Multimedia SystemsGeneral... ...The successful candidate will work with systems, software, and integration/test engineers... ...for single and multi-sensor fusion using ML/AI techniques. Typical development flow uses...Work experience placementWork from home$181.1k - $318.4k
...technology company in Cupertino is seeking a Machine Learning Engineer to build infrastructure for product-focused machine learning projects... .... The ideal candidate will have a strong background in backend systems development and solid knowledge of machine learning...$204k - $259k
...autonomous driving technology company is looking for an experienced engineer to improve compute performance in machine learning systems. This hybrid role involves collaboration with a world-class ML team and requires strong expertise in ML software or systems. The ideal...- Rhoda is building the next generation of generalist robotic systems in Mountain View, CA. We are seeking a senior or staff-level Research Engineer or ML Systems Engineer to make the robot-learning pipeline reliable and measurable from end to end. You will own the supported...
- Rhoda AI is hiring a Senior/Staff-level Research Engineer to ensure our robot-learning pipeline is reliable from data collection through... ..., and real-robot evaluation. You will build validation systems, observability, and robust operating practices to distinguish model...
- General Motors’ Data Labeling Engineering team is building cutting‑edge labeling tools and pipelines that power autonomous vehicle ML models. The role sits at the intersection of software... ...and ML, focusing on scalable labeling systems and foundations for foundation‑model...
- Apple seeks a senior engineer to design the core auction system powering its ads platform at global scale. You will drive optimal outcomes for advertisers, users, and the platform, applying advanced auction theory and machine learning in production environments. You will...
$181.1k - $318.4k
...lives? We truly believe it can. We are the System Intelligent and Machine Learning (SIML) group... ...learning industry pro, or an outstanding engineer wanting to expand your horizons and learn more about the nitty gritty of how ML projects are launched from inception to release...Relocation- Rhoda AI in Mountain View is seeking a Staff / Principal ML Training Systems Engineer to lead the performance of large-scale multimodal training systems. This role involves improving training efficiency and collaborating closely with research teams to accelerate model...
- Apple in Cupertino is seeking a robotics engineer to design and implement AI/ML systems for real-world products. You will work with a team of engineers and scientists to bring new experiences to Apple hardware and software. The role requires a strong foundation in robotics...
- ServiceNow is seeking a Senior Staff engineer to shape the next generation of agentic systems. You will own cross-cutting architecture, tool boundaries, and streaming protocols while collaborating with teams across Java/Spring Boot and React/TypeScript frontend. Strong...Remote job
- ServiceNow is seeking a Senior/Staff-level engineer to shape the AI agent platform. You’ll handle multi-agent orchestration, tool boundaries, and prompt context architecture while reading Spring Boot services and unblocking frontend integrations. The role spans Python...
$169k - $338k
...Segment: Home OfficePosition Summary...As a Distinguished AI/ML Engineer within Walmart Global Tech's Site Reliability Engineering organization... ...lead the technical development of next-generation agentic AI systems and intelligent automation solutions that ensure mission-...Full timeTemporary workPart time- Apple Inc. in Cupertino, CA is seeking a senior machine learning and AI expert to design and optimize the core auction system powering our advertising platform. You will deliver high-quality production code and experiments to improve marketplace value for advertisers,...
$174k - $253k
...agents and LLM-powered journeys that evaluate, gate, and improve engineering artifacts.Engineer and refine skills and context provided to... ...information retrieval, distributed computing, large-scale system design, networking and data storage, security, artificial intelligence...$184k - $287.5k
...building intelligent video analytics and perception solutions for smart cities, industrial automation, and autonomous systems at scale. We seek a Senior ML Engineer to compose and deliver next-generation Metropolis solutions powered by Cosmos, NVIDIA's world foundation model...Full time$165.2k - $223.6k
...team is actively seeking skilled compiler engineers to join our efforts in developing a state... ...forefront of AWS innovation for advanced ML capabilities, powering solutions like... ...in compiler technology and deep-learning systems software. Additionally, you'll collaborate...InternshipLocal areaFlexible hours$165.2k - $223.6k
...forefront of maximizing performance for AWS's custom ML accelerators. Working at the hardware-software boundary, our engineers craft high-performance kernels for ML functions... ...of software, hardware, and machine learning systems, you'll bring expertise in low-level...InternshipLocal areaWork from homeFlexible hours- Rhoda AI is seeking a Staff / Principal ML Training Systems Engineer in Mountain View to enhance the training systems performance. This role focuses on large-scale multimodal training, driving efficiency, and scalability in compute use across thousands of GPUs. The ideal...
- A leading technology company in Sunnyvale, CA is seeking a Software Engineer III to develop next-generation technologies related to AI and ML. This role requires a Bachelor's degree and experience in programming, particularly with Python or C++, along with expertise in...
- ...An innovative AI startup is seeking an experienced Machine Learning Engineer to design and deploy production-grade ML systems. This role involves using AI to enhance sales data insights through real-time audio understanding. Candidates should have a relevant degree and...
- A leading technology company in Cupertino is seeking a Senior ML Software Engineer to develop innovative machine learning features for Apple Watch. You will work on optimizing ML algorithms using multimodal data and collaborate with cross-functional teams to enhance user...
- General Motors is seeking a Staff ML Systems Engineer to help our Data Labeling Engineering team advance self-driving vehicle capabilities. You will own end-to-end platform projects, define roadmaps, and work across ML, data science, and product teams to deliver scalable...Remote job
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to ML Systems Engineer. Be the first to apply!
- machine learning ai engineer San Jose, CA
- computer vision machine learning engineer San Jose, CA
- machine learning engineer San Jose, CA
- ai ml engineer San Jose, CA
- machine learning software engineer San Jose, CA
- senior ml engineer San Jose, CA
- system engineer remote San Jose, CA
- senior windows systems engineer San Jose, CA
- senior linux systems engineer San Jose, CA
- ground systems engineer San Jose, CA

