Staff Machine Learning Infrastructure Engineer
$224k - $280kAtoms
Who we are Atoms is building the machines that power the next era of progress. Over the last decade, software has transformed the digital world. But the physical world, where food is made, minerals are mined, goods are moved, and industries are run, remains far less intelligent, far less efficient, and far more constrained. We’re changing that. Atoms builds Physical AI – real‑world robots for the industries that move civilization forward, starting with food, mining, and transport. Our systems are designed to understand, predict, and control the real world with precision, turning complex physical operations into something more reliable, more scalable, and more productive. This work requires more than robotics. It requires deep integration across hardware, software, AI, operations, manufacturing, and real estate. We don’t just build machines in a lab. We deploy them into real environments, operate them, learn from them, and improve them until they work at scale. We are roboticists, engineers, operators, and builders. We believe the next great technology companies will not only transform information, but the physical systems that shape everyday life. If you want to work on hard problems with real‑world impact, join us. What you’ll do We are seeking a foundational Machine Learning Infrastructure Engineer to design and build the large‑scale ML training infrastructure that powers our next‑generation autonomous transport models. In this role, you will design the high‑performance training pipelines and validation environments that enable our world‑class robotics and ML researchers to iterate rapidly. You will own the challenge of scaling distributed GPU workloads to support a high volume of concurrent training runs across an expanding vehicle fleet, building a platform that can flexibly run on whatever GPU capacity is available, regardless of provider or environment, directly accelerating innovation across the platform. Training Infrastructure: Design, implement, and scale repeatable machine learning infrastructure utilizing Kubernetes to support large‑scale distributed GPU training of novel neural networks. Distributed Computing & Orchestration: Leverage distributed compute frameworks to efficiently manage and execute a high volume of complex ML training jobs concurrently across large GPU clusters. Experiment Tracking & MLOps: Integrate advanced model management and experiment tracking tools to provide researchers with deep observability into training metrics and run performance. Data Engineering Pipelines: Build and optimize high‑throughput data ingestion pipelines to seamlessly stream petabyte‑scale multi‑sensor vehicle logs into training environments. Validation at Scale: Architect robust infrastructure for autonomous model validation and continuous integration testing, ensuring new vehicle policy releases are entirely regression‑free. Cross‑Functional Collaboration: Partner closely with core robotics engineers and machine learning researchers to eliminate workflow bottlenecks and accelerate the deploy‑to‑vehicle lifecycle. What we’re looking for 8+ years of professional software engineering career experience Strong backend systems programming skills with proficiency in Go, Python, Java or similar (with familiarity or exposure to Rust considered a plus). Proficiency with Kubernetes for container orchestration and building cloud‑agnostic environments from scratch. Experience implementing distributed ML compute frameworks (e.g., Ray) to coordinate large pools of GPUs for heavy, multi‑node workloads. Hands‑on experience building MLOps pipelines, metadata tracking architectures, and model registries using platforms like MLflow. Prior experience managing high‑throughput data pipelines using modern distributed data engines to feed data‑hungry neural network architectures. Why join us At Atoms, you’ll work on one of the defining challenges of our time – bringing automation into the physical world to drive real, lasting impact. We exist to uncover valuable unknown truths and turn them into progress, which means constantly pushing beyond what’s known and building what doesn’t yet exist. The work is ambitious and often challenging, but it’s grounded in a shared sense of purpose and a team committed to seeing it through together. Our work only matters if it serves others, and we know that meaningful progress depends on the trust of the people we serve and the strength of our team – so we invest in both, creating an environment where you can do your best work and grow. What else you need to know This role is based in our San Francisco office. Atoms is a company driven by invention and continuous change – we are constantly reimagining our industries, building new products, and refining how we operate. We do our best work together. That’s why all of our office‑based teams work onsite, five days a week. Base salary range for this role is $224,000 - $280,000 per year. Actual compensation will be determined on an individual basis and may vary depending on experience, skills, and qualifications. Base salary is just one part of your total rewards package. You may also be eligible for equity awards and an annual performance‑based bonus. Benefits Summary (USA Full‑Time Exempt Employees) Medical, Dental, Vision, Disability, and Life Insurance Flexible Spending Account / Health Savings Account Options 401(k) Equity Sick Time, Unlimited Flexible Time Off, and Paid Holidays Paid Parental Leave Pre‑Tax Commuter Benefit Plan Team lunch in our SoMa office every Tuesday and Thursday Benefits are subject to change at the company’s discretion. Atoms accepts applications on an ongoing basis. Ready to join us as we serve those who serve others?
- LI-Onsite
- J-18808-Ljbffr Atoms
- ...looking for people that have done genuinely amazing work in infrastructure that are interested in a challenge, working with both traditional... ...., as well as very different infrastructure around inference engines and GPU loads. This is a role that will inherently require...Suggested
$183.7k - $248.6k
The opportunity Unity is looking for a Senior Machine Learning Infrastructure Engineer to join our Vector Ads team, where we build the real‑time systems that power Unity's global advertising platform. This is a high‑scale, low‑latency environment — processing billions...SuggestedWork at officeRemote workWorldwideRelocation package$213k - $263k
...simulation across 15+ U.S. states. The Simulation ML Infrastructure team builds scalable AI/ML infrastructure to... ...systems, and weather. We seek an experienced Senior Machine Learning Infrastructure Engineer to lead the development of advanced AI/ML infrastructure...SuggestedFull timeRemote work$224k - $280k
Staff Machine Learning Infrastructure Engineer Who we are Atoms is building the machines that power the next era of progress. Over the last decade, software has transformed the digital world. But the physical world, where food is made, minerals are mined, goods are moved...SuggestedFull timeContract workFor contractorsFor subcontractorWork at officeFlexible hours$166k - $225k
...4 Summary Databricks Mosaic AI is hiring experienced machine learning platform engineers to build out our customer-facing generative AI platform... ...implementation Design and build the core platform infrastructure that supports our customer-facing product features Ensure...SuggestedLocal area$208k - $263.5k
...Atoms is building the machines that power the next era of progress. Over the last decade, software... ...into real environments, operate them, learn from them, and improve them until they work at scale. We are roboticists, engineers, operators, and builders. We believe the...Full timeInternshipWork at officeFlexible hours$200k - $400k
...top interpretability researchers and engineers from organizations like OpenAI and DeepMind... .... About the role We’re looking for Machine Learning Engineers to help build our platform... ...Interpretability tools – Building the tools and infrastructure to support dissection and design of...- ...and b.) everything about the fruit they are seeing. We are looking for a Senior Machine Learning Engineer to build creative, practical, and robust solutions to ML/CV software and infrastructure problems, relating to training edge ML models on massive amounts of real-world...Full timeWork at officeFlexible hoursWeekend work
$159.18k - $295.62k
...Senior Machine Learning Engineer – Data & Audience Platform (DAP) Senior, high‑ownership US‑based role... ...that sits between our Senior MLE and Staff MLE levels. The Engineer owns design... ...builds in Data Clean Rooms. MLOps & Infrastructure Champion MLOps best practices: model...Temporary work- ...firm). We've raised over 100M from world class investors like A16z, Benchmark, and First Round Capital, and are hiring a Machine Learning Engineer to help us train and deploy the models critical to the performance of our core product. We would love to meet you if you:...Work at officeLocal area
$160k - $235k
...to deliver actionable insights to customers. As a Senior Machine Learning Engineer , you will collaborate with data engineers, software engineers... ...systems: Architect and launch ranking and recommendation infrastructure from scratch, initially via integrated off-the-shelf...Work at officeRemote workWorldwideFlexible hours2 days per week3 days per week- ...build a legacy that inspires future generations. Senior Machine Learning Engineer Primary: Bay Area (San Francisco / Peninsula) | Secondary... ...and ship ML solutions. Drive architecture decisions for ML infrastructure and platform capabilities, and cut deployment cycle time....Work at officeRelocation packageFlexible hours3 days per week
$232k - $313k
...in late 2020 by a small group of machine learning researchers, Mosaic AI enables companies... ...Databricks is hiring a Senior Staff Machine Learning Engineer to play a key role in building our... ...and implementation of our ML infrastructure and GenAI platform software technologies...Local area- ...Mosaic AI, now part of Databricks, is hiring experienced machine learning platform engineers to build out our generative AI platform for the ML... ...of these designs. Design and build the core platform infrastructure that supports our customer‑facing product features. Ensure...Full time
$295k - $405.5k
About this role As the Senior Staff Machine Learning Platform Engineer, you will own the technical vision and evolution of Faire’s ML platform. You... ...workloads Strong background in distributed systems, ML infrastructure, and cloud architecture. Demonstrated technical...Work experience placementWork at officeLocal areaRemote workMonday to FridayFlexible hours3 days per week- Stripe is seeking a Staff Engineer for its ML Platform team in San Francisco, California. You will lead the architectural design and technical... ...qualifications include experience with large-scale ML infrastructure and familiarity with AI tools and cloud services. #J-18808-...
- Atoms, based in San Francisco, is seeking a Machine Learning Infrastructure Engineer to design the ML training infrastructure for autonomous transport models. The role involves scaling distributed GPU workloads, leading to innovations across their platform. Ideal candidates...
- Mach9 is looking for an experienced ML infrastructure engineer in San Francisco to build and maintain systems powering production AI models. This mid-career role involves designing centralized versioning systems, optimizing real-time inference services, and ensuring reliable...
- Gridware in San Francisco is looking for a Senior ML Infrastructure Engineer to design, build, and maintain their ML deployment infrastructure. This role will involve creating monitoring systems and collaborating with engineering teams to improve CI/CD pipelines. The ideal...
- We are looking to recruit an exceptional Infrastructure Engineer to own and build the backend systems that power machine learning at Maven Robotics. In this role, you will design and scale the core infrastructure used by our AI and robotics teams to manage data, run compute...
- ...innovation through advanced hardware engineering and AI solutions. Our mission is to... ...lasting impact. We emphasize continuous learning and growth, fostering cross-... ...Job Summary We are seeking a Senior Machine Learning Infrastructure Engineer to join our team. The person...Flexible hours
$149k - $186k
...looking for brilliant software engineers who have a passion for... ...think about car ownership. Learn more about our Engineering team... ...will do: Build a world-class machine learning platform that will... ...systems, and/or machine learning infrastructure, with an understanding of...Full timeWork at officeLocal area$295k - $380k
OpenAI is seeking a Systems Engineer to work on ML training infrastructure in San Francisco. This role focuses on creating reliable systems for large-scale training, enabling new model approaches while maintaining performance and debuggability. The ideal candidate will...- We’re looking for an experienced HPC infrastructure engineer to lead bringup, administration, and operations on is probably the largest anime... ...serve as the bridge between our researchers and the bare GPU machines, helping to make sure that SLURM jobs are running, parallel...Work at officeVisa sponsorship
- Chalk, based in San Francisco, is seeking an exceptional software engineer to partner with machine learning teams. In this role, you will develop bespoke solutions, build efficient feature pipelines, and collaborate closely with customers in healthcare and finance. The...
$224k - $280k
ATOMS Careers page is seeking a Staff Machine Learning Infrastructure Engineer to design and implement ML training infrastructure in San Francisco. The role involves building high-performance training pipelines and managing distributed GPU workloads. Ideal candidates should...$190k - $260k
Role Description As a Senior ML Infrastructure Engineer, you will work directly in the Automation org with the core ML, Ops, and Analytics teams... ...a year Senior ML Engineer Base Salary- $190,000-$210,000 Staff ML Engineer Base Salary- $245,000-$260,000. #J-18808-Ljbffr...- Baseten is looking for a Software Engineer to join their Training Infrastructure team in San Francisco. You'll architect and lead the development of the training platform, enabling developers to deploy and monitor workloads efficiently. The ideal candidate has a Bachelor...Flexible hours
$275k
Series C Startup | AI-Powered 3D & Avatar Platform | Hybrid (LA or SF) We’re hiring a Senior ML Infrastructure / Backend Engineer to join a well-funded AI company building the visual and interaction layer for the next generation of AI-powered digital identities. This team...Work at office- ...edge compute resources. Responsibility The AI Infrastructure team at Zensors builds the engine that powers our visual sensing platform. We provide... ...across thousands of video streams. As a Machine Learning Engineer in ML Runtime & Optimization , you will develop...Work at office
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Staff Machine Learning Infrastructure Engineer. Be the first to apply!
- assistant chief engineer San Francisco, CA
- technology administrator San Francisco, CA
- project engineer assistant project manager San Francisco, CA
- staff security engineer San Francisco, CA
- engineering aide San Francisco, CA
- senior staff systems engineer San Francisco, CA
- assistant electrical engineer San Francisco, CA
- staff design engineer San Francisco, CA
- research assistant engineering San Francisco, CA
- staff data engineer San Francisco, CA

