Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Software Engineer: ML Optimization

Full-time

The Generalist

About the Role

We internally call this team MBMB (More Big More Better). You will own optimizations on both the training and on-robot inference stacks. We are still in a regime of step-function, not incremental, gains.

You’ll be responsible for:

  • Making GPUs go brrrrr

  • Implementing ML, hardware, and software changes that lead to step-function gains

  • Optimizing both the inference and training stacks

You might thrive in this role if you:

  • Are proficient and stay current with the latest ML techniques for training and inference optimizations in transformer and diffusion based architectures

  • Will chase ML optimizations anywhere: From the CUDA kernels, to ML architecture, to frontend or backend network bottlenecks, CPU bottlenecks, NVLink and comms, to torch, numpy, and Python inefficiencies.


About Generalist

At Generalist, we are on a mission to make general-purpose robots a reality. We believe the industries and homes of the future will depend on humans and machines working together in new ways. Robots can help us build more and get more done.

We build embodied foundation models, starting with a focus on dexterity. This requires advancing the frontiers of data, models, and hardware, to enable robots to intelligently interact with the physical world.

The company embraces both large-scale AI and robotics as core to its DNA. Our team of researchers, roboticists, and company builders come from OpenAI, Boston Dynamics, Google DeepMind, and other frontier labs—with a track record of shipping AI breakthroughs. Before Generalist, we pioneered large embodied multimodal models and vision-language-action models (PaLM-E, RT-2 , Gemini Robotics ), launched and scaled ChatGPT and GPT-4 to hundreds of millions of users, engineered the foundations of autonomous driving, built next-generation robots ( Atlas , Spot , Stretch ) and pushed the limits of what they can do (from parkour to manipulation , and testing robustness ).

We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic.

Vacancy posted 8 hours ago
Similar jobs that could be interesting for youBased on the Software Engineer: ML Optimization in San Francisco, CA vacancy
  •  ...Google Workspace.  What you'll do As a Software Engineer on our Site Reliability team at Sierra,...  ...and collaborating across product, ML, and core engineering teams. ~ Degree...  ...Experience with LLM infrastructure — optimizing inference performance, managing fine-tuned... 
    Suggested
    Full time
    Flexible hours

    Sierra

    San Francisco, CA
    24 days ago
  •  ...library that empowers all our engineers and designers. By delivering...  ...accessibility, and code quality. You’ll optimize our frontend for speed (...  ...who share our passion for software craftsmanship and getting...  ...experience : Exposure to AI or ML products, especially building... 
    Suggested
    Full time
    Flexible hours

    Sierra

    San Francisco, CA
    24 days ago
  • $229.9k - $262.4k

     ...Sr. Lead AI Engineer (Inference Optimization, FM hosting, AI Platform) Overview: At Capital One, we are...  ...real time, our applications of AI & ML are bringing humanity and simplicity...  ...develop, test, deploy, and support AI software components including foundation model... 
    Suggested
    Full time
    Part time
    Local area

    Capital One

    San Francisco, CA
    2 days ago
  •  ...with hands-on support from AMD engineers the team is scaling rapidly to...  ...learn how large AI models are optimized and deployed at scale, and collaborate closely with ML researchers and experienced systems...  ...experience) ~3+ years of software engineering experience, with a... 
    Suggested
    Full time
    Work at office
    Flexible hours

    Sciforium

    San Francisco, CA
    8 hours ago
  • $202k - $237k

     ...computing and make it accessible to software developers of all skill levels...  ...data scientist can scale an ML application from their laptop...  ...seeking a Backend Software Engineer to join our team focused on building...  ...enhance user experience and optimize workflows. Compensation... 
    Suggested
    Full time
    Work at office
    Flexible hours

    Anyscale

    San Francisco, CA
    8 hours ago
  • $300k - $320k

     ...group of committed researchers, engineers, policy experts, and business...  ...is looking for backend software engineers to work across our...  ...inference and safeguards to optimize the full stack. API Capabilities...  ..., SSO, RBAC) Exposure to ML/AI systems or an understanding... 
    Full time
    Work at office
    Visa sponsorship
    Flexible hours

    Anthropic

    San Francisco, CA
    8 hours ago
  •  ...Minerals Mariana Minerals is a software-first, vertically integrated...  ...Senior Full Stack Software Engineer to lead critical technical initiatives...  ...collaboration with ML engineers to integrate LLMs for...  ...reliability and performance optimization for critical systems. What... 
    Full time

    Mariana Minerals

    San Francisco, CA
    8 hours ago
  • $180k - $220k

     ...the AI era. We’re looking for a backend engineer to join our small team and help lead the...  ..., performance, and innovation around optimal team communication. Who you are An...  ...GraphQL APIs, with emphasis on integrating AI/ML inference endpoints and ensuring... 
    Full time
    Work at office
    Local area
    Immediate start
    Flexible hours

    Glu Mobile Inc.

    San Francisco, CA
    3 hours ago
  • $170k - $195k

     ...looking for a Senior Backend Engineer to join Tatari. As a Senior Engineer...  ...-relational data models for optimal storage, retrieval, and...  ...Qualifications ~6+ years developing software with object-oriented and/or...  ...) ~ Experience developing AI/ML-enabled features ~... 
    Full time
    Work at office
    Work from home
    2 days per week

    Tatari

    San Francisco, CA
    3 hours ago
  • $126k

     ...About the team: The ML Content Understanding team powers...  ...intersection of machine learning, data engineering, and distributed systems,...  ...Overview: We’re seeking a Software Engineer II with strong...  ...role, you’ll design, build, and optimize distributed systems that... 
    Full time
    Local area
    Worldwide
    Home office
    Flexible hours

    Scribd, Inc.

    San Francisco, CA
    8 hours ago
  •  ..., will be as an Infrastructure / Backend Engineer on our founding team. You're the right...  ...powered agents in production Deploying and optimizing AI-heavy services with high availability...  ...Experience deploying and productionizing ML or AI-heavy workloads Expertise in... 
    Full time

    Vooma

    San Francisco, CA
    8 hours ago
  • $136k - $160.5k

     ...About the Role: We are seeking a Software Engineer, Control Plane to help build and scale...  ...systems that manage our global fleet of AI-optimized compute, network, and storage resources...  ...differentiated cloud solutions for AI/ML customers. Operational Excellence:... 
    Full time
    Temporary work
    Immediate start

    Crusoe

    San Francisco, CA
    8 hours ago
  •  ...work sits at the intersection of production ML systems, developer platforms, model...  ...the Role We’re looking for a backend engineer who can quickly understand OpenAI’s models...  ...runtime abstractions for a stateful, cloud-optimized agentic platform. Partner closely with... 
    Full time
    Internship

    OpenAI

    San Francisco, CA
    8 hours ago
  •  ...reliable, and unblocked. You will work across engineering and infrastructure problems as they...  ...with experience in some layer of ML infrastructure. - Have worked on RL, inference...  ...stacks. - Background in performance optimization, scaling, or production-critical infrastructure... 
    Full time

    OpenAI

    San Francisco, CA
    8 hours ago
  • $177k - $245k

     ...there were no similarly great tools to help ML practitioners build better models....  ...practitioners and our business. As a Senior Software Engineer, you will lead initiatives in pricing,...  ..., integrating with billing vendors for optimal user experience. Collaborate with... 
    Remote job
    Full time
    Temporary work
    Work at office
    Home office
    Flexible hours

    Weights & Biases

    San Francisco, CA
    8 hours ago
  • $152.5k - $287.5k

     ...the first time, all from batteries we already have.   Software Engineer, ML/Computer Vision (Battery Sorting) The Battery Sorting team...  ...as MLflow ~ Familiarity with edge deployment or model optimization techniques for inference (e.g., quantization, TensorRT, ONNX... 
    Hourly pay
    Full time
    Immediate start
    Shift work

    Redwood Materials

    San Francisco, CA
    8 hours ago
  •  ...builds foundational components that power OpenAI’s ML training infrastructure. We focus on developing...  ...development at scale. About the Role As a software engineer on the Scaling team, you’ll help build and optimize the low-level stack that orchestrates computation... 
    Full time
    Work at office
    Local area
    Relocation package
    3 days per week

    OpenAI

    San Francisco, CA
    8 hours ago
  •  ...We're looking for a strong generalist software engineer to help us ship the next version of our...  ...bottleneck Collaborate closely with ML, hardware, and clinical teams to ship end...  ...them ~ Pragmatic instincts about when to optimize, when to ship, and when to rewrite ~ Strong... 
    Permanent employment
    Full time
    Remote work

    Remedy Robotics

    San Francisco, CA
    8 hours ago
  •  ...cost are spent, then turn that understanding into performance optimizations and models that project performance and capacity needs for...  ...benchmarking, analysis, and optimization. Enjoy collaborating with engineering and research teams to improve real production systems.... 
    Full time

    OpenAI

    San Francisco, CA
    8 hours ago
  •  ...responsible for the architectural and engineering backbone of OpenAI’s...  ...models. Our work spans system software, networking, platform...  ...monitoring, and performance optimization. About the Role We’re hiring...  ...~5+ years in one or more of: ML systems, performance engineering... 
    Full time

    OpenAI

    San Francisco, CA
    8 hours ago
  •  ...are easy for researchers to use and maximally utilized Optimizing and improving ML data loading transport and storage in highly distributed...  ...scaled ChatGPT and GPT-4 to hundreds of millions of users, engineered the foundations of autonomous driving, built next-... 
    Full time

    The Generalist

    San Francisco, CA
    8 hours ago
  •  ...the boundaries of data, scaling laws, optimization techniques, model architectures, and efficiency...  ...About the Role We’re looking for a Software Engineer focused on building and scaling...  ...Familiarity with embedding-based or ML-powered systems. Experience with performance... 
    Full time

    OpenAI

    San Francisco, CA
    8 hours ago
  •  ...Conviction. Join us and help build the platform engineers turn to to ship AI products. THE...  ...that powers modern AI workloads, optimizing every microsecond of computation to enable...  ...implement high-performance GPU kernels for key ML operations, including matrix... 
    Full time
    Flexible hours

    Baseten

    San Francisco, CA
    8 hours ago
  • $320k

     ...group of committed researchers, engineers, policy experts, and business...  ...and unattended. As a Software Engineer on the Launch Engineering...  ...is a resource-constrained optimization problem at its core: validation...  ...Have Experience with ML inference or training... 
    Full time
    Work at office
    Visa sponsorship
    Flexible hours
    Shift work

    Anthropic

    San Francisco, CA
    8 hours ago
  •  ...or as self-hosted and on-premises software, with unmatched accuracy, low...  ...Opportunity: We are seeking a Software Engineer to join Deepgram for Restaurants...  ...of our backend with our ML models and client devices Monitor and optimize the performance of our backend systems... 
    Full time
    Home office
    Flexible hours

    Deepgram

    San Francisco, CA
    8 hours ago
  • $325k

     ...group of committed researchers, engineers, policy experts, and business...  ...for reliability-minded software engineers and SREs. Are curious...  ...experience with one or more ML hardware accelerators (GPUs,...  ...Understand ML-specific networking optimizations like RDMA and InfiniBand.... 
    Full time
    Work at office
    Visa sponsorship
    Flexible hours

    Anthropic

    San Francisco, CA
    8 hours ago
  • $220k - $247.5k

     ...games. We are looking for a Senior Machine Learning Engineer to join our Revenue ML team at Discord. This role sits at the intersection of Discord...  ...targeting, audience segmentation, and campaign optimization. Familiarity with loyalty, rewards, or incentive pro... 
    Full time
    Seasonal work

    Discord

    San Francisco, CA
    8 hours ago
  •  ...Tech Lead to drive the design, optimization, and scaling of our inference...  .... In this role, you’ll lead engineering efforts to ensure our largest...  ...layer of large-scale ML infrastructure. Understand the...  ...performance issues across hardware and software layers. Have strong... 
    Full time

    OpenAI

    San Francisco, CA
    8 hours ago
  • $115k - $140k

     ...is the critical hardware & software necessary to augment any off-...  ...the Job As a   Software Engineer - Perception   at Lodestar, you...  ...deployment Design, develop ML algorithms from the ground up...  ...in addition to build on and optimizing our existing models for edge... 
    Permanent employment
    Full time
    Flexible hours

    Lodestar Corporation

    San Francisco, CA
    8 hours ago
  •  ...Conviction. Join us and help build the platform engineers turn to to ship AI products. THE...  ..., runtime tuning, and server-level optimizations. Build large-scale, real-time infrastructure...  ...developer-oriented tools Interest in ML/AI infrastructure and willingness to... 
    Full time
    Flexible hours

    Baseten

    San Francisco, CA
    8 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Software Engineer: ML Optimization. Be the first to apply!