Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Lead AI Systems Engineer — Scale Core ML & GPU Clusters

Transluce

Transluce, a fast-moving research lab in San Francisco, is seeking an exceptional AI systems engineer to lead the design and development of our core ML stack, building scalable systems that can leverage thousands of GPUs and handle trillion-token databases. As an early member of a highly collaborative team, you will move fast, innovate from the ground up, and deliver high-impact tooling with cross-organizational reach, including open-source components that support AI evaluation and public policy #J-18808-Ljbffr Transluce

Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Lead AI Systems Engineer — Scale Core ML & GPU Clusters in San Francisco, CA vacancy
  • Linuxcareers in San Francisco is building AI research infrastructure. You will design, deploy, and operate large-scale GPU clusters powering training, evaluation, and serving for the research team. The role emphasizes extending orchestration with Kubernetes/Slurm, building... 
    Suggested

    Linuxcareers

    San Francisco, CA
    6 days ago
  • $250k

     ...Join a rapidly scaling AI cloud infrastructure...  ...a next-generation GPU platform designed...  ...Site Reliability Engineer to support and scale...  ...closely with platform, ML, and...  ...frameworks for GPU compute clusters Collaborate with...  ...available infrastructure systems Improve CI/CD... 
    Suggested
    Full time
    Remote work
    San Francisco, CA
    more than 2 months ago
  •  ...s most dynamic AI companies, like...  ...build the platform engineers turn to to ship...  ...operating system for distributed...  ...modal workloads scale, the network is...  ...foundational engineers to lead our GPU Networking...  ...bleeding-edge clusters (H100/H200, B20...  ...a variety of ML startups,... 
    Suggested
    Full time
    Flexible hours

    Baseten

    San Francisco, CA
    8 hours ago
  • Salesforce is seeking an accessibility engineer to accelerate the development of accessible AI powered products and services across the enterprise...  .... As a subject matter expert with AI/ML experience, you design prompt systems, build evaluation frameworks, and enable scalable... 
    Suggested

    Salesforce

    San Francisco, CA
    3 days ago
  •  ...progress of AI applications...  ...scientist can scale an ML application from...  ...to the cluster without needing...  ...distributed systems expert....  ...looking for engineers with systems...  ...About the Ray Core Team The Ray...  ...you will: Leading cross-team projects...  ...Knowledge of GPU programming is... 
    Suggested
    Full time
    Work experience placement

    Anyscale

    San Francisco, CA
    8 hours ago
  •  ...software development with AI-powered formal...  ...Join our team as an AI Engineer and help us push the boundaries...  ...refine the data and ML pipelines for scaled distributed training...  ...with Kubernetes clusters and distributed compute...  ...Multi-node and multi-GPU training Mathematical... 
    Full time
    Contract work

    Logical Intelligence

    San Francisco, CA
    8 hours ago
  • $240k - $280k

     ...a Software Engineer to build the systems that treat infrastructure...  ...inference clusters without a...  ...stand up, scale, or tear...  ...functioning AI cluster for...  ...bring-up to GPU driver/CUDA...  ...the inference/ML platform...  ...Requirements Core requirements...  ...to leading open-source... 
    Full time

    Together Ai

    San Francisco, CA
    8 hours ago
  • $220k

     ...Perplexity is looking for an engineer to join their team in San Francisco. You will work on...  ...engine, supporting new models, migrating GPU kernels, and developing a Rust-based serving...  ...experience in software engineering with a focus on ML inference, familiarity with deep learning... 

    Perplexity

    San Francisco, CA
    4 days ago
  • $220k - $300k

     ...so companies can scale without losing...  ...raised $204M from leading venture capital...  ...customer operations. AI is reshaping...  ...effort is our Core AI Platform team...  ...applied AI and engineering talent whose work...  ...pipelines, evaluation systems) that all...  ...technical fluency in AI/ML systems,... 
    Work at office
    Immediate start
    Remote work
    Work from home
    Monday to Friday

    FrontApp

    San Francisco, CA
    3 days ago
  • iframe.ai is hiring a Customer Cluster Engineer to own three to five reserved-capacity accounts, managing their training and inference performance end...  ...kernel tuning. You have 5+ years in distributed- or ML-systems engineering, strong PyTorch/FSDP/Megatron-LM debugging... 
    Remote job

    iFrame

    San Francisco, CA
    3 days ago
  •  ...the web by building AI agents that can reliably...  ...Responsibilities: Scale infra for post-training...  ...closely with product engineers to translate cutting‑...  ...: Experience with ML infrastructure (GPU clusters) and supporting networking...  ...latency) Low level systems experience (Triton,... 
    Work at office
    Relocation
    Visa sponsorship

    Yutori

    San Francisco, CA
    2 days ago
  • $250k

    A Series A Funded start-up in California is seeking a Systems Engineer to design and optimize systems handling complex ML pipelines. The role involves building scalable infrastructure, developing CI/CD pipelines, and ensuring system performance. Key qualifications include... 

    Acceler8 Talent

    San Francisco, CA
    4 days ago
  • $207k - $290k

     ...About JazzX AI:   Vision:...  ...enterprises don't scale expertise—...  ...bringing AI systems to market that...  ...experienced AI Engineer with deep...  ...You will lead the design, development...  ...product, core platform engineering...  ...in AI/ML engineering,...  ...(Kubernetes, GPU/TPU clusters, and cloud ML... 
    Worldwide
    Flexible hours

    JazzX AI

    San Francisco, CA
    8 days ago
  •  ...to incubate and scale revolutionary AI‑powered enterprise...  ...and intelligent systems. The result is a...  ...pacesetters who lead their industries...  ...an experienced AI Engineer with deep expertise...  ...experience in AI/ML engineering, including...  ...(Kubernetes, GPU/TPU clusters, and cloud ML platforms... 
    Flexible hours

    JazzX AI

    San Francisco, CA
    3 days ago
  • $106.9k - $176.5k

     ...world. The opportunity We are seeking AI Systems Engineers to own the security and trust fabric of...  ...and cryptographic lifecycle at production scale across multiple environments. Strong understanding...  ...OPA), and API gateways. Exposure to AI/ML workloads and the specific trust... 
    Work experience placement
    Summer holiday
    Remote work
    Flexible hours

    EY

    San Francisco, CA
    2 days ago
  • $350k

     ...tech stack for understanding and debugging AI systems. We build world-class, AI-backed...  ...are looking for an exceptional AI systems engineer to lead the design and development of our core ML stack , building systems that can scale to thousands of GPUs and performantly query... 
    Visa sponsorship
    Flexible hours

    Transluce

    San Francisco, CA
    3 days ago
  •  ...world's most dynamic AI companies, like...  ...the platform engineers turn to to ship AI...  ...building its own GPU infrastructure for large-scale inference. As we...  ...high-density NVIDIA systems, the hardest...  ...We are hiring a Lead Software Engineer...  ...to a variety of ML startups, offering... 
    Full time
    Flexible hours

    Baseten

    San Francisco, CA
    8 hours ago
  •  ...software development with AI-powered formal...  ...our team as an AI Engineer and help us push...  ...LLMs) pipelines for scaled distributed training...  ...optimizing machine learning systems, including general...  ...experience in ML Infra, DataOps,...  ...Proficiency with Kubernetes clusters and distributed... 
    Full time
    Contract work

    Logical Intelligence

    San Francisco, CA
    8 hours ago
  • $200k - $350k

     ...operations across multiple systems and platforms....  ...seeking a Senior AI Engineer to develop, deploy...  ...AI, and modern ML frameworks. Build...  ...infrastructure. GPU optimization. RAG...  ...systems. Startup or scale-up experience....  ...technology that sits at the core of critical... 
    Remote job
    Full time
    Immediate start

    Pragmatike

    San Francisco, CA
    8 hours ago
  • AI Systems Engineer - Codex Core Agents Location San Francisco Employment Type Full time...  ...virtualization, cloud platforms, or ML systems. Enjoy working...  ...strong ownership, and can lead scoped or multi-team AI...  ...runtimes, inference optimization, GPU systems, benchmarking,... 
    Full time
    Work at office
    Local area
    Relocation package
    Flexible hours

    Slope

    San Francisco, CA
    3 days ago
  • $230k - $385k

    AI Systems Engineer - Codex Core Agents About The Team The Codex Core Agents team builds...  ...across low-level systems and ML workflows, able to debug Codex...  ..., inference/runtime stack, GPU fleet, and product surface....  ...show strong ownership, and can lead scoped or multi-team AI... 

    OpenAI

    San Francisco, CA
    4 days ago
  • Innovaccer is seeking a Principal AI Engineer to build production-grade AI systems, including LLM-based solutions, AI agents, and RAG-driven workflows. You...  ...design, develop, and deploy AI-powered applications at scale. The ideal candidate will own the full AI development... 

    Innovaccer

    San Francisco, CA
    2 days ago
  • $200k - $350k

     ...operations across multiple systems and platforms....  ...seeking a Staff AI Engineer to lead the design and...  ...production AI systems at scale. What You'll Do...  ...of engineering or ML experience. ~...  ...Kubernetes and GPU infrastructure....  ...that sits at the core of critical business... 
    Remote job
    Full time
    Immediate start

    Pragmatike

    San Francisco, CA
    8 hours ago
  •  ...About the Team The Codex Core Agent team builds the kernel of...  ...That means working across the systems that make Codex actually function...  ...We’re looking for applied AI engineers to help bring Codex agents from...  ...Python and comfortable with modern ML tooling. Have worked on... 
    Full time

    OpenAI

    San Francisco, CA
    8 hours ago
  •  ...Limited is seeking an ambitious infrastructure engineer to help build and operate automated systems that bring GPU clusters from bare machines to customer-ready...  ...tenant lifecycle, run Kubernetes and Postgres at scale, and contribute to Kubernetes operators that power... 
    Remote job

    Hamilton Barnes Associates Limited

    San Francisco, CA
    2 days ago
  •  ...client is a well-funded AI startup building production-grade ML infrastructure used by...  ...for a Senior AI/ML Engineer to own model training pipelines, evaluation systems, and inference serving at scale. Full-time, on-site in...  ...distributed training, GPU optimization, or inference... 
    Full time

    Clera

    San Francisco, CA
    8 hours ago
  •  ...da Vinci surgical system and Ion—have transformed...  ....We’re a team of engineers, clinicians, and...  ...a Senior Systems GPU Engineer - AI & Robotics, you...  ...research, SW/ HW/ ML engineering, regulatory...  ...& Strategy: Lead complex, end-to-end...  ...while serving as a core contributor to team... 
    Local area
    Worldwide
    Flexible hours

    Intuitive Surgical

    San Francisco, CA
    4 days ago
  •  ...The AI Infrastructure team at Zensors builds the engine that powers our visual sensing...  ...Engineer in ML Runtime & Optimization...  ...: Optimizing Core ML Pipelines:...  ..., and operating systems to facilitate...  ...understanding of GPU hardware performance...  ...or cloud-scale inference serving... 
    Full time

    Zensors

    San Francisco, CA
    8 hours ago
  • $350k

     ...tools to make AI work for their...  ...are scientists, engineers, and builders who...  ...and systems engineers to help...  ...architecting and scaling the core infrastructure...  ...infrastructure for the clusters to reliably and...  ...clusters with GPU workloads, or building...  ...with GPU/ML workflows or... 
    Full time
    Local area
    Immediate start
    Visa sponsorship
    Work visa
    Relocation package
    Flexible hours

    Thinking Machines Lab

    San Francisco, CA
    8 hours ago
  •  ...to be the world's leading generative AI studio — we're...  ...AI Infrastructure Engineer to join us in building...  ...out end-to-end ML infrastructure to...  ...distributed systems. You have familiarity...  ...and multi-cloud clusters. But most...  ...understanding of GPU’s handling large... 
    Work experience placement
    Work at office
    Visa sponsorship

    Spellbrush

    San Francisco, CA
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Lead AI Systems Engineer — Scale Core ML & GPU Clusters. Be the first to apply!