Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Staff AI Infrastructure Engineer Orchestration & Inference

Hamilton Barnes Associates Limited

Hamilton Barnes Associates Limited is seeking a Staff-level engineer to architect and evolve the AI infrastructure software stack for large-scale GPU workloads. You’ll drive orchestration, scheduling, and agentic operations across bare metal, Slurm, Kubernetes, and InfiniBand in a live production environment. You’ll contribute to inference platforms, model serving, and high-utilization microservices, while building IaC tooling to automate deployment across thousands of servers. #J-18808-Ljbffr Hamilton Barnes Associates Limited

Vacancy posted 10 hours ago
Similar jobs that could be interesting for youBased on the Staff AI Infrastructure Engineer Orchestration & Inference in San Francisco, CA vacancy
  • $200k - $230k

     ...DescriptionDirector, AI Platform EngineeringLocations...  ..., and advanced orchestration on top of our...  ...Director of AI Platform Engineering to lead the design, development...  ..., architect critical infrastructure, and drive the...  ...(new LLM providers, inference optimization, agentic... 
    Suggested
    Ongoing contract
    Full time
    Casual work
    Work at office
    Flexible hours

    SS&C Technologies

    San Francisco, CA
    2 days ago
  • Perplexity is building a self-serve compute platform that lets inference engineers run training jobs and inference services without worrying...  ...across workloads. You’ll manage the Kubernetes for GPU orchestration, multi-cloud GPU fleets, and capacity planning, aiming for... 
    Suggested

    Neura Market

    San Francisco, CA
    3 days ago
  • $156.06k - $211.14k

    Afresh, the AI platform for grocery, began by tackling the...  ...grocery market. Our platform now orchestrates billions of decisions from...  ...not.As a Senior AI Platform Engineer, you build the AI and data...  ...the evaluation and serving infrastructure underneath. Your "customers"... 
    Suggested
    Full time
    Live in
    Work at office
    Local area
    Remote work
    Work from home
    Home office
    Flexible hours
    3 days per week

    Afresh

    San Francisco, CA
    20 hours ago
  • $142.2k - $204.6k

     ...About This RoleAs a software engineer for GenAI inference, you will help design,...  ...from kernels and runtimes to orchestration and memory management.What...  ..., distributed inference infrastructure - orchestrate across nodes...  ...DatabricksDatabricks is the data and AI company. More than 10,000... 
    Suggested
    Local area
    Worldwide

    DataBricks

    San Francisco, CA
    20 hours ago
  • $250k - $300k

     ...intelligence. As the only vertically integrated AI infrastructure company built from the ground up, we...  ...in production. That means owning the inference stack end to end: profiling where time...  ...will also work directly with customer engineering teams to tailor deployments to their... 
    Suggested
    Temporary work

    Crusoe

    San Francisco, CA
    3 days ago
  • A leading cloud infrastructure company is seeking a Senior Engineer 2 to join their AI Inference Optimization team. The role involves leading the technical strategy for performance architecture and addressing complex performance issues ensuring industry-leading service... 
    Remote work

    DigitalOcean

    San Francisco, CA
    1 day ago
  •  ...Workato delivers enterprise infrastructure for the agentic era,...  ...applications, processes, and AI into a single, governed platform...  ...to power real‑time orchestration at scale. With enterprise-...  ...looking for an exceptional Staff Software Engineer to help define and build the... 
    Immediate start
    Remote work
    Flexible hours

    Workato

    San Francisco, CA
    2 days ago
  • $269.1k - $307.2k

    Distinguished AI Engineer (Agentic AI Platform)...  ...investments in technology infrastructure and world-class talent...  ...APIs that cover agent orchestration, RAG pipelines, prompt...  ...hours, mentoring Staff, Principal and Senior...  ...technologies (e.g. LLM Inference, Similarity Search and... 
    Full time
    Part time
    Work at office
    Local area

    Capital One Financial Corporation

    San Francisco, CA
    20 hours ago
  • $175k - $300k

     ...intersection of platform engineering, site reliability,...  ...of Meshy's AI model serving stack,...  ...with core engineering infrastructure. The team operates a...  ...capabilities for the AI inference platform, including key...  ...scheduling, service orchestration, elastic scaling, and... 
    Work at office
    Remote work
    Flexible hours

    Meshy

    San Francisco, CA
    4 days ago
  • $151.8k - $265.35k

     ...content effortlessly. The AI for Engineering team builds a scalable, production...  ....We are seeking a Staff Engineer - AI for Engineering...  ...dynamics, failure modes, and orchestration strategies — and can...  ...integration, memory systems, inference services, data flows, evaluation... 
    Full time
    Temporary work
    Local area
    Worldwide

    Adobe Systems

    San Francisco, CA
    20 hours ago
  • $220k

    Perplexity is looking for an engineer to join their team in San Francisco. You will work on building and operating the inference engine, supporting new models, migrating GPU kernels, and developing a Rust-based serving runtime. The ideal candidate has 3+ years of experience... 

    Perplexity

    San Francisco, CA
    1 day ago
  • Infinity Artificial Intelligence Institute in San Francisco Bay Area seeks an AI Infrastructure Engineer to advance an inference stack that spans multiple chips and models. You will build optimization kernels, a scalable library generator, and a benchmark-driven pipeline... 

    Touring Capital

    San Francisco, CA
    2 days ago
  • $300k

     ...interpretable, and steerable AI systems. We want...  ...researchers, engineers, policy experts,...  ...The Cloud Inference team scales and optimizes...  ..., and make smart infrastructure decisions that...  ...or container orchestration Have strong interest...  ..., we expect all staff to be in one of... 
    Full time
    Work at office
    Visa sponsorship
    Flexible hours

    Anthropic

    San Francisco, CA
    20 hours ago
  • $180k - $225k

    As a Software Engineer on the ML Infrastructure team, you will design and build platforms...  ...with containers and orchestration tools (e.g., Docker, Kubernetes...  ...-LLM, or text-generation-inference.Compensation packages at...  ...mission is to develop reliable AI systems for the world's... 
    Full time

    Scale AI

    San Francisco, CA
    20 hours ago
  •  ...Using frontier causal inference-based econometric...  ...managers, economists, and engineers from Google, Netflix,...  ...library, a science orchestration library, and a Metaflow...  ...team that leans into AI-assisted development...  ..., A/B testing infrastructure, or causal inference... 
    Full time
    Work at office
    Work from home
    Worldwide
    Flexible hours

    Haus Analytics

    San Francisco, CA
    20 hours ago
  •  ...products.As a Senior Lead Software Engineer at JPMorgan Chase within the Corporate Sector, Infrastructure Platforms team, you are an...  ...cloud platforms optimized for AI/ML workloads.Partner with AI teams...  ...architecture, ML training, and inference.Experience with Infrastructure... 
    For contractors

    JP Morgan Chase

    San Francisco, CA
    1 day ago
  • Artificial Analysis, the leading AI benchmarking company, is hiring a Member of Technical Staff to own serverless inference coverage and drive performance benchmarks across provider...  ...serving performance. You will work with engineers and leadership to direct the roadmap of... 

    Artificial Analysis

    San Francisco, CA
    1 day ago
  • Sail builds the world’s most efficient software for inference and agent hosting. In this role, you’ll own token processing down to the lowest layers of the stack, optimize kernel performance, develop new request scheduling and parallelism strategies, and help us use a heterogeneous... 

    SAIL

    San Francisco, CA
    1 day ago
  •  ...to be the world's leading generative AI studio — we're the team behind niji・journey...  ...niji・journey , is looking for an AI Infrastructure Engineer to join us in building out end-to-end...  ...implement and run our next-generation inference architecture for running all our... 
    Work experience placement
    Work at office
    Visa sponsorship

    Spellbrush

    San Francisco, CA
    3 days ago
  •  ...We're a team of ex-Google engineers who built some of the largest defensive platforms on...  ...: stopping the new wave of adversarial AI attacks already hitting organizations today...  .... These agents will operate under an orchestration layer designed for rapid iteration and adaptive... 
    Full time
    Flexible hours

    Aegis Ai

    San Francisco, CA
    20 hours ago
  • $150k - $237.5k

     ...full stack: compute infrastructure, telemetry, asset modeling...  ...operational control orchestration, and integration with...  ...trust, learning, and engineering excellence, and we...  ...distributed systems AI-accelerated...  ...employment information, and inferences drawn from your PI. We... 
    Full time
    Flexible hours

    Redwood Materials

    San Francisco, CA
    20 hours ago
  • $197.3k - $225.1k

     ...Overview Lead AI Engineer (MLX, Agentic AI, Gen AI platform Services) At Capital...  ...experiences. Our investments in technology infrastructure and world-class talent — along with...  ...model training, large language model inference, similarity search, guardrails, model... 
    Full time
    Part time
    Local area

    Capital One

    San Francisco, CA
    5 days ago
  • $229.9k - $262.4k

    Senior Lead AI Engineer (Gen AI Platform Services) Overview: At Capital One, we are creating...  .... Our investments in technology infrastructure and world-class talent — along with our...  ...model training, large language model inference, similarity search, guardrails, model... 
    Full time
    Part time
    Local area

    Capital One

    San Francisco, CA
    3 days ago
  •  ...& Tech, Enterprise Software, AI and Machine Learning, SaaS &...  ...and process to power real-time orchestration at scale. With enterprise-...  ...RoleOur client is looking for a Staff Product Manager for Developer...  ..., business goals, and engineering teams to drive the vision and... 
    Permanent employment
    Full time
    Immediate start
    Remote work
    Flexible hours

    Chronos Consulting

    San Francisco, CA
    3 days ago
  • $150k - $200k

     ...We're a small, fast-moving AI infrastructure company building an open-source...  ...demanding AI workloads including inference pipelines, distributed training, storage orchestration, and GPU scheduling. Our...  ...cluster in days. As a Platform Engineer, you'll work shoulder-to-... 
    Full time
    Visa sponsorship

    Clera

    San Francisco, CA
    6 days ago
  •  ...Known - Conversational AI Engineer, System Prompt and Orchestration ~ San Francisco, CA (In-Person) ~170k-220k Cash + Equity Known is a matchmaker that talks to users and supports them like a friend. Our mission is to empower humanity by applying general intelligence... 
    Full time

    Known

    San Francisco, CA
    20 hours ago
  • $102k - $125k

     ...and purpose. By blending design, product engineering, analytics, and automation, we build the...  ..., you’ll design and deliver innovative AI/ML and agentic AI solutions as part of intelligent...  ...frameworks using cutting-edge orchestration frameworks, tool-use protocols, and cloud... 
    Temporary work
    Work at office
    Local area
    Flexible hours

    Slalom

    San Francisco, CA
    2 days ago
  • $170k - $216k

     ...across 15+ U.S. states. The Simulation Infrastructure team creates reliable, scalable, and...  ...for a broad range of customers Software Engineers, Product, Data Science, System...  ...You will: Build and evolve ML inference infrastructure for simulations. Be responsible... 
    Full time
    Remote work

    Waymo

    San Francisco, CA
    20 hours ago
  • $181.37k - $235.02k

     ...customers. We want 6sense to be the best chapter of your career. Staff Enterprise AI Engineer, IT OperationsAbout the RoleWe’re looking for a Staff...  ...frameworks, Microsoft 365 Copilot, and ClaudeBuild and orchestrate AI agents that automate cross-functional workflows across... 
    Full time

    6sense

    San Francisco, CA
    1 day ago
  • $161.7k - $303.3k

     ...marketing teams operate — not by layering AI tools on top of existing workflows, but...  ...a team of 6-8 Forward-Deployed AI Engineers embedded across GMI's paid media, lifecycle...  ...agentic systems, RAG, LLM evaluation, tool orchestration. Your team will bring you real problems... 
    Full time
    Temporary work
    Local area
    Worldwide

    Adobe Systems

    San Francisco, CA
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Staff AI Infrastructure Engineer Orchestration & Inference. Be the first to apply!