Staff AI Infrastructure Engineer Orchestration & Inference
Hamilton Barnes Associates Limited
Hamilton Barnes Associates Limited is seeking a Staff-level engineer to architect and evolve the AI infrastructure software stack for large-scale GPU workloads. You’ll drive orchestration, scheduling, and agentic operations across bare metal, Slurm, Kubernetes, and InfiniBand in a live production environment. You’ll contribute to inference platforms, model serving, and high-utilization microservices, while building IaC tooling to automate deployment across thousands of servers. #J-18808-Ljbffr Hamilton Barnes Associates Limited
$200k - $230k
...DescriptionDirector, AI Platform EngineeringLocations... ..., and advanced orchestration on top of our... ...Director of AI Platform Engineering to lead the design, development... ..., architect critical infrastructure, and drive the... ...(new LLM providers, inference optimization, agentic...SuggestedOngoing contractFull timeCasual workWork at officeFlexible hours- Perplexity is building a self-serve compute platform that lets inference engineers run training jobs and inference services without worrying... ...across workloads. You’ll manage the Kubernetes for GPU orchestration, multi-cloud GPU fleets, and capacity planning, aiming for...Suggested
$156.06k - $211.14k
Afresh, the AI platform for grocery, began by tackling the... ...grocery market. Our platform now orchestrates billions of decisions from... ...not.As a Senior AI Platform Engineer, you build the AI and data... ...the evaluation and serving infrastructure underneath. Your "customers"...SuggestedFull timeLive inWork at officeLocal areaRemote workWork from homeHome officeFlexible hours3 days per week$142.2k - $204.6k
...About This RoleAs a software engineer for GenAI inference, you will help design,... ...from kernels and runtimes to orchestration and memory management.What... ..., distributed inference infrastructure - orchestrate across nodes... ...DatabricksDatabricks is the data and AI company. More than 10,000...SuggestedLocal areaWorldwide$250k - $300k
...intelligence. As the only vertically integrated AI infrastructure company built from the ground up, we... ...in production. That means owning the inference stack end to end: profiling where time... ...will also work directly with customer engineering teams to tailor deployments to their...SuggestedTemporary work- A leading cloud infrastructure company is seeking a Senior Engineer 2 to join their AI Inference Optimization team. The role involves leading the technical strategy for performance architecture and addressing complex performance issues ensuring industry-leading service...Remote work
- ...Workato delivers enterprise infrastructure for the agentic era,... ...applications, processes, and AI into a single, governed platform... ...to power real‑time orchestration at scale. With enterprise-... ...looking for an exceptional Staff Software Engineer to help define and build the...Immediate startRemote workFlexible hours
$269.1k - $307.2k
Distinguished AI Engineer (Agentic AI Platform)... ...investments in technology infrastructure and world-class talent... ...APIs that cover agent orchestration, RAG pipelines, prompt... ...hours, mentoring Staff, Principal and Senior... ...technologies (e.g. LLM Inference, Similarity Search and...Full timePart timeWork at officeLocal area$175k - $300k
...intersection of platform engineering, site reliability,... ...of Meshy's AI model serving stack,... ...with core engineering infrastructure. The team operates a... ...capabilities for the AI inference platform, including key... ...scheduling, service orchestration, elastic scaling, and...Work at officeRemote workFlexible hours$151.8k - $265.35k
...content effortlessly. The AI for Engineering team builds a scalable, production... ....We are seeking a Staff Engineer - AI for Engineering... ...dynamics, failure modes, and orchestration strategies — and can... ...integration, memory systems, inference services, data flows, evaluation...Full timeTemporary workLocal areaWorldwide$220k
Perplexity is looking for an engineer to join their team in San Francisco. You will work on building and operating the inference engine, supporting new models, migrating GPU kernels, and developing a Rust-based serving runtime. The ideal candidate has 3+ years of experience...- Infinity Artificial Intelligence Institute in San Francisco Bay Area seeks an AI Infrastructure Engineer to advance an inference stack that spans multiple chips and models. You will build optimization kernels, a scalable library generator, and a benchmark-driven pipeline...
$300k
...interpretable, and steerable AI systems. We want... ...researchers, engineers, policy experts,... ...The Cloud Inference team scales and optimizes... ..., and make smart infrastructure decisions that... ...or container orchestration Have strong interest... ..., we expect all staff to be in one of...Full timeWork at officeVisa sponsorshipFlexible hours$180k - $225k
As a Software Engineer on the ML Infrastructure team, you will design and build platforms... ...with containers and orchestration tools (e.g., Docker, Kubernetes... ...-LLM, or text-generation-inference.Compensation packages at... ...mission is to develop reliable AI systems for the world's...Full time- ...Using frontier causal inference-based econometric... ...managers, economists, and engineers from Google, Netflix,... ...library, a science orchestration library, and a Metaflow... ...team that leans into AI-assisted development... ..., A/B testing infrastructure, or causal inference...Full timeWork at officeWork from homeWorldwideFlexible hours
- ...products.As a Senior Lead Software Engineer at JPMorgan Chase within the Corporate Sector, Infrastructure Platforms team, you are an... ...cloud platforms optimized for AI/ML workloads.Partner with AI teams... ...architecture, ML training, and inference.Experience with Infrastructure...For contractors
- Artificial Analysis, the leading AI benchmarking company, is hiring a Member of Technical Staff to own serverless inference coverage and drive performance benchmarks across provider... ...serving performance. You will work with engineers and leadership to direct the roadmap of...
- Sail builds the world’s most efficient software for inference and agent hosting. In this role, you’ll own token processing down to the lowest layers of the stack, optimize kernel performance, develop new request scheduling and parallelism strategies, and help us use a heterogeneous...
- ...to be the world's leading generative AI studio — we're the team behind niji・journey... ...niji・journey , is looking for an AI Infrastructure Engineer to join us in building out end-to-end... ...implement and run our next-generation inference architecture for running all our...Work experience placementWork at officeVisa sponsorship
- ...We're a team of ex-Google engineers who built some of the largest defensive platforms on... ...: stopping the new wave of adversarial AI attacks already hitting organizations today... .... These agents will operate under an orchestration layer designed for rapid iteration and adaptive...Full timeFlexible hours
$150k - $237.5k
...full stack: compute infrastructure, telemetry, asset modeling... ...operational control orchestration, and integration with... ...trust, learning, and engineering excellence, and we... ...distributed systems AI-accelerated... ...employment information, and inferences drawn from your PI. We...Full timeFlexible hours$197.3k - $225.1k
...Overview Lead AI Engineer (MLX, Agentic AI, Gen AI platform Services) At Capital... ...experiences. Our investments in technology infrastructure and world-class talent — along with... ...model training, large language model inference, similarity search, guardrails, model...Full timePart timeLocal area$229.9k - $262.4k
Senior Lead AI Engineer (Gen AI Platform Services) Overview: At Capital One, we are creating... .... Our investments in technology infrastructure and world-class talent — along with our... ...model training, large language model inference, similarity search, guardrails, model...Full timePart timeLocal area- ...& Tech, Enterprise Software, AI and Machine Learning, SaaS &... ...and process to power real-time orchestration at scale. With enterprise-... ...RoleOur client is looking for a Staff Product Manager for Developer... ..., business goals, and engineering teams to drive the vision and...Permanent employmentFull timeImmediate startRemote workFlexible hours
$150k - $200k
...We're a small, fast-moving AI infrastructure company building an open-source... ...demanding AI workloads including inference pipelines, distributed training, storage orchestration, and GPU scheduling. Our... ...cluster in days. As a Platform Engineer, you'll work shoulder-to-...Full timeVisa sponsorship- ...Known - Conversational AI Engineer, System Prompt and Orchestration ~ San Francisco, CA (In-Person) ~170k-220k Cash + Equity Known is a matchmaker that talks to users and supports them like a friend. Our mission is to empower humanity by applying general intelligence...Full time
$102k - $125k
...and purpose. By blending design, product engineering, analytics, and automation, we build the... ..., you’ll design and deliver innovative AI/ML and agentic AI solutions as part of intelligent... ...frameworks using cutting-edge orchestration frameworks, tool-use protocols, and cloud...Temporary workWork at officeLocal areaFlexible hours$170k - $216k
...across 15+ U.S. states. The Simulation Infrastructure team creates reliable, scalable, and... ...for a broad range of customers Software Engineers, Product, Data Science, System... ...You will: Build and evolve ML inference infrastructure for simulations. Be responsible...Full timeRemote work$181.37k - $235.02k
...customers. We want 6sense to be the best chapter of your career. Staff Enterprise AI Engineer, IT OperationsAbout the RoleWe’re looking for a Staff... ...frameworks, Microsoft 365 Copilot, and ClaudeBuild and orchestrate AI agents that automate cross-functional workflows across...Full time$161.7k - $303.3k
...marketing teams operate — not by layering AI tools on top of existing workflows, but... ...a team of 6-8 Forward-Deployed AI Engineers embedded across GMI's paid media, lifecycle... ...agentic systems, RAG, LLM evaluation, tool orchestration. Your team will bring you real problems...Full timeTemporary workLocal areaWorldwide
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Staff AI Infrastructure Engineer Orchestration & Inference. Be the first to apply!
- software engineer staff San Francisco, CA
- assistant engineer San Francisco, CA
- engineering aide San Francisco, CA
- staff engineer San Francisco, CA
- staff security engineer San Francisco, CA
- assistant mechanical engineer San Francisco, CA
- assistant engineering manager San Francisco, CA
- senior staff systems engineer San Francisco, CA
- technology administrator San Francisco, CA
- project engineer assistant project manager San Francisco, CA



