Staff Engineer, AI Inference & Distributed Systems
Sail Research Inc.
Sail is building the world's most efficient software for inference and agent hosting. In this role, you'll be one of the first engineers on Sailboxes, contributing across the stack—from low-level performance work on our custom networking stack to building large-scale systems that maximize efficiency. You'll focus on distributed systems, cloud infrastructure, and scalable, ergonomic tooling. The SF office offers a collaborative environment with generous equipment and setup to accelerate #J-18808-Ljbffr Sail Research Inc.
- ...in San Francisco is seeking a talented engineer to design and implement robust systems that ensure fast and cost-efficient AI inference at global scale. You will be responsible... ...ideal candidate has a strong background in distributed systems and is eager to engage in...Suggested
- ...Francisco, CA, is seeking a Member of Technical Staff for distributed systems to design, build, and operate the platform that schedules AI workloads across thousands of nodes. This... .... You will collaborate with founders and engineers from Nvidia, Google AI, Intel, and Pixie...Suggested
- Sail is hiring for an engineering role in San Francisco to design and implement high-performance... ...fleet. You will also build LV routing systems to dispatch workloads with latency... ...caching for memory/compute trade-offs in LLM inference stacks. You will contribute to deep...Suggested
- Kindredventures is recruiting infrastructure engineers to scale large-scale inference and evaluation around a physics-based LPM initiative. You will work on high-throughput systems, latency-optimized serving, and distributed orchestration across Kubernetes, Ray, and Slurm...Suggested
- ...in San Francisco is seeking a Member of Technical Staff to lead distributed systems at the core of our AI infrastructure. You’ll design, build, and operate scheduling... ..., GPUs, and accelerators. You’ll work closely with inference, runtimes, compilers, kernels, and hardware teams...Suggested
- ...in San Francisco is seeking a Member of Technical Staff to design and build distributed systems for AI workloads. The role involves developing scheduling... ...APIs. Ideal candidates should have strong software engineering skills and experience with distributed systems. This...
- ...seeking a highly motivated Member of Technical Staff to join our growing engineering team. You will contribute to inference benchmarks, system modeling, and the preparation of technical... ...with a global team across multiple AI and hardware domains. #J-18808-Ljbffr S27...Daily paid
$150k - $350k
...Inc. is seeking a Member of Technical Staff to focus on distributed systems in San Francisco, California. This... ...and building the core platform for AI workloads, developing resource management... ...should have strong software engineering fundamentals and experience with distributed...- Harrison Clarke is working with a high-growth startup in San Francisco seeking a Staff Distributed Systems Engineer. This hands-on role involves designing and building core systems for a cutting-edge AI code generation product. The ideal candidate will have excellent coding...
- Together AI in San Francisco is seeking a Staff ML Systems Engineer to design and prototype algorithms, architectures, and scheduling for low-latency, high-throughput inference. You will implement changes in production-grade inference engines, including kernel backends...
$200k - $400k
A leading AI technology company located in San Francisco is seeking an infrastructure engineer to build distributed systems for their AI inference engine. The role involves designing systems that ensure minimal latency and maximum reliability. Candidates should have a...Visa sponsorship$150k - $250k
Asari AI in San Francisco is looking for a skilled individual to build the supercomputing infrastructure that runs AI agents,... ...performance workloads. Your role will involve designing cloud compute, distributed systems, and sandboxed tooling to ensure efficiency and scalability....$190.9k - $232.8k
A leading data and AI company is seeking a Staff Software Engineer for GenAI inference to lead the architecture and optimization of the inference engine. The... ...requires expertise in CUDA, GPU programming, and distributed systems design. Ideal candidates will have a strong...$225k
Dormont Manufacturing Co is looking for a Software Engineer on the Inference & RL Systems team in San Francisco. The role involves designing distributed systems, optimizing performance, and ensuring high reliability for RL and post-training workflows. The ideal candidate...- Eventual is seeking a Member of Technical Staff to build Eventual's core products and... ...autonomously solving problems. We value engineers who can scope tasks and implement efficient... ..., with a focus on performance and reliability across distributed #J-18808-Ljbffr Mixpeek
- B Capital is seeking a skilled engineer for GPU infrastructure in San Francisco. This... ...designing and operating high-performance systems for model inference, synthetic data generation, and... ...a passion for working in cutting-edge AI. Benefits include top-tier compensation...
- SemiAnalysis LLC is hiring a Member of Technical Staff to develop training and inference benchmarks and system modelling for AI hardware. You will contribute to writing our newsletter articles with potential authorship recognition and collaborate with a global team of...
- ...building the best way to talk to AI and humans together —... .... Member of Technical Staff is the title we use for engineers who own hard problems end... ...Representative projects Designing systems that give 3M+ AI agents... ...fine-tuning, evaluation, inference, or RAG at scale High-...
- ...Francisco is hiring Members of Technical Staff to build systems that accelerate LLM inference and own customer workloads end to... ...-performance kernels, inference engine internals, and production... ...collaborating with a fast-growing AI inference company. #J-18808-Ljbffr...
- Modal is building an infrastructure layer for AI workloads, covering training, deployment, observation, and inference. You will perform hands-on inference research, selecting... ...work with customers alongside Forward Deployed Engineers to deploy and tune models, while expanding...
$220k
We build and run the inference engine behind every Perplexity query and deploy dozens of model... ...CUTLASS, or similar). Any other deep systems programming experience is a plus. You... ...You've built and operated production distributed systems under real load - ideally performance...- ...building sophisticated infrastructure and software systems. You will join as a Member of Technical Staff to lead high‑impact technical initiatives... ...teams. The role emphasizes architecture, distributed systems, and engineering leadership, with opportunities to influence...
- ...company. We are seeking a Member of Technical Staff to lead architecture and drive high-impact initiatives across multiple systems and teams. You will design scalable distributed systems, own initiatives end-to-end, mentor engineers, and collaborate with Product,...Remote job
$215k - $260k
...the only vertically integrated AI infrastructure company built... .... That means owning the inference stack end to end: profiling where... ...not good enough. This is core systems and performance work on some... ...work directly with customer engineering teams to tailor deployments to...Temporary work- Pragmatike is seeking a Staff level Member of Technical Staff to lead architecture and... ...The role focuses on designing scalable distributed systems, resolving bottlenecks, and shaping... ...and strong leadership, you will mentor engineers and influence critical technology decisions...Remote job
- Inception is seeking engineers and scientists to design, optimize, and scale the diffusion LLM serving systems powering production inference. Your work will help make inference faster, more... ...(Kubernetes, Ray, SLURM) for distributed inference, evaluation, and large-batch...
- ...trusted decision-ready AI to the world's most... ...on. As a Staff Machine Learning Engineer, you’ll own AI-driven... ...works to the production system the mission depends... ...scale, drawing on real distributed-systems experience.... ...latency, high-concurrency inference (Triton, vLLM, GPU-...Full timeContract workRemote workFlexible hours
$155k - $269k
...Description Job Description Waabi, founded by AI visionary Raquel Urtasun, is the leader... ...team of Research Scientists and Engineers building the content backbone of a best-in... ...scale, not just prototypes. You have shipped systems that other engineers and researchers...Full timeWork at officeWork from homeFlexible hours$150k - $225k
...with unprecedented speed and accuracy. Our AI-enabled platform turns siloed and... ...team — which means you're not inheriting a system, you're designing one. You'll be the second... ...identity, SaaS, and cloud platforms.The work is engineering-first: workflow automation, system...Local area$203.5k - $299.3k
...the next generation of causal decisioning systems for New Verticals: grocery, convenience,... ...are hiring a Causal Machine Learning Engineer to help build the causal ML foundation behind... ...…Deep practical experience with causal inference, econometrics, experimentation, or...Hourly payWork at officeLocal areaRemote workFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Staff Engineer, AI Inference & Distributed Systems. Be the first to apply!
- project engineer assistant project manager San Francisco, CA
- senior staff systems engineer San Francisco, CA
- staff data engineer San Francisco, CA
- assistant chief engineer San Francisco, CA
- assistant engineer San Francisco, CA
- assistant electrical engineer San Francisco, CA
- engineering aide San Francisco, CA
- software engineer staff San Francisco, CA
- staff design engineer San Francisco, CA
- research assistant engineering San Francisco, CA


