Staff Engineer, Frontier AI Inference
$350kMirendil
Mirendil in San Francisco is searching for an engineer to develop and optimize inference systems for cutting-edge AI models. You will handle the complete inference stack, enhancing performance and reliability. The role involves partnering with teams to deploy new architectures and implement optimizations such as quantization and caching strategies. With a focus on innovation, you will contribute to groundbreaking AI research. A competitive base salary of $350,000–$500,000 USD along with equity and benefits is offered. #J-18808-Ljbffr Mirendil
$252k - $315k
About Scale AIScale AI is the data foundation for AI, helping organizations build... ...their AI transformation through frontier AI systems that solve real business problems... ...problems remains one of the hardest engineering challenges.As a Staff Frontier Agent Engineer (Applied AI),...SuggestedFull time- ...Francisco is hiring Members of Technical Staff to build systems that accelerate LLM inference and own customer workloads end to... ...-performance kernels, inference engine internals, and production... ...collaborating with a fast-growing AI inference company. #J-18808-Ljbffr...Suggested
- B Capital is seeking a skilled engineer for GPU infrastructure in San Francisco. This role... ...operating high-performance systems for model inference, synthetic data generation, and... ...and a passion for working in cutting-edge AI. Benefits include top-tier compensation,...Suggested
$200k - $400k
A leading AI technology company located in San Francisco is seeking an infrastructure engineer to build distributed systems for their AI inference engine. The role involves designing systems that ensure minimal latency and maximum reliability. Candidates should have a...SuggestedVisa sponsorship$190.9k - $232.8k
A leading data and AI company is seeking a Staff Software Engineer for GenAI inference to lead the architecture and optimization of the inference engine. The role requires expertise in CUDA, GPU programming, and distributed systems design. Ideal candidates will have a strong...Suggested- Sail Research in San Francisco is seeking a talented engineer to design and implement robust systems that ensure fast and cost-efficient AI inference at global scale. You will be responsible for building high-performance schedulers and optimizing global routing while focusing...
$250k - $300k
...intelligence. As the only vertically integrated AI infrastructure company built from the... ...in production. That means owning the inference stack end to end: profiling where time... ...you will also work directly with customer engineering teams to tailor deployments to their...Temporary work- Kindredventures is recruiting infrastructure engineers to scale large-scale inference and evaluation around a physics-based LPM initiative. You will work on high-throughput systems, latency-optimized serving, and distributed orchestration across Kubernetes, Ray, and Slurm...
- Sail is hiring for an engineering role in San Francisco to design and implement high-performance schedulers that optimize admission control... ..., and explore KV caching for memory/compute trade-offs in LLM inference stacks. You will contribute to deep observability, tracing...
- Halluminate, an applied AI and data company in San Francisco, seeks a leader to drive... ...and post-training efforts. You will train frontier models, evaluate quality, and build... ...contribute to platform tooling, and help grow the engineering team. Expect startup speed, strong...
- Together AI is building the best inference infrastructure for voice applications. We seek a Staff ML Engineer to own the model serving stack and optimize latency and throughput for real-time voice workloads. You'll work with state-of-the-art accelerators and collaborate...
- ...running the world’s best data and AI infrastructure platform so... ...their business. Founded by engineers — and customer obsessed — we... ...production, across model serving, inference, retrieval, and agent... ...Background working at or with a frontier lab or AI-native cloud provider...Worldwide
- ...providing trusted decision-ready AI to the world's most... ...real consequences on. As a Staff Machine Learning Engineer, you’ll own AI-driven... .... You’re energized by what frontier models and agents make newly... ...-latency, high-concurrency inference (Triton, vLLM, GPU-backed serving...Full timeContract workRemote workFlexible hours
$220k - $280k
...About the Role Together AI is building the best inference infrastructure for voice applications.... ...reliability. We're looking for a Staff ML Engineer to drive the model serving layer... ...pushing latency and throughput to the frontier. You'll profile GPU utilization,...Full time$203.5k - $299.3k
...question.About the RoleWe are hiring a Causal Machine Learning Engineer to help build the causal ML foundation behind how DoorDash... ...about you because you have…Deep practical experience with causal inference, econometrics, experimentation, or causal ML.Experience shipping...Hourly payWork at officeLocal areaRemote workFlexible hours- Sail builds the world’s most efficient software for inference and agent hosting. In this role, you’ll own token processing down to the lowest layers of the stack, optimize kernel performance, develop new request scheduling and parallelism strategies, and help us use a heterogeneous...
- Hamilton Barnes Associates Limited is seeking a Staff-level engineer to architect and evolve the AI infrastructure software stack for large-scale GPU workloads... ...a live production environment. You’ll contribute to inference platforms, model serving, and high-utilization...
- ...leader to own a self-serve GPU compute platform for training and inference workloads. You will design and operate the system that lets... ...tolerance, observability, and a coherent platform roadmap for scalable AI workloads. #J-18808-Ljbffr United States Digital Space LLC
- Together AI in San Francisco is seeking a Staff ML Systems Engineer to design and prototype algorithms, architectures, and scheduling for low-latency, high-throughput inference. You will implement changes in production-grade inference engines, including kernel backends...
$227.33k - $312.58k
We’re looking for a Staff ML Data Engineer to join Procore’s AI & Frontier Models organization. In this role, you’ll be responsible for designing and building... ...support machine learning training, evaluation, or inference workflows.Solid understanding of data modeling, dataset...Full timeWork at officeLocal areaImmediate start3 days per week- ...DesignArena, invites you to join a talent-dense team in San Francisco with 5.5M+ users and a rapidly growing platform. You will define how frontier AI models are measured, design new benchmarks, run experiments, and publish analyses that become industry gold standards. We sponsor...RelocationVisa sponsorship
- Modal is building an infrastructure layer for AI and is seeking strong engineers to optimize ML systems for performance at scale. You will contribute to Modal’s container runtime and open-source projects, pushing language and diffusion models toward higher throughput and...
- ...empowering and governing autonomous AI agents across industries.... ...The Role We are seeking a Staff Research Engineer, AI/ML & Cybersecurity to... ...grade ML components Improve inference performance, observability,... ...infrastructure Operate at the frontier of neural systems and...
- ...Francisco, CA, is seeking a Member of Technical Staff for distributed systems to design, build, and operate the platform that schedules AI workloads across thousands of nodes. This... .... You will collaborate with founders and engineers from Nvidia, Google AI, Intel, and Pixie...
- Jaide Health is seeking an engineer for their Model Efficiency team in San Francisco. The role focuses on building reliable ML systems... ...plus strong skills in C++ or Python and insights into the LLM inference ecosystem. A commitment to diversity and inclusive work...Remote job
- ...’s Possible. At Pinterest, AI isn't just a feature, it's a powerful... .... We are looking for a Staff MLE to lead the technical... ...better understand intention and infer interests from online activity... ...lifecycle. Coach and mentor engineers while collaborating with...Full timeWork at officeRemote workRelocationRelocation package
$180k - $336k
...Your Impact at LILA We are growing our Applied AI org and seeking talented Senior/Staff Machine Learning Engineers with expertise in LLM training, evaluation, and... ...scientific needs, with a focus on turning frontier model capabilities into reliable workflows that...Full timeWork at officeLocal areaFlexible hours$137.1k - $201.6k
...forming a new team that will leverage AI and advanced ML to power decision... ...About the Role We’re looking for a Staff Machine Learning Engineer to drive the design and development of... ...you will: Contribute to Causal inference modeling to measure the incremental impact...Hourly payFull timeWork at officeLocal areaRemote workFlexible hours- ...opportunityWe are building the next generation of AI-driven game experiences, running... ...that runtime. As a Senior Machine Learning Engineer for On-Device & Mobile AI, you will take... ...deployment of significant parts of the inference stack — from a trained checkpoint leaving...Full timeWork at officeRemote workWorldwide
$197.3k - $313.7k
...DetailsAbout SalesforceSalesforce is the #1 AI CRM, where humans with agents drive... ...OPPORTUNITIES*Slack is looking for a Staff Machine Learning Engineer with deep expertise in model training... ...with model optimization for inference (quantization, pruning, speculative decoding...Full time
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Staff Engineer, Frontier AI Inference. Be the first to apply!
- assistant engineering manager San Francisco, CA
- assistant civil engineer San Francisco, CA
- assistant mechanical engineer San Francisco, CA
- assistant engineer San Francisco, CA
- staff engineer San Francisco, CA
- staff data engineer San Francisco, CA
- software engineer staff San Francisco, CA
- assistant electrical engineer San Francisco, CA
- assistant chief engineer San Francisco, CA
- staff design engineer San Francisco, CA

