Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Staff Engineer, Frontier AI Inference

$350k

Mirendil

Mirendil in San Francisco is searching for an engineer to develop and optimize inference systems for cutting-edge AI models. You will handle the complete inference stack, enhancing performance and reliability. The role involves partnering with teams to deploy new architectures and implement optimizations such as quantization and caching strategies. With a focus on innovation, you will contribute to groundbreaking AI research. A competitive base salary of $350,000–$500,000 USD along with equity and benefits is offered. #J-18808-Ljbffr Mirendil

Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Staff Engineer, Frontier AI Inference in San Francisco, CA vacancy
  • $252k - $315k

    About Scale AIScale AI is the data foundation for AI, helping organizations build...  ...their AI transformation through frontier AI systems that solve real business problems...  ...problems remains one of the hardest engineering challenges.As a Staff Frontier Agent Engineer (Applied AI),... 
    Suggested
    Full time

    Scale AI

    San Francisco, CA
    3 days ago
  •  ...Francisco is hiring Members of Technical Staff to build systems that accelerate LLM inference and own customer workloads end to...  ...-performance kernels, inference engine internals, and production...  ...collaborating with a fast-growing AI inference company. #J-18808-Ljbffr... 
    Suggested

    Simplify

    San Francisco, CA
    1 day ago
  • B Capital is seeking a skilled engineer for GPU infrastructure in San Francisco. This role...  ...operating high-performance systems for model inference, synthetic data generation, and...  ...and a passion for working in cutting-edge AI. Benefits include top-tier compensation,... 
    Suggested

    B Capital

    San Francisco, CA
    4 days ago
  •  ...an infrastructure layer for AI workloads, covering training,...  ...deployment, observation, and inference. You will perform hands-on inference...  ...alongside Forward Deployed Engineers to deploy and tune models,...  ...external labs and turning frontier techniques into usable #J-18... 
    Suggested

    Mixpeek

    San Francisco, CA
    5 days ago
  • $200k - $400k

    A leading AI technology company located in San Francisco is seeking an infrastructure engineer to build distributed systems for their AI inference engine. The role involves designing systems that ensure minimal latency and maximum reliability. Candidates should have a... 
    Suggested
    Visa sponsorship

    Inferact

    San Francisco, CA
    4 days ago
  • $190.9k - $232.8k

    A leading data and AI company is seeking a Staff Software Engineer for GenAI inference to lead the architecture and optimization of the inference engine. The role requires expertise in CUDA, GPU programming, and distributed systems design. Ideal candidates will have a strong... 

    Jobleads-US

    San Francisco, CA
    6 days ago
  • Sail Research in San Francisco is seeking a talented engineer to design and implement robust systems that ensure fast and cost-efficient AI inference at global scale. You will be responsible for building high-performance schedulers and optimizing global routing while focusing... 

    Sail Research

    San Francisco, CA
    3 days ago
  • Kindredventures is recruiting infrastructure engineers to scale large-scale inference and evaluation around a physics-based LPM initiative. You will work on high-throughput systems, latency-optimized serving, and distributed orchestration across Kubernetes, Ray, and Slurm... 

    Kindredventures

    San Francisco, CA
    3 days ago
  • Halluminate, an applied AI and data company in San Francisco, seeks a leader to drive...  ...and post-training efforts. You will train frontier models, evaluate quality, and build...  ...contribute to platform tooling, and help grow the engineering team. Expect startup speed, strong... 

    Halluminate

    San Francisco, CA
    3 days ago
  • $250k - $300k

     ...intelligence. As the only vertically integrated AI infrastructure company built from the...  ...in production. That means owning the inference stack end to end: profiling where time...  ...you will also work directly with customer engineering teams to tailor deployments to their... 
    Temporary work

    Crusoe

    San Francisco, CA
    4 days ago
  •  ...running the world’s best data and AI infrastructure platform so...  ...their business. Founded by engineers — and customer obsessed — we...  ...production, across model serving, inference, retrieval, and agent...  ...Background working at or with a frontier lab or AI-native cloud provider... 
    Worldwide

    DataBricks

    San Francisco, CA
    3 days ago
  •  ...providing trusted decision-ready AI to the world's most...  ...real consequences on. As a Staff Machine Learning Engineer, you’ll own AI-driven...  .... You’re energized by what frontier models and agents make newly...  ...-latency, high-concurrency inference (Triton, vLLM, GPU-backed serving... 
    Full time
    Contract work
    Remote work
    Flexible hours

    Primer.ai

    San Francisco, CA
    13 hours ago
  • $220k - $280k

     ...About the Role Together AI is building the best inference infrastructure for voice applications....  ...reliability. We're looking for a Staff ML Engineer to drive the model serving layer...  ...pushing latency and throughput to the frontier. You'll profile GPU utilization,... 
    Full time

    Together Ai

    San Francisco, CA
    13 hours ago
  • $203.5k - $299.3k

     ...question.About the RoleWe are hiring a Causal Machine Learning Engineer to help build the causal ML foundation behind how DoorDash...  ...about you because you have…Deep practical experience with causal inference, econometrics, experimentation, or causal ML.Experience shipping... 
    Hourly pay
    Work at office
    Local area
    Remote work
    Flexible hours

    Doordash

    San Francisco, CA
    4 hours ago
  • Sail builds the world’s most efficient software for inference and agent hosting. In this role, you’ll own token processing down to the lowest layers of the stack, optimize kernel performance, develop new request scheduling and parallelism strategies, and help us use a heterogeneous... 

    Sail

    San Francisco, CA
    2 days ago
  •  ...building the best way to talk to AI and humans together —...  .... Member of Technical Staff is the title we use for engineers who own hard problems end...  ...evaluating, and routing across frontier models like Opus 4.7 and...  ...fine-tuning, evaluation, inference, or RAG at scale High-performance... 

    Shapes, Inc

    San Francisco, CA
    6 days ago
  • Hamilton Barnes Associates Limited is seeking a Staff-level engineer to architect and evolve the AI infrastructure software stack for large-scale GPU workloads...  ...a live production environment. You’ll contribute to inference platforms, model serving, and high-utilization... 

    Hamilton Barnes Associates Limited

    San Francisco, CA
    1 day ago
  • Description We are seeking a Member of Technical Staff - Mechanical Engineer to design and develop robotic manipulation hardware within a frontier AI and robotics research lab. You will own the mechanical design of manipulation-focused hardware end-to-end: grippers, end... 
    Local area

    Amazon Science

    San Francisco, CA
    5 days ago
  • $220k

    We build and run the inference engine behind every Perplexity query and deploy dozens of model architectures at scale with tight latency and cost budgets. Our stack is Rust, Python, CUDA, and CuTe DSL - and we need another engineer to join us. What you will work on Examples... 

    Perplexity

    San Francisco, CA
    3 days ago
  •  ...leader to own a self-serve GPU compute platform for training and inference workloads. You will design and operate the system that lets...  ...tolerance, observability, and a coherent platform roadmap for scalable AI workloads. #J-18808-Ljbffr United States Digital Space LLC

    United States Digital Space LLC

    San Francisco, CA
    3 days ago
  • Together AI in San Francisco is seeking a Staff ML Systems Engineer to design and prototype algorithms, architectures, and scheduling for low-latency, high-throughput inference. You will implement changes in production-grade inference engines, including kernel backends... 

    Together

    San Francisco, CA
    2 days ago
  •  ...DesignArena, invites you to join a talent-dense team in San Francisco with 5.5M+ users and a rapidly growing platform. You will define how frontier AI models are measured, design new benchmarks, run experiments, and publish analyses that become industry gold standards. We sponsor... 
    Relocation
    Visa sponsorship

    Intelligence

    San Francisco, CA
    4 days ago
  • $227.33k - $312.58k

    We’re looking for a Staff ML Data Engineer to join Procore’s AI & Frontier Models organization. In this role, you’ll be responsible for designing and building...  ...support machine learning training, evaluation, or inference workflows.Solid understanding of data modeling, dataset... 
    Full time
    Work at office
    Local area
    Immediate start
    3 days per week

    Procore Technologies

    San Francisco, CA
    2 days ago
  •  ...empowering and governing autonomous AI agents across industries....  ...The Role We are seeking a Staff Research Engineer, AI/ML & Cybersecurity to...  ...grade ML components Improve inference performance, observability,...  ...infrastructure Operate at the frontier of neural systems and... 

    Ephapsys

    San Francisco, CA
    1 day ago
  • Modal is building an infrastructure layer for AI and is seeking strong engineers to optimize ML systems for performance at scale. You will contribute to Modal’s container runtime and open-source projects, pushing language and diffusion models toward higher throughput and... 

    Modal

    San Francisco, CA
    3 days ago
  •  ...Francisco, CA, is seeking a Member of Technical Staff for distributed systems to design, build, and operate the platform that schedules AI workloads across thousands of nodes. This...  .... You will collaborate with founders and engineers from Nvidia, Google AI, Intel, and Pixie... 

    Acceler8 Talent

    San Francisco, CA
    2 days ago
  • Jaide Health is seeking an engineer for their Model Efficiency team in San Francisco. The role focuses on building reliable ML systems...  ...plus strong skills in C++ or Python and insights into the LLM inference ecosystem. A commitment to diversity and inclusive work... 
    Remote job

    Jaide Health

    San Francisco, CA
    1 day ago
  •  ...’s Possible. At Pinterest, AI isn't just a feature, it's a powerful...  .... We are looking for a Staff MLE to lead the technical...  ...better understand intention and infer interests from online activity...  ...lifecycle. Coach and mentor engineers while collaborating with... 
    Full time
    Work at office
    Remote work
    Relocation
    Relocation package

    Pinterest

    San Francisco, CA
    13 hours ago
  • $137.1k - $201.6k

     ...forming a new team that will leverage AI and advanced ML to power decision...  ...About the Role We’re looking for a Staff Machine Learning Engineer to drive the design and development of...  ...you will: Contribute to Causal inference modeling to measure the incremental impact... 
    Hourly pay
    Full time
    Work at office
    Local area
    Remote work
    Flexible hours

    From Restaurants Near You

    San Francisco, CA
    13 hours ago
  • $180k - $336k

     ...Your Impact at LILA We are growing our Applied AI org and seeking talented Senior/Staff Machine Learning Engineers with expertise in LLM training, evaluation, and...  ...scientific needs, with a focus on turning frontier model capabilities into reliable workflows that... 
    Full time
    Work at office
    Local area
    Flexible hours

    Lila Sciences

    San Francisco, CA
    13 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Staff Engineer, Frontier AI Inference. Be the first to apply!