Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

AI Performance Engineer

$215k - $285k
Full-time

Applied Intuition

Applied Intuition, Inc. is powering the future of physical AI. Founded in 2017 and now valued at $15 billion, the Silicon Valley company is creating the digital infrastructure needed to bring intelligence to every moving machine on the planet. Applied Intuition services the automotive, defense, trucking, construction, mining and agriculture industries in three core areas: tools and infrastructure, operating systems, and autonomy. Eighteen of the top 20 global automakers, as well as the United States military and its allies, trust the company’s solutions to deliver physical intelligence. Applied Intuition is headquartered in Sunnyvale, California, with offices in Washington, D.C.; San Diego; Ft. Walton Beach, Florida; Ann Arbor, Michigan; London; Stuttgart; Munich; Stockholm; Bangalore; Seoul; and Tokyo. Learn more at applied.co.

We are an in-office company, and our expectation is that full-time employees primarily work from their Applied Intuition office 5 days a week. However, we also recognize the importance of flexibility and trust our employees to manage their schedules responsibly. This may include occasional remote work, starting the day with morning meetings from home before heading to the office, or leaving earlier when needed to accommodate family commitments. This in-office expectation does not apply to contractor positions

About The Role

We are looking for a performance engineer who specializes in making large-scale machine learning workloads fast and cost-efficient in the datacenter. This role is focused on distributed training runs spanning many nodes, and high-throughput batch inference sweeping petabytes of real-world autonomy logs for auto-labeling, data mining, ground-truth generation, and evaluation.

The optimization target here is not tail latency on a vehicle - it is throughput, cluster goodput, and cost per unit of data processed. A training run that wastes 30% of its GPU-hours on stalled data loaders, or an offline inference sweep that takes a week instead of a day, directly slows down how fast the whole company can iterate. You will own the gap between what our fleet of accelerators is theoretically capable of and what our workloads actually achieve: profiling across the stack, finding where the compute and the wall-clock time actually go, and closing the difference.

You will work at the intersection of accelerators, ML frameworks, and large-scale data infrastructure, partnering with the teams who own each layer to land wins that show up in training time-to-result and offline processing cost. At Applied, we encourage all engineers to take ownership over technical and product decisions, closely interact with users to collect feedback, and contribute to a thoughtful, dynamic team culture.

At Applied, You Will

  • Profile and optimize distributed training end to end - data loading and preprocessing, augmentation, kernel execution, gradient communication, and checkpointing
  • Optimize large-scale offline and batch inference over petabyte-scale sensor logs: batching and scheduling strategies, quantization and low-precision execution, graph optimization, and accelerator saturation across long-running sweeps
  • Establish roofline and performance models for our workloads, quantify the gap between achieved and theoretical performance, and stack-rank optimization opportunities by impact and effort
  • Improve multi-node scaling efficiency: sharding and parallelism strategies, collective communication, interconnect utilization, and memory-bandwidth and kernel-fusion bottlenecks
  • Drive cluster goodput - reduce GPU idle time from input pipeline stalls, storage and network I/O, scheduling gaps, stragglers, and failure recovery on long-running jobs
  • Build the benchmarking, observability, and regression-detection tooling that keeps performance from silently degrading as models and code evolve
  • Collaborate with engineers across functions to solve complex data and compute problems at scale
  • Contribute to a team culture that values effective collaboration, technical excellence, and innovation

We're Looking For Someone Who Has

  • Hands-on ML performance engineering experience: profiling, roofline analysis, throughput optimization, and root-cause investigation in production systems
  • Experience with distributed multi-node training at scale (FSDP, DeepSpeed, Megatron, NCCL, or equivalent), including diagnosing scaling inefficiency as node count grows
  • Deep familiarity with GPU or accelerator performance concepts - memory bandwidth, kernel launch overhead, occupancy, quantization, collective communication
  • Experience with high-throughput or batch inference systems (NVIDIA Triton Inference Server, TensorRT, ONNX Runtime, Ray, or similar)
  • Fluency in Python and proficiency in C++ or another systems language
  • Excellent debugging, analytical, and problem-solving skills
  • A deep understanding of machine learning foundations, and the ability to develop technical solutions for problems with no established playbook

Nice To Have

  • GPU kernel development experience: CUDA, Triton, CUTLASS, or hand-tuned attention implementations
  • Experience with profiling toolchains such as Nsight Systems/Compute, PyTorch Profiler, or perf
  • Experience with GPU scheduling and orchestration on Kubernetes, Slurm, or Ray, including multi-tenant cluster utilization
  • Experience with fault tolerance and elastic training for long-running jobs - checkpointing strategy, straggler mitigation, preemption recovery
  • Familiarity with autonomy or robotics data (ROS, OpenCV, multi-sensor log formats)

Don’t meet every single requirement? If you’re excited about this role but your past experience doesn’t align perfectly with every qualification in the job description, we encourage you to apply anyway. You may be just the right candidate for this or other roles.

Applied Intuition is an equal opportunity employer and federal contractor or subcontractor. Consequently, the parties agree that, as applicable, they will abide by the requirements of 41 CFR 60-1.4(a), 41 CFR 60-300.5(a) and 41 CFR 60-741.5(a) and that these laws are incorporated herein by reference. These regulations prohibit discrimination against qualified individuals based on their status as protected veterans or individuals with disabilities, and prohibit discrimination against all individuals based on their race, color, religion, sex, sexual orientation, gender identity or national origin. These regulations require that covered prime contractors and subcontractors take affirmative action to employ and advance in employment individuals without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, protected veteran status or disability. The parties also agree that, as applicable, they will abide by the requirements of Executive Order 13496 (29 CFR Part 471, Appendix A to Subpart A), relating to the notice of employee rights under federal labor laws.

FOR US-BASED ROLES: Applied Intuition is committed to providing an accessible and inclusive application and interview experience to applicants who are disabled veterans and other applicants with disabilities or medical conditions. Reasonable accommodations are available, requesting an accommodation will not affect your candidacy in any way, and you are not required to disclose the nature of your disability or medical condition in order to make a request.

If you require an accommodation please contact View email address on us.fitly.work . We will work with you!

Compensation Range: $215K - $285K

Vacancy posted 6 days ago
Similar jobs that could be interesting for youBased on the AI Performance Engineer in California vacancy
  •  ...intelligence layer that compounds. The Platform engineer creates the tools, you create the...  ...systems, or automated recommendation loops. AI-first development workflow and ability to...  ...) Background in growth engineering, performance marketing, or data science Experience... 
    Performance
    Full time

    Hellyeah Ai

    San Francisco, CA
    1 day ago
  •  ...For phase 1, we're training foundational AI models that create building code-compliant...  ...models, geometric and physics engines from scratch to transform how the world designs...  ...validation pipelines that measure not just model performance, but constructibility and real-world... 
    Performance
    Full time
    Visa sponsorship

    Augrade Private Limited

    San Francisco, CA
    1 day ago
  • $180k - $300k

     ...About The Role You'll own the core AI systems that power Gamma: the models, prompts...  ...evaluating, and fine-tuning for maximum performance across Gamma's product surface. You'll...  ...our AI stack. You'll work closely with engineering and product to ship improvements that millions... 
    Performance
    Full time
    Work at office
    Immediate start
    Work from home

    Gamma

    San Francisco, CA
    1 day ago
  •  ...We’re hiring an AI Engineer to build the intelligence layer for the leading AI companion for language learning. You’ll own the core AI...  ...systems to measure conversation quality. Optimize real-time performance — drive sub-200ms latency and systems that scale efficiently... 
    Performance
    Full time

    Pingo پینگو

    San Francisco, CA
    1 day ago
  •  ...exceptional and thrive at work through human, AI, and software-based coaching. We’re on a...  ...for a product-minded senior software engineer with experience working with data to help...  ..., teams, and companies to unlock their performance, growth, and how we all work together.... 
    Performance
    Remote job
    Full time
    Home office

    株式会社mento

    San Francisco, CA
    1 day ago
  •  ...Forward Deployed AI Engineer The opportunity We are looking for a Forward Deployed AI Engineer to serve as the critical bridge between...  ...Ensure deployments meet enterprise standards for security, performance and reliability. Customer advocacy & product feedback:... 
    Performance
    Full time
    Shift work

    Latent Labs

    San Francisco, CA
    1 day ago
  •  ...Parasail is redefining AI infrastructure by enabling seamless deployment across a distributed...  ...network of GPUs, optimizing for cost, performance, and flexibility. Our mission is to...  ...We’re looking for a hungry, creative engineer who thrives in a high-trust, high-velocity... 
    Performance
    Full time

    Parasail

    San Francisco, CA
    1 day ago
  • $130k - $196k

     ...analytics infrastructure, and next-generation AI to solve real problems and accelerate...  ...in space. The AI, Data and Platform Engineering team leads Relativity's initiative to...  ...service architectures that support high-performance, user-facing applications Experience... 
    Performance
    Full time

    Relativity Space

    Long Beach, CA
    1 day ago
  • $200k - $350k

     ...multiple systems and platforms. We are seeking a Senior AI Engineer to develop, deploy, and optimize production AI systems. What...  ...Build inference and evaluation pipelines. Optimize model performance, latency, and cost. Integrate AI capabilities into... 
    Performance
    Remote job
    Full time
    Immediate start

    Pragmatike

    San Francisco, CA
    1 day ago
  • $155k - $180k

     ...complexities of risk-based contracting. Arbital AI allows users to interact with complex VBC contracts & performance data in natural language and visualize it on...  ...actionable.  We are looking for an AI Engineer to help build our next-generation conversational... 
    Performance
    Full time
    Contract work
    Work at office
    Remote work
    Flexible hours
    2 days per week

    Arbital Health

    San Francisco, CA
    1 day ago
  • $150k - $350k

     ...About Collate   Collate is an AI document generation platform for life sciences...  ...and founder of Lever. Our AI researchers, engineers, and designers have worked at Google,...  ...applied in Life sciences — a space where performance, safety, and trust matter as much as innovation... 
    Performance
    Full time

    Collate

    San Francisco, CA
    1 day ago
  •  ...Everyone's building AI agents, but almost nobody gets them to production. Building...  ..., governance, and trust become the real engineering challenge. Arcade is the MCP runtime...  ...range below and determined based on a candidate's background, experience, and performance.... 
    Performance
    Full time
    Shift work

    Arcade

    San Francisco, CA
    1 day ago
  • $155k - $180k

     ...our cross-disciplinary skills (embedded AI, high-tech manufacturing automation, eco-...  ...the employee will assume the role of AI engineer, with the following main responsibilities...  ...implement data processing architectures Perform business and technical analysis of data to... 
    Performance
    Full time

    Kickmaker

    San Francisco, CA
    1 day ago
  • $25k

     ...We are looking for a Marketing AI Engineer to sit at the intersection of marketing strategy and cutting-edge AI technology. This 4-month...  ..., ensuring reliable integrations and optimized query performance for each connector.   Skills & Governance Set up centralized... 
    Performance
    Remote job
    Full time
    Contract work
    Temporary work
    Work at office
    Worldwide

    Azul

    Sunnyvale, CA
    1 day ago
  • $50k - $120k

     ...Who are we? Mission Altimate AI, founded in 2022 in San Francisco, is revolutionizing...  ...tools like VSCode, Git, and Slack, performing tasks ranging from data documentation to...  ...at the forefront of the AI-powered data engineering revolution. You can read more about us... 
    Performance
    Full time
    Worldwide

    Pa Early Stage Partners

    Sunnyvale, CA
    1 day ago
  • $200k - $400k

     ...intelligent machines at scale. At Scout AI, we’re developing Fury, the first robotic...  ...We're looking for a Senior or Staff AI Engineer to join the Fury Orchestration Team with...  ...and mission operations to validate system performance e under real-world constraints Qualifications... 
    Performance
    Full time
    Relocation package

    Scout Ai

    Sunnyvale, CA
    1 day ago
  • $160k - $180k

     ...value candidates who are actively using AI tools to enhance productivity, automate...  ...equivalent experience in Computer Science, Engineering, Machine Learning, or other relevant...  ...for this position may also include annual performance bonus, stock, benefits and/or other applicable... 
    Performance
    Full time

    Appzen

    San Jose, CA
    1 day ago
  • $150k - $250k

     ...About Distyl AI Distyl is an applied AI technology company partnering with the world...  ...0+ F500s. What We Are Looking For AI Engineers build and operate production AI systems that...  ...to work directly on AI systems that must perform reliably under enterprise constraints.... 
    Performance
    Full time
    Work at office
    Flexible hours
    3 days per week

    Distyl Ai

    San Francisco, CA
    1 day ago
  •  ...revolutionizing software development with AI-powered formal verification. We've...  ...About the role Join our team as an AI Engineer and help us push the boundaries of what's...  ...programming languages and tools critical for high-performance computing in Python/C++ and machine... 
    Performance
    Full time
    Contract work

    Logical Intelligence

    San Francisco, CA
    1 day ago
  • Do you want to work at the intersection of AI, product, engineering, and real -world operations? \ \ Join a fast -growing technology startup...  ...next -generation software. You will join a small, high -performing team of exceptional builders and work directly with leadership... 
    Performance
    Full time
    Immediate start

    Eth Juniors

    San Francisco, CA
    1 day ago
  • $250k - $300k

     ...About Us At You.com, we are building the AI Search Infrastructure that powers modern...  ..., and useful. Our team includes engineers, researchers, product builders, and operators...  ...and token usage, and getting strong performance out of smaller, cheaper models. Contribute... 
    Performance
    Full time
    Immediate start
    Remote work
    Work from home
    Flexible hours

    You.com

    San Francisco, CA
    1 day ago
  • $90k - $140k

     ...for modern healthcare organizations. Our AI platform streamlines critical workflows across...  ...As a Forward Deployed AI Engineer at Charta Health, you will be at the forefront...  ...experience + Equity + Benefits ~ Annual performance based bonus up to 30%   Our Commitment... 
    Performance
    Full time
    Work at office
    Immediate start

    Charta Health

    San Francisco, CA
    1 day ago
  •  ...Inc. is powering the future of physical AI. Founded in 2017 and now valued at $15 billion...  ...commitments. Meet our software engineers! Meet some of our software engineers...  ...attainment, skill level requirements, interview performance, and the level and scope of the position... 
    Performance
    Full time
    For contractors
    For subcontractor
    Casual work
    Work at office
    Remote work
    Day shift

    Applied Intuition

    Sunnyvale, CA
    1 day ago
  •  ...close that gap. We're building the first AI-powered platform built for hardware product...  ...This is an early role for a hands-on AI engineer to design and build the agents at the...  ...and observability infrastructure for agent performance, including tracing, cost tracking, latency... 
    Performance
    Remote job
    Permanent employment
    Full time
    Contract work
    Shift work

    Re:build Manufacturing

    Los Angeles, CA
    1 day ago
  •  ...The Role As an Applied AI Engineer, you'll bring the frontier of AI research and engineering to Rowspace, transforming how our platform...  ...to customer feedback, design evaluation frameworks for AI performance, and implement advanced retrieval systems to make the most of... 
    Performance
    Full time
    Work at office

    Rowspace

    San Francisco, CA
    1 day ago
  •  ...Description Cooperidge Consulting Firm is seeking an AI Agent Software Engineer for a high-momentum AI platform leader in San Francisco...  ...understandable, debuggable, and continuously improving. Performance Engineering: Analyze large-scale performance data to... 
    Performance
    Full time
    Apprenticeship
    Flexible hours

    Cooperidge Consulting Firm

    San Francisco, CA
    1 day ago
  •  ...FurtherAI, we’re building the next generation of AI agents for the insurance industry - a...  ..., we’re looking for an exceptional AI Engineer to join our agent team and help shape...  ...-to-end evals to measure & improve agent performance. Experiment with new agentic techniques... 
    Performance
    Full time
    Work at office

    Further AI

    San Francisco, CA
    1 day ago
  • $229.9k - $262.4k

    {"description": "Senior Lead AI Engineer (MLX, Agentic AI, Gen AI platform Services) Overview: At Capital One,...  ...capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the... 
    Performance
    Full time
    Part time
    Local area

    Capital One Financial Corporation

    San Jose, CA
    1 day ago
  •  ...revolutionizing software development with AI-powered formal verification. We've...  ...About the role Join our team as an AI Engineer and help us push the boundaries of what's...  ...programming languages and tools critical for high-performance computing in Python/C++ and machine... 
    Performance
    Full time
    Contract work

    Logical Intelligence

    San Francisco, CA
    1 day ago
  •  ...Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This...  ...Role You'll own model quality and performance for Cerebras' inference offerings. You will...  ...this on a loop." You'll sit between engineering, product, and customer-facing teams.... 
    Performance
    Full time

    Cerebras Systems

    Sunnyvale, CA
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to AI Performance Engineer. Be the first to apply!