AI Performance Engineer
$215k - $285kApplied Intuition
We are an in-office company, and our expectation is that full-time employees primarily work from their Applied Intuition office 5 days a week. However, we also recognize the importance of flexibility and trust our employees to manage their schedules responsibly. This may include occasional remote work, starting the day with morning meetings from home before heading to the office, or leaving earlier when needed to accommodate family commitments. This in-office expectation does not apply to contractor positions About The Role
We are looking for a performance engineer who specializes in making large-scale machine learning workloads fast and cost-efficient in the datacenter. This role is focused on distributed training runs spanning many nodes, and high-throughput batch inference sweeping petabytes of real-world autonomy logs for auto-labeling, data mining, ground-truth generation, and evaluation. The optimization target here is not tail latency on a vehicle - it is throughput, cluster goodput, and cost per unit of data processed. A training run that wastes 30% of its GPU-hours on stalled data loaders, or an offline inference sweep that takes a week instead of a day, directly slows down how fast the whole company can iterate. You will own the gap between what our fleet of accelerators is theoretically capable of and what our workloads actually achieve: profiling across the stack, finding where the compute and the wall-clock time actually go, and closing the difference. You will work at the intersection of accelerators, ML frameworks, and large-scale data infrastructure, partnering with the teams who own each layer to land wins that show up in training time-to-result and offline processing cost. At Applied, we encourage all engineers to take ownership over technical and product decisions, closely interact with users to collect feedback, and contribute to a thoughtful, dynamic team culture. At Applied, You Will- Profile and optimize distributed training end to end - data loading and preprocessing, augmentation, kernel execution, gradient communication, and checkpointing
- Optimize large-scale offline and batch inference over petabyte-scale sensor logs: batching and scheduling strategies, quantization and low-precision execution, graph optimization, and accelerator saturation across long-running sweeps
- Establish roofline and performance models for our workloads, quantify the gap between achieved and theoretical performance, and stack-rank optimization opportunities by impact and effort
- Improve multi-node scaling efficiency: sharding and parallelism strategies, collective communication, interconnect utilization, and memory-bandwidth and kernel-fusion bottlenecks
- Drive cluster goodput - reduce GPU idle time from input pipeline stalls, storage and network I/O, scheduling gaps, stragglers, and failure recovery on long-running jobs
- Build the benchmarking, observability, and regression-detection tooling that keeps performance from silently degrading as models and code evolve
- Collaborate with engineers across functions to solve complex data and compute problems at scale
- Contribute to a team culture that values effective collaboration, technical excellence, and innovation
- Hands-on ML performance engineering experience: profiling, roofline analysis, throughput optimization, and root-cause investigation in production systems
- Experience with distributed multi-node training at scale (FSDP, DeepSpeed, Megatron, NCCL, or equivalent), including diagnosing scaling inefficiency as node count grows
- Deep familiarity with GPU or accelerator performance concepts - memory bandwidth, kernel launch overhead, occupancy, quantization, collective communication
- Experience with high-throughput or batch inference systems (NVIDIA Triton Inference Server, TensorRT, ONNX Runtime, Ray, or similar)
- Fluency in Python and proficiency in C++ or another systems language
- Excellent debugging, analytical, and problem-solving skills
- A deep understanding of machine learning foundations, and the ability to develop technical solutions for problems with no established playbook
- GPU kernel development experience: CUDA, Triton, CUTLASS, or hand-tuned attention implementations
- Experience with profiling toolchains such as Nsight Systems/Compute, PyTorch Profiler, or perf
- Experience with GPU scheduling and orchestration on Kubernetes, Slurm, or Ray, including multi-tenant cluster utilization
- Experience with fault tolerance and elastic training for long-running jobs - checkpointing strategy, straggler mitigation, preemption recovery
- Familiarity with autonomy or robotics data (ROS, OpenCV, multi-sensor log formats)
- ...intelligence layer that compounds. The Platform engineer creates the tools, you create the... ...systems, or automated recommendation loops. AI-first development workflow and ability to... ...) Background in growth engineering, performance marketing, or data science Experience...PerformanceFull time
- ...For phase 1, we're training foundational AI models that create building code-compliant... ...models, geometric and physics engines from scratch to transform how the world designs... ...validation pipelines that measure not just model performance, but constructibility and real-world...PerformanceFull timeVisa sponsorship
$180k - $300k
...About The Role You'll own the core AI systems that power Gamma: the models, prompts... ...evaluating, and fine-tuning for maximum performance across Gamma's product surface. You'll... ...our AI stack. You'll work closely with engineering and product to ship improvements that millions...PerformanceFull timeWork at officeImmediate startWork from home- ...We’re hiring an AI Engineer to build the intelligence layer for the leading AI companion for language learning. You’ll own the core AI... ...systems to measure conversation quality. Optimize real-time performance — drive sub-200ms latency and systems that scale efficiently...PerformanceFull time
- ...exceptional and thrive at work through human, AI, and software-based coaching. We’re on a... ...for a product-minded senior software engineer with experience working with data to help... ..., teams, and companies to unlock their performance, growth, and how we all work together....PerformanceRemote jobFull timeHome office
- ...Forward Deployed AI Engineer The opportunity We are looking for a Forward Deployed AI Engineer to serve as the critical bridge between... ...Ensure deployments meet enterprise standards for security, performance and reliability. Customer advocacy & product feedback:...PerformanceFull timeShift work
- ...Parasail is redefining AI infrastructure by enabling seamless deployment across a distributed... ...network of GPUs, optimizing for cost, performance, and flexibility. Our mission is to... ...We’re looking for a hungry, creative engineer who thrives in a high-trust, high-velocity...PerformanceFull time
$130k - $196k
...analytics infrastructure, and next-generation AI to solve real problems and accelerate... ...in space. The AI, Data and Platform Engineering team leads Relativity's initiative to... ...service architectures that support high-performance, user-facing applications Experience...PerformanceFull time$200k - $350k
...multiple systems and platforms. We are seeking a Senior AI Engineer to develop, deploy, and optimize production AI systems. What... ...Build inference and evaluation pipelines. Optimize model performance, latency, and cost. Integrate AI capabilities into...PerformanceRemote jobFull timeImmediate start$155k - $180k
...complexities of risk-based contracting. Arbital AI allows users to interact with complex VBC contracts & performance data in natural language and visualize it on... ...actionable. We are looking for an AI Engineer to help build our next-generation conversational...PerformanceFull timeContract workWork at officeRemote workFlexible hours2 days per week$150k - $350k
...About Collate Collate is an AI document generation platform for life sciences... ...and founder of Lever. Our AI researchers, engineers, and designers have worked at Google,... ...applied in Life sciences — a space where performance, safety, and trust matter as much as innovation...PerformanceFull time- ...Everyone's building AI agents, but almost nobody gets them to production. Building... ..., governance, and trust become the real engineering challenge. Arcade is the MCP runtime... ...range below and determined based on a candidate's background, experience, and performance....PerformanceFull timeShift work
$155k - $180k
...our cross-disciplinary skills (embedded AI, high-tech manufacturing automation, eco-... ...the employee will assume the role of AI engineer, with the following main responsibilities... ...implement data processing architectures Perform business and technical analysis of data to...PerformanceFull time$25k
...We are looking for a Marketing AI Engineer to sit at the intersection of marketing strategy and cutting-edge AI technology. This 4-month... ..., ensuring reliable integrations and optimized query performance for each connector. Skills & Governance Set up centralized...PerformanceRemote jobFull timeContract workTemporary workWork at officeWorldwide$50k - $120k
...Who are we? Mission Altimate AI, founded in 2022 in San Francisco, is revolutionizing... ...tools like VSCode, Git, and Slack, performing tasks ranging from data documentation to... ...at the forefront of the AI-powered data engineering revolution. You can read more about us...PerformanceFull timeWorldwide$200k - $400k
...intelligent machines at scale. At Scout AI, we’re developing Fury, the first robotic... ...We're looking for a Senior or Staff AI Engineer to join the Fury Orchestration Team with... ...and mission operations to validate system performance e under real-world constraints Qualifications...PerformanceFull timeRelocation package$160k - $180k
...value candidates who are actively using AI tools to enhance productivity, automate... ...equivalent experience in Computer Science, Engineering, Machine Learning, or other relevant... ...for this position may also include annual performance bonus, stock, benefits and/or other applicable...PerformanceFull time$150k - $250k
...About Distyl AI Distyl is an applied AI technology company partnering with the world... ...0+ F500s. What We Are Looking For AI Engineers build and operate production AI systems that... ...to work directly on AI systems that must perform reliably under enterprise constraints....PerformanceFull timeWork at officeFlexible hours3 days per week- ...revolutionizing software development with AI-powered formal verification. We've... ...About the role Join our team as an AI Engineer and help us push the boundaries of what's... ...programming languages and tools critical for high-performance computing in Python/C++ and machine...PerformanceFull timeContract work
- Do you want to work at the intersection of AI, product, engineering, and real -world operations? \ \ Join a fast -growing technology startup... ...next -generation software. You will join a small, high -performing team of exceptional builders and work directly with leadership...PerformanceFull timeImmediate start
$250k - $300k
...About Us At You.com, we are building the AI Search Infrastructure that powers modern... ..., and useful. Our team includes engineers, researchers, product builders, and operators... ...and token usage, and getting strong performance out of smaller, cheaper models. Contribute...PerformanceFull timeImmediate startRemote workWork from homeFlexible hours$90k - $140k
...for modern healthcare organizations. Our AI platform streamlines critical workflows across... ...As a Forward Deployed AI Engineer at Charta Health, you will be at the forefront... ...experience + Equity + Benefits ~ Annual performance based bonus up to 30% Our Commitment...PerformanceFull timeWork at officeImmediate start- ...Inc. is powering the future of physical AI. Founded in 2017 and now valued at $15 billion... ...commitments. Meet our software engineers! Meet some of our software engineers... ...attainment, skill level requirements, interview performance, and the level and scope of the position...PerformanceFull timeFor contractorsFor subcontractorCasual workWork at officeRemote workDay shift
- ...close that gap. We're building the first AI-powered platform built for hardware product... ...This is an early role for a hands-on AI engineer to design and build the agents at the... ...and observability infrastructure for agent performance, including tracing, cost tracking, latency...PerformanceRemote jobPermanent employmentFull timeContract workShift work
- ...The Role As an Applied AI Engineer, you'll bring the frontier of AI research and engineering to Rowspace, transforming how our platform... ...to customer feedback, design evaluation frameworks for AI performance, and implement advanced retrieval systems to make the most of...PerformanceFull timeWork at office
- ...Description Cooperidge Consulting Firm is seeking an AI Agent Software Engineer for a high-momentum AI platform leader in San Francisco... ...understandable, debuggable, and continuously improving. Performance Engineering: Analyze large-scale performance data to...PerformanceFull timeApprenticeshipFlexible hours
- ...FurtherAI, we’re building the next generation of AI agents for the insurance industry - a... ..., we’re looking for an exceptional AI Engineer to join our agent team and help shape... ...-to-end evals to measure & improve agent performance. Experiment with new agentic techniques...PerformanceFull timeWork at office
$229.9k - $262.4k
{"description": "Senior Lead AI Engineer (MLX, Agentic AI, Gen AI platform Services) Overview: At Capital One,... ...capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the...PerformanceFull timePart timeLocal area- ...revolutionizing software development with AI-powered formal verification. We've... ...About the role Join our team as an AI Engineer and help us push the boundaries of what's... ...programming languages and tools critical for high-performance computing in Python/C++ and machine...PerformanceFull timeContract work
- ...Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This... ...Role You'll own model quality and performance for Cerebras' inference offerings. You will... ...this on a loop." You'll sit between engineering, product, and customer-facing teams....PerformanceFull time
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to AI Performance Engineer. Be the first to apply!
- ai engineer California
- ai engineer remote California
- ai developer California
- senior ai engineer California
- senior performance engineer California
- acting performance California
- performance test architect California
- performance improvement consultant California
- senior performance tester California
- performance windows California

