ML Runtime Optimization Engineer
$159.05k - $199.3kFull-time
Applied Intuition
About Applied Intuition Applied Intuition, Inc. is powering the future of physical AI. Founded in 2017 and now valued at $15 billion, the Silicon Valley company is creating the digital infrastructure needed to bring intelligence to every moving machine on the planet. Applied Intuition services the automotive, defense, trucking, construction, mining and agriculture industries in three core areas: tools and infrastructure, operating systems, and autonomy. Eighteen of the top 20 global automakers, as well as the United States military and its allies, trust the company’s solutions to deliver physical intelligence. Applied Intuition is headquartered in Sunnyvale, California, with offices in Washington, D.C.; San Diego; Ft. Walton Beach, Florida; Ann Arbor, Michigan; London; Stuttgart; Munich; Stockholm; Bangalore; Seoul; and Tokyo. Learn more at applied.co. We are an in-office company, and our expectation is that employees primarily work from their Applied Intuition office 5 days a week. However, we also recognize the importance of flexibility and trust our employees to manage their schedules responsibly. This may include occasional remote work, starting the day with morning meetings from home before heading to the office, or leaving earlier when needed to accommodate family commitments. About The Role We are looking for a software engineer with deep experience in optimizing ML models and deploying them on production-grade embedded runtime environments. You’ll work across the entire ML framework stack (e.g. PyTorch, JAX, ONNX, TensorRT, CUDA, XLA, Triton). At Applied Intuition, You Will
- Drive ML performance optimization on multiple technologies for on-road and off-road ADAS / AD stacks targeting deployment on a variety of embedded compute platforms
- Develop compute usage strategies to optimize efficiency and latency of model inference for compute boards selected by our customers
- Work on model pruning and quantization, and support deployment on memory constrained platforms
- Collaborate closely with ML engineers and software developers on technical efforts to find and optimize efficient model architecture solutions
- Set up methodologies to profile the model performance on target embedded compute platforms and identify performance bottlenecks as part of stack integration
- Bachelors in Electrical Engineering or Computer Science, OR B.Sc. in Computer Science, Mathematics, Physics or a related field
- 3+ years of experience with ML accelerators, GPU, CPU, SoC architecture and micro-architecture
- Strong software development skills with the focus on embedded programming
- Experience profiling and optimizing model performance on embedded compute platforms
- Experience in working with deep learning frameworks (e.g., PyTorch, JAX, ONNX, etc.)
- M.Sc or PhD in a ML related area
- Built an ML optimization framework from scratch before
- Deployed ML solutions to embedded chips for real time robotics applications
Vacancy posted 5 days ago
Similar jobs that could be interesting for youBased on the ML Runtime Optimization Engineer in California vacancy
$166k - $244k
...the Role:We are looking for a Software Engineer, Edge Systems & Runtime to join our team, focusing on high-... ...edge GPUs.Your primary focus will be optimizing our production C++ runtime infrastructure... .../smart memory management.On-Device ML Deployment: 3+ years hands-on...SuggestedFull time$195.2k - $361.2k
...future of AI should belong to the people it servesRole SummaryMake models fast on the hardware people actually own. You optimize inference engines (llama.cpp, vLLM) for constrained local and edge environments — GPU/iGPUs, Vulkan backends — not datacenter H100 environment...SuggestedFull timeInternshipLocal areaImmediate startShift work$120k - $275k
...and software to train and run the largest ML workloads for AGI. MatX is seeking silicon micro-architects and design engineers to join our team as we create best-in-class... ...dynamic and static power reduction.Drive power optimization across compute, memory, interconnect, PCIe,...SuggestedFull timeWork experience placementLocal areaRemote workMonday to FridayFlexible hours$130k - $200k
...programmable using standard high-level programming languages and AI/ML frameworks. This level of efficiency makes perpetual,... ...computing revolution About the Role We are seeking Software Optimization Engineers to join our growing team. Efficient’s Optimization Engineers...SuggestedImmediate start$266.05k - $396k
.... Job Summary Distinguished Engineer - AI Infrastructure We are seeking... ...Engineer with unrivaled depth in AI/ML inferencing at scale and the... ...inference engines (TensorRT, vLLM, ONNX Runtime, Triton), model optimization (quantization, pruning, distillation...SuggestedPart timeWork at officeLocal area$226k - $307k
...autonomous system intelligence.As a Machine Learning and System Optimization Engineer, you will orchestrate and allocate overall system capacity to... ...-constrained vehicle SoCs.In addition, you will optimize ML models, write custom CUDA kernels, and build highly concurrent...Full timeTemporary workRelocation package$184k - $287.5k
We're now looking for a Sr. Inference Engineer, for GPU Kernel Optimization! What does it take to push every LLM inference operation to its performance... ...across kernel execution, compiler decisions, and runtime scheduling.Direct experience with LLM inference frameworks...Full time$184k - $287.5k
...We are seeking an AI Compiler Engineer with deep expertise in... ...implement end-to-end compiler optimization workflows, from feature engineering... ...software engineering and AI/ML experience, preferably in tools... ...measurable outcomes such as runtime gains, compile-time...Full timeRemote work$286.2k - $326.7k
...Senior Distinguished Engineer, AI Compute (Remote Eligible) At Capital... ..., our applications of AI & ML are bringing humanity and... ...the high-scale developer and runtime environments required to build... ...partners across Capital One to help optimize business outcomes while...Full timePart timeRemote work- ...patients worldwide.We’re a team of engineers, clinicians, and innovators... ...alongside research, SW/ HW/ ML engineering, regulatory,... ...algorithms into performance optimized, robust, validated and scalable... ...management, and containerized GPU runtime environments (e.g., Docker,...Local areaWorldwideFlexible hours
$136k - $218.5k
...closely collaborate with HW/ML experts and infrastructure teams... ...power patterns, generate optimized code, and provide actionable... ...structures, algorithms, and software engineering principles.Familiarity with... ..., and comment on their runtime and memory complexities.Desire...Full time$124k - $195.5k
...computing, and low-level hardware optimization has never been more critical.... ...exploratory tools and runtime systems to profile and accelerate... ...Computer Science, Computer Engineering, Electrical Engineering, or related... ...deep learning compilers and ML systems, including graph-...Full time$224k - $356.5k
...Performance Senior Software Engineer to join our energetic team. You... ...be doing:Play a key role in optimizing system software for Nvidia automotive... ...teams to track key boot & runtime performance benchmarks.Ensure... ...software efficiency.AI/ML experience is highly desirable...Full time$218.8k - $335.3k
...with cutting‑edge robotics, optimization, and machine learning to build... ...looking for a Staff Software Engineer to provide technical leadership... ...robustness, and predictable runtime behavior under tight latency... ...mix of analytical models and ML‑based forecasting, including...Full timeLocal areaRemote workWork from homeFlexible hours$124k - $195.5k
NVIDIA is recruiting a Senior Inference Performance Engineer to push NVIDIA's performance limits on large-scale AI inference benchmarks... ...position provides an outstanding opportunity to employ your optimization knowledge in an autonomous optimization framework. AI agents use...Full time$100k
...and looking for contributors of all seniorities.As a Software Engineer on the Metal Runtime team at Tenstorrent, you’ll work on the low-level software that powers our AI accelerators. You’ll build and optimize high-performance runtime systems that execute directly on the...Permanent employment$184k - $287.5k
...best work. Come join the team and see how you can make a lasting impact on the worldWe are looking for an experienced Compiler Optimization Engineer for an exciting role in our Compute Compiler Team. We deliver features and improvements to CUDA and other compute compilers...Full time- ...fusing strategy, consulting, and customer experience with agile engineering and problem-solving creativity. United by our core values and... ...truly value.OverviewWe’re looking for a Senior Adobe Journey Optimizer engineer who can take a business flow from a Miro board to a live...
- ...video generation.We're seeking a GPU Performance Engineer to squeeze every last FLOP from our H100 infrastructure and optimize our model serving stack to its absolute limits.... ...fusion, and GPU utilizationCollaborate with ML engineers to optimize model implementationsDebug...
$120k - $250k
...compiler, and kernels so each layer benefits from the others. The runtime owns the host-side stack and the contracts that bind those... ...and debuggers — perf counters, traces, and the Python surfaces ML engineers actually use — and hit measurable performance targets on...Full timeContract workWork experience placementLocal areaRemote workMonday to FridayFlexible hours$152k - $241.5k
...reinvented itself over two decades, inventing the GPU in 1999 to reshape PC gaming and modern computer graphics. As an engineer in our EDA Workflow Optimization team, you will partner closely with our engineering teams worldwide. You will understand workflows covering the...Full timeWorldwide$151.19k - $214.59k
...to name a few).Responsibilities:Manage the entire ML lifecycle from data collection to deployment and monitoringCollaborate... ...across teams such as DS, QA, Infra and other engineering teams to productionize ML modelsWrite and optimize code for production environments, ensuring the...Work experience placement$148k - $222k
...of customer missionsDevelop and leverage software tooling to optimize launch and orbital trajectories across key customer missionsLead... ...You:An undergraduate or graduate degree (BS/MS/PhD) in Engineering, Computer Science, Physics, or a related field2+ years of relevant...Full timeWork experience placement$166k - $220k
...autonomous software that powers them. As we scale, we are seeking a visionary Air-Vehicle Multidisciplinary Design Analysis and Optimization (MDAO) Engineer to join our fast-growing team. You will be instrumental in driving the rapid development of cutting-edge weapon systems...Full timeWork experience placementImmediate start$160k - $230k
...inference for large language models (LLMs). Our mission is to optimize inference frameworks, algorithms, and infrastructure, pushing the... ....We are seeking anInference Frameworks and Optimization Engineer to design, develop, and optimize distributed inference engines...Full time$150k - $250k
Job Title: Senior Analog Mixed-Signal Engineer - AI HardwareJob Location: Palo Alto, CA, Austin... ...teams, and process/device engineers to optimize performance, power, and area for AI... ...optimization of algorithms and circuits for ML workloads.Strong programming skills in Python...Remote work$140k - $170k
Job Summary:The Sr Engineer, AI & ML will be part of the R&D team at Masimo with focus on design and creation of next-generation health... ...deep learning models, digital signal processing techniques, optimization and numerical modeling in MATLAB, Python or similar softwareResponsible...Work at officeFlexible hours$163k - $400k
...search ranking, two-sided marketplace matching, and dispatch optimization. This is the engine of a local marketplace: learning-to-rank and retrieval... ...prove causal impact in a marketplace setting.Bring modern ML/LLM techniques into ranking and matching where they add real...Full timeLocal areaWork from home- ...THE ROLEAMD is seeking an AI Systems Engineer to help develop and optimize machine learning workloads on next-... ...software, designing high-performance ML operator kernels, optimizing... ...collaborate closely with compiler, runtime, silicon, and architecture teams while...Worldwide
$198k - $326k
...work is centered on trust and optimized for culture, connection,... ...for a Senior Staff Software Engineer with deep expertise at the intersection... ...how models interact with runtimes, compilers, and hardware, and... ...closely with ML, infrastructure, and product...For contractorsWork at officeFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to ML Runtime Optimization Engineer. Be the first to apply!
Related searches
- machine learning engineer California
- computer vision machine learning engineer California
- machine learning research scientist California
- data engineer machine learning California
- internship machine learning California
- machine learning scientist California
- machine learning part time California
- machine learning remote California
- machine learning intern California
- machine learning California


