ML Engineer - Foundation Models Infra (Low Latency)
Apple
Apple’s Foundation Model Services team builds the frameworks, services, and tools that run Apple's largest foundation models in production. You’ll work with product teams to deploy real-time model serving, and you’ll partner with researchers to prototype inference for state-of-the-art architectures. You will create tooling to understand bottlenecks across hardware and use cases. Expect to write high-quality code, move quickly in a fast-evolving field, and grow your impact as the system scales to #J-18808-Ljbffr Apple
- Apple Inc. seeks a Sr. Machine Learning Engineer for Foundation Models Inference in Cloud OS & Inference to advance private, scalable AI inference... ...at planetary scale. You will own high-throughput, low-latency serving, mentor engineers, and collaborate with security...Foundation
- ...better place to do it than Apple. The Foundation Model Services team builds the frameworks, services... ...millions of queries at incredibly low latency while drawing every ounce of performance... ...Docker. 5 year+ industry experience in ML technologies (LLMs, Machine Learning,...Foundation
$222.72k - $389.75k
...abilities, we’ll explore your foundational skills and how you... ...Machine Learning Engineer to lead the technical... ...our Ads Conversion Core Modeling team, building the state... ...-of-the-art applied ML projects for ads conversion... ...prediction with low latency. Mine text, visual, and...FoundationWork at officeLocal areaRelocationRelocation package- Apple Inc. in Seattle is seeking a Software Engineer focused on Machine Learning and AI to... ...production-grade systems that serve real-time model inferences at scale. You will collaborate... ...years shipping production software, 5+ years ML/AI experience, and a BS in CS or related...Foundation
$184.7k - $324.8k
Sr. Machine Learning Engineer, Foundation Models Inference - Cloud OS & Inference Santa Clara, California, United States Machine Learning and... ...iMessage, Photos, Camera, Spotlight & Safari, at remarkably low latency with every ounce of compute extracted from the hardware...FoundationWorldwideRelocation$175k - $280k
...consisting of a variety of LLM, speech, and vision models. Partner with ML infrastructure and training engineers to build a fast, cost‑effective, accurate, and... ...for serving reliably at high throughput, with low latency. Significant systems programming experience; ex...Contract workFlexible hours- ...building the first real-time human foundation model — unifying text, speech, and... ...scalable data pipelines for ML, including preprocessing,... ...and shipping ultra-low-latency ML products to millions. We... ...research with the ruthless engineering needed for consumer-grade, real...FoundationShift work
$141.9k - $190.3k
...global organization of engineers, product developers,... ...advance the technological foundation and consumer media... ...help build services and models enabling efficient ad... ...testable software and apply ML solutions where needed... ...in a high-throughput, low-latency microservices...FoundationWork experience placementWork at office- Xaira Therapeutics is seeking a Senior Software Engineer to join our Platform team to design, build, and deploy the AI infrastructure... ...thousands of GPUs for training and inference of biology foundation models in Seattle. This role covers MLOps for GPU clusters and backend...Foundation
$229k - $343k
...digital services.Snap Engineering teams build fun... ...machine learning models that determine the... ...fairness, safety, latency, and business goalsPartner... ...working on ML ranking... ...monetizationExperience with LLMs, foundation models, semantic... ...building low-latency ML serving...FoundationFull timeLive inWork at officeLocal area- ...Machine Learning Engineer As a Machine Learning Engineer... ...You won't just train models in isolation; you will... ...production scale with low latency. You will work on an... ..., ensuring our core ML services scale... ...Solid Programming Foundations: Strong proficiency in...Foundation
$190.2k - $345.65k
...image, video, and 3D models built on each customer... ...Staff Machine Learning Engineer to architect and lead... ...labels) into structured, low-latency, searchable... ...loops that improve them.ML Engineering leadership... ...Strong data-engineering foundations — large-scale batch and...FoundationFull timeTemporary workLocal areaWorldwide- ...Sr Machine Learning Engineer Technology is at the... ...advance the technological foundation and consumer media... ...help build services and models enabling efficient ad... ...testable software and apply ML solutions where needed... ...in a high-throughput, low-latency microservices...FoundationWork experience placementWork at office
- Uber is seeking a Senior Software Engineer for ML Data & Backend in Seattle. You will architect mission-critical ML infrastructure, from low-latency data pipelines to scalable backend systems powering trust across our marketplace. You’ll design and implement long-lasting...
$150k - $300k
...Senior Staff ML Engineer At GEICO, we offer a rewarding career where... ...search capabilities as their foundation. You bring a passion for... ...evaluation systems for AI/ML models and LLMs used in production systems... ...workflows via both no code/low code and traditional high-code...FoundationHourly payWork experience placementLocal area- ...for a Machine Learning Engineer who will deliver... ...problem discovery through model deployment and monitoring... ...and operating ML-powered features that... ...SageMaker) for high-traffic, low-latency, large-data applications... .... Understanding of foundation models and the open-source...FoundationRemote work
- ...compilation technologies for AI foundation models. Responsibilities As an... ...performance, including latency, throughput, and system stability... ...with researchers and engineers to translate model requirements... ...pursuing long-term work in ML systems or AI infrastructure...FoundationInternship
$232.56k - $427.5k
Research Engineer - LLM/VLM Inference Optimization (Seed Infra) Location: Seattle Team: Technology... ...state-of-the-art model inference engines... ...development, low‑precision... ...and Python; solid foundations in algorithms, data... ...demonstrated impact on latency, throughput, or serving...FoundationTemporary workLocal area- ...training systems (e.g., data/model/pipeline parallelism,... ...Improve inference performance, latency, and throughput for foundation models Develop compiler... ...Science, Electrical Engineering, or related technical fields... ...Experience working on large-scale ML systems or infrastructure...FoundationInternship
- AI/ML Engineer**** Please note: This role is not eligible for 100% remote... ...of experience implementing models or machine learning... ...stakeholders.* Solid technical foundation, combined with constant curiosity... ...accuracy, groundedness, safety, latency, cost, and drift.* Experience...FoundationTemporary workWork at officeLocal area3 days per week
$202k
Senior Software Engineer - ML Infra About the Role & Team Engineering at Uber... ...systems that power the foundation of trust across our global... ...problems at the intersection of low latency and high correctness, often... ...key frameworks while role-modeling coding best practices and...FoundationFull timeWork at officeRemote work$266.72k - $350.07k
...delivery of a trusted unified data foundation, AI-driven data analytics and... ...As a Principal AI/ML Engineer, you will define and drive implementation... ...frameworks and tools for model development, training... ...offs involving model quality, latency, reliability, scalability, cost...FoundationPermanent employmentFull timePart timeWork visa$184.7k - $324.8k
LLM Machine Learning Engineer, Models and Agent Science, AIML Cupertino, California, United States... ...for on-device and server-based Apple Foundation Models and Apple Intelligence features.... ...field Publication record at top AI/ML venues Experience with post-training LLMs...FoundationRelocation$198.36k - $416.1k
...seeking a Machine Learning Platform Engineer to develop and maintain our... ...supports deep learning models for code development, testing,... ...learning platform and serves as a foundation for recommendation, advertising... ...developing key components of ML infrastructure and mentoring interns...FoundationFull timeTemporary workLocal areaImmediate start$90.1k - $191.8k
...world! The Data Labeling Engineering team designs, builds,... ...machine learning models across General Motors... ...data engineering , and ML , defining labeling strategies... ...labeling solutions at foundation‑model scale . We... ...hit quality, cost, and latency goals. Champion AI‑assisted...FoundationWork experience placementLocal areaWork from homeFlexible hours- Pangleglobal is seeking a Student Researcher in Seattle to conduct research on infrastructure for AI foundation models. This role requires pursuing a PhD in computer science and strong programming skills, focusing on efficiency and reliability in large-scale systems. Interns...FoundationInternship
- ...identifying algorithms for risk/violation/low-quality issues in e-commerce scenarios... ...CoT, alignment, and other work for large models in the e-commerce domain, aiming for ultimate... ...proficient in Python, Go, or C++- Solid foundation in data structures/algorithms, proficient...FoundationOverseas
$228.7k - $306.7k
...global organization of engineers, product developers,... ...advance the technological foundation and consumer media... ...unblock and guide our ML and Research teams to... ...maintainable, and testable models and pipelines.Daily,... ...in a high throughput, low latency environment.Mentoring...Foundation$228.7k - $306.7k
...Principal Machine Learning Engineer, Ad PlatformsReq ID:10... ...the technological foundation and consumer media... ...unblock and guide our ML and Research teams to... ...maintainable, and testable models and pipelines.Daily,... ...in a high throughput, low latency environment.Mentoring...FoundationFull time$185.6k - $255k
...not a feature; it’s the foundation everything else stands on... ...without waiting on an engineering queue. The bet underneath... ...you.The RoleAt Amperity, ML Engineers work in small,... ..., and predictive models at scale.Improve model inference latency to deliver predictions that...FoundationWork at officeLocal areaRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to ML Engineer - Foundation Models Infra (Low Latency). Be the first to apply!
- data scientist machine learning engineer Seattle, WA
- machine learning ai engineer Seattle, WA
- computer vision machine learning engineer Seattle, WA
- machine learning engineer Seattle, WA
- ai ml engineer Seattle, WA
- graduate machine learning engineer Seattle, WA
- machine learning software engineer Seattle, WA
- senior ml engineer Seattle, WA
- foundation manager Seattle, WA
- foundation program officer Seattle, WA

