Distributed LLM Inference Engineer (San Francisco)
$170k - $245kAnyscale
About AnyscaleAt Anyscale, we're on a mission to democratize distributed computing and make it accessible to software developers of all skill levels. We’re commercializing Ray, a popular open-source project that's creating an ecosystem of libraries for scalable machine learning. Companies like OpenAI, Uber, Spotify, Instacart, Cruise, and many more, have Ray in their tech stacks to accelerate the progress of AI applications out into the real world.With Anyscale, we’re building the best place to run Ray, so that any developer or data scientist can scale an ML application from their laptop to the cluster without needing to be a distributed systems expert.Proud to be backed by Andreessen Horowitz, NEA, and Addition with $250+ million raised to date.About the roleAs a Distributed LLM Inference Engineer, you will help systems and optimizations that push the boundaries of performance for inference at large scale. This is an incredibly critical role to Anyscale as it allows us to achieve a market leading position for AI infrastructure.As part of this role, you willIterate very quickly with product teams to ship the end to end solutions for Batch and Online inference at high scale which will be used by open-source Ray users and customers of AnyscaleWork across the stack integrating Ray Data and LLM engine providing optimizations achieving low cost solutions for large scale ML inference Integrate with Open source software like vLLM, work closely with the community to adopt these techniques in Anyscale solutions, and also contribute improvements to open sourceFollow the latest state-of-the-art in the open source and the research community, implementing and extending best practices We'd love to hear from you if you haveFamiliarity with running ML inference at large scale with high throughput and low latencyFamiliarity with deep learning and deep learning frameworks (e.g. PyTorch)Solid understanding of distributed systems, ML inference challengesBonus points!ML Systems knowledgeExperience using Ray Work closely with community on LLM engines like vLLM, TensorRT-LLMContributions to deep learning frameworks (PyTorch, TensorFlow)Contributions to deep learning compilers (Triton, TVM, MLIR)Prior experience working on GPUs / CUDACompensationAt Anyscale, we take a market-based approach to compensation. We are data-driven, transparent, and consistent. As the market data changes over time, the target salary for this role may be adjusted. This role is also eligible to participate in Anyscale's Equity and Benefits offerings, including the following:Stock OptionsHealthcare plans, with premiums covered by Anyscale at 99% for both employees and dependents401k Retirement PlanEducation & Wellbeing StipendPaid Parental LeaveFertility BenefitsPaid Time OffCommute reimbursement100% of in-office meals coveredAnyscale Inc. is an Equal Opportunity Employer. Candidates are evaluated without regard to age, race, color, religion, sex, disability, national origin, sexual orientation, veteran status, or any other characteristic protected by federal or state law.Anyscale Inc. is an E-Verify company and you may review the Notice of E-Verify Participation and the Right to Work posters in English and SpanishCompensation Range: $170K - $245KLocationSan Francisco; Palo AltoEmployment TypeFull timeLocation TypeHybridDepartmentEngineeringCompensationTarget Base Salary:$170K – $245K • Offers EquityAt Anyscale, we take a market-based approach to compensation. We are data-driven, transparent, and consistent. As the market data changes over time, the target salary for this role may be adjusted.This role is also eligible to participate in Anyscale's Equity and Benefits offerings, including the following:Stock OptionsHealthcare plans, with premiums covered by Anyscale at 99%401k Retirement PlanWellness & Education StipendPaid Parental LeaveFertility BenefitsPaid Time OffCommute reimbursement100% of in office meals covered
$227.2k - $417k
...the Role:As a Software Engineer on the ML... ...class machine learning inference platforms. These platforms... ...support Deep Learning, LLM, and Search models. This... ...throughput, and low latency distributed systems using... ...fans. Headquartered in San Francisco and founded in 2014, Tubi...SuggestedFull timeTemporary workPart timeLocal areaFlexible hours$172.5k - $260.1k
...immediate opportunities for Lead software engineers who want their lines of code to have... ...and exciting components/frameworks in distributed filesystems in an ever-growing and evolving... ...found at the following link: to the San Francisco Fair Chance Ordinance and the Los...SuggestedFull timePart timeImmediate start$132.1k - $258.4k
...in 2020 with office hubs in San Francisco, New York City, Seattle, Austin... ...? As a software quality engineer at WRITER, you'll play a critical... ...plans for our AI agents and LLM-powered applications,... ...strong focus on testing complex distributed systems or AI/ML...SuggestedFull timePart timeWork at officeLocal area$36.06 - $40.87 per hour
...Technical Support Field Engineer - San Francisco, CA Dentsply Sirona is the world’s largest manufacturer of professional dental products and... ...references, certifications, transcripts and languages spoken); and inferences from personal information collected (e.g., a profile...SuggestedHourly payWork experience placementWork at officeRemote workWorldwideFlexible hoursNight shift$190k - $252.5k
...Job Description Staff BESS Electrical Design Engineer (EPC) — Architect the Grid's Battery Future San Francisco, CA | Full-Time | $190,000 – $252,500 The Opportunity... ...Push the frontier: Lead R&D into MV-DC distribution for next-generation storage integration...SuggestedFull timeShift work$190k - $252.5k
...Description Job Description Staff Electrical Design Engineer (EPC) — Design It All, Own It All San Francisco, CA | Full-Time | $190,000 – $252,500 The... ...across the board — MV feeds up to 34.5 kV, LV distribution, lighting, raceway, grounding, communications...Full time$190k - $252.5k
...Description Staff High Voltage Electrical Engineer (EPC) — Transmission-Scale Design, Mission-Scale Impact San Francisco, CA | Full-Time | $190,000 – $252,500 The... ...Innovation on the side: Lead R&D into MV-DC distribution for next-generation BESS integration...Full time- ...Salesforce is seeking a senior engineering candidate to join the Site Reliability organization in San Francisco. Working closely with... ...platforms.Strong knowledge of distributed systems and Linux/Unix internals... ..., predictive analysis, LLM-based automation, and prompt...Part timeWorldwideWeekend work
$206.4k - $379.1k
...seeking a Principal Service Engineer to serve as the technical lead... ...products.Design and architect inference infrastructure for enterprise... ...expertise with Kubernetes, distributed systems, and MLOps platforms.... ...civil liability.SummaryLocation: San Jose; Seattle; San...Full timeTemporary workPart timeLocal areaWorldwide- ...powers mission-critical inference for the world's most... ...help build the platform engineers turn to to ship AI... ...operating system for distributed, heterogeneous AI hardware... .... We believe that as LLM and multi-modal workloads... ...requirements of the San Francisco Fair Chance Ordinance,...Full timeFlexible hours
$300 per month
...Senior or Staff Hardware Systems Engineer to strengthen Crusoe’s... ...studies across training and inference - dense, MoE, long-context,... ...workloads.Hands-on experience with distributed training and/or inference... ...or regulation.LocationSan Francisco, CA - US; Sunnyvale, CA - USEmployment...Temporary workPart time$211.7k - $302.4k
...advantage.We're looking for an Automation Engineering Managerto lead the team responsible... ...a hybrid role based out of either our San Francisco or Toronto office. You must be willing... ...tech industry.Experience with distributed system testing or large-scale performance...Full timeTemporary workPart timeWork at officeLocal areaImmediate startFlexible hours2 days per week- ...SoFi’s SeniorStaff AI Engineer is a hands-on AI engineering... ...frameworks and LLM deployment patterns.Advanced... ...problems at scale.Distributed Agent Memory & State:... ...throughput, low-latency inference across diverse... ...characteristics.Pursuant to the San Francisco Fair Chance Ordinance,...Part timeRemote work
$295k - $405.5k
...Machine Learning Platform Engineer, you will own the... ...platform including training, inference, feature management,... ...background in distributed systems, ML infrastructure... ...authorityExperience integrating LLM workflows into... ...have headquarters in San Francisco and Kitchener-Waterloo...Part timeWork experience placementWork at officeLocal areaRemote workMonday to FridayFlexible hours3 days per week$117.2k - $313.7k
...the future of AI, and you are the future of Salesforce.Distributed Systems Software Engineer - Public Cloud (Senior/Lead/Principal) Note: By... ...benefits can be found at the following link: to the San Francisco Fair Chance Ordinance and the Los Angeles Fair Chance...Full timePart time$243k - $284k
...the world.The RoleWe're hiring a Staff Engineer, Security Automation to anchor automation... ...native role. You'll build with and on AI: LLM-augmented workflows, agentic automations... ...in-office presence 2 days a week in our San Francisco, CA office.To join our team, you should...Part timeWork at office2 days per week$148.5k - $223.9k
...infusing AI into every aspect of our software engineering. While AI is writing the code,... ...automation plans, focusing heavily on validating LLM integrations, prompt accuracy, and... ...be found at the following link: to the San Francisco Fair Chance Ordinance and the Los Angeles...Full timePart time$197.3k - $313.7k
...team is building a highly scalable and distributed load balancing and gateway service to front... ...to add experienced distributed systems engineers who are passionate, hungry for new... ...be found at the following link: to the San Francisco Fair Chance Ordinance and the Los Angeles...Full timePart time$192k - $240k
...support you need to grow your career.Engineering at BrexEngineering at Brex is about building... ...ll workThis role will be based in our San Francisco office. We are a hybrid environment... ...languages Experience with securing distributed systems in AWS, cloud and Kubernetes environmentsContributions...Part timeWork at officeRemote workWork from home$250k - $285k
...Staff Product Security Engineer with deep AI/ML... ..., infrastructure, and distributed AI systems. This is a... ...end-to-end, including LLM pipelines, vector databases... ...stack, including MLOps, inference architectures, vector... ...regulation.LocationSan Francisco, CA - USEmployment TypeFull...Temporary workPart time$180k - $247k
....The Staff Product Security Engineer OpportunityThe Security team... ..., orchestration layers, and LLM-integrated services.Deliver... ...testing, applied to complex distributed systems.Knowledge of authentication... ...candidates located in the San Francisco Bay area is between: $180,00...Part timeLocal areaWorldwideFlexible hours$190k - $230k
...’re seeking a Senior Systems Engineer to play a key role in executing... ...)Experience working with LLM APIs (OpenAI, Anthropic, Google... ...engineeringExperience designing scalable, distributed systems in cloud environments... ...or regulation.LocationSan Francisco, CA - USEmployment TypeFull...Temporary workPart time$180k - $230k
...San Francisco, CAProduct Systems Engineering – Product Systems /Full time /On-siteWanna join the adventure?Loft Orbital builds a space infrastructure... ...responsibilitiesStrong understanding of software-heavy systems, including distributed components, APIs, processing chains, and integration...Full timeTemporary workPart time$220k - $235k
...future of our cloud platform and champion engineering excellence across Ironclad. In this... ...resilience, and solving complex distributed systems at scale.Roles & Responsibilities... ...range for this position based at our San Francisco headquarters. The actual base salary offered...Full timeContract workPart timeWork at office$194k - $267k
...technical StaffObservabilitySite Reliability Engineer with a specialty in Splunk to own and... ...agents and collectors across complex distributed systems.Key ResponsibilitiesAutomated... ...must attend in person onboarding in our San Francisco office the first week of employment....Permanent employmentPart timeWork at officeLocal areaWorldwideFlexible hours$153k - $376k
...infrastructure is at the heart of everything we build. As a Software Engineer on our Infrastructure team, you’ll help design, build, and... .... We’re scaling fast, and we’re looking for experienced distributed systems engineers across a variety of teams. Whether you’re passionate...Minimum wageFull timePart timeLocal areaRemote workWorldwideFlexible hours$10 per hour
...Autonomous Freight Systems team is a brand new, AI-first engineering team in San Francisco, building Flexport's client-facing rates platform and self... ...familiarity to hands-on implementation of agent workflows and LLM-backed automation in production.Drive technical design for...Part timeFlexible hours$227.2k - $324.5k
...About the Role:Site Reliability Engineering (SRE) at Tubi is not a traditional operations... ...of building and running large-scale, distributed systems. Our mission is to engineer resilience... ...to Los Angeles, New York City, and San Francisco$227,200—$324,500 USDTubi is a division...Full timeContract workTemporary workPart timeLocal areaFlexible hours$195k - $225k
...seeking a highly motivated Electrical Design Engineer to act as the primary technical bridge... ...You will ensure that large-scale power distribution systems are designed for safety,... ...in office with options to be located in San Francisco, Denver, or Dallas. Travel will vary, but...Temporary workPart timeWork at office- ...RoleThe Director of Platform & Reliability Engineering will lead a critical engineering... ...requires 2-3 days a week in office in San Francisco, CA.ResponsibilitiesThis leader will be... ...practices.Strong technical judgment in distributed systems, production operations,...Part timeWork at officeLocal area2 days per week3 days per week
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Distributed LLM Inference Engineer (San Francisco). Be the first to apply!


















