ML Engineer, Inference & Optimization (Palo Alto)
$250k - $350kPika
About the RoleWe are seeking Senior/Staff level Inference Engineers to accelerate the performance of Pika's AI-driven products. In this highly technical role, you will operate at the intersection of cutting-edge inference acceleration, GPU parallelism, advanced model deployment, and video generation technologies. Your expertise will drive significant improvements to model speed and efficiency, ensuring our creative AI systems deliver industry-leading user experiences at scale.You will design and optimize inference pipelines, implement state-of-the-art acceleration techniques, and work closely with researchers and engineers across the team to push the boundaries of what’s possible in real-time AI deployment. Your efforts will play a foundational role in powering the next generation of Pika’s video and language models.What You’ll DoAccelerate Inference: Lead and implement advanced inference acceleration techniques, including attention optimization and quantization for efficient model serving.Maximize GPU Parallelism: Engineer and optimize GPU strategies across tensor, sequence, and pipeline parallelism (TP, SP, PP) for maximal efficiency and scalability.Programming for Performance: Develop and optimize high-performance computing kernels and distributed workloads using CUDA and NCCL.Advance AI Deployment: Collaborate with research and engineering teams to bring state-of-the-art videogen and large language models into production.Improve Training Efficiency: (Bonus) Contribute to improvements in model training speed, stability, and resource utilization as part of our deployment lifecycle.Technical Excellence: Drive rigorous code reviews, participate in technical discussions, and mentor fellow engineers on best practices in inference and GPU programming.What We’re Looking ForExperience: 5+ years engineering experience, with a strong track record in inference acceleration and model deployment at scale.Inference Mastery: Proven expertise in inference optimization, including quantization, attention acceleration, and deep learning compiler stacks.GPU & Parallelism: Deep knowledge of GPU programming (CUDA, NCCL) and experience with SP, TP, PP, and other forms of parallelism for distributed inference.AI Domain Knowledge: Familiarity with video generation (videogen) models and large language models (LLMs).Collaboration: Strong cross-discipline communication skills; able to drive shared goals across research and engineering functions.Ownership Mindset: Self-driven, solutions-oriented, and capable of managing ambiguity in a fast-paced startup environment.Bonus: Experience in enhancing training efficiency, stability, or resource optimization for large models.Nice to HaveExperience with high-throughput video or real-time streaming model deploymentFamiliarity with distributed training and optimization toolkitsContributions to open source projects in AI infrastructure or deep learning compilersStartup or rapid prototyping experienceWhat We OfferCompetitive salary in the AI industryEquity in a fast-growing startup shaping the future of AIComprehensive health benefits, monthly stipends, company retreatsA supportive and collaborative office culture—we’re all building and launching togetherAbout PikaAt Pika, we're crafting a future where video creation is seamless, intuitive, and universally accessible. Our mission is to empower creativity by breaking down technical barriers using the transformative power of AI. We’re a tight-knit, energetic team based in Palo Alto, CA, valuing efficiency, curiosity, and the ambition to make a meaningful impact on the world.We work from our Palo Alto office 3–5 days a week and welcome applicants who are eager to contribute onsite.Compensation Range: $250K - $350KLocationPalo Alto HQEmployment TypeFull timeLocation TypeOn-siteDepartmentResearchCompensationbase salary $250K – $350K
$209k - $313k
...digital services.Snap Engineering teams build fun and technically... ...causal impact, optimize decision-making, and... ...understanding of causal inference and modern approaches... ...and leveraging causal ML in production... ...Francisco, California; Palo Alto, California; New York,...SuggestedFull timePart timeLive inWork at officeLocal area$230k - $260k
...a Principal Machine Learning Engineer, you will operate at the company... ...the design of large-scale ML systems and shared platforms... ...platforms (training, evaluation, inference, safety) used across multiple... ...a hybrid role based in our Palo Alto HQ. We collaborate in-office...SuggestedPart timeWork at officeImmediate start3 days per week$222.72k - $389.75k
...building a new Programmatic Ads ML team to bring in exchange-... ...We’re looking for a Staff ML engineer to develop core bidding and ranking systems that help us optimally buy and sell inventory across... ...following offices: San Francisco, Palo Alto, Seattle.#LI-SM4At Pinterest we...SuggestedPart timeWork at officeLocal areaRelocationRelocation package$190k - $234k
...What You’ll Do As a StaffMachine Learning Engineer/Applied Scientist, you will be... ...fieldExperience with building and evolving ML Training and Inferencing systems at significant... ...applicationsLocationThis is a hybrid role based in our Palo Alto HQ. We collaborate in-office 3 days a...SuggestedPart timeWork at officeLocal area3 days per week$276k - $414k
...other digital services.Snap Engineering teams build fun and technically... ...Engineer to join the Content ML team at Snap! We build large-... ...design, train, deploy, and optimize state-of-the-art machine learning... ...; San Francisco, California; Palo Alto, California; New York, New...SuggestedFull timePart timeLive inWork at officeLocal area$229k - $343k
...other digital services.Snap Engineering teams build fun and technically... ...mentor engineers working on ML ranking systemsStay current with... ..., embeddings, deep learning, optimization, evaluation, and... ...form of RSUs.SummaryLocation: Palo Alto, California; Seattle, Washington...Full timePart timeLive inWork at officeLocal area$185k - $225k
...building a data-driven decision engine that powers every aspect of... ...feature success and optimizing user experiences, your work will... ...Location: This role is based in Palo Alto, CA and involves a hybrid work... ...testing, cohort analysis, causal inference).Excellent communication...Part timeWork at officeRemote work- ...democratize AI through high-performance, optimized, open-source and cutting-edge models,... ...Mistral AI is seeking a Applied AI Engineer to facilitate the adoption of its products... ...open source codebases for tasks such as inference and fine-tuning. • You’ll be involved...Full timeWork at officeVisa sponsorship
$222.72k - $389.75k
...for a Staff Machine Learning Engineer to lead the technical vision for... ...of state-of-the-art applied ML projects for ads conversion. Design... ...understand intention and infer interests from online activity... ...following offices: San Francisco, Palo Alto, Seattle.#LI-HYBRID #LI-SM4At...Part timeWork at officeLocal areaRelocationRelocation package- ...Rubrik's Semantic AI Governance Engine, which is the first system... ...lead even further.As an Applied ML Engineer on the SAGE team,... ...supervised fine-tuning, preference optimization (DPO/RLAIF), and distillation... ...Model Serving and Inference Infrastructure (25% of time)Designing...Permanent employmentPart time
$280.15k
...looking for a Distinguished Engineer to set the technical... ..., raising the bar on ML and systems excellence... ..., multi-objective optimization) that materially improve... ..., and causal inference for discovery and generative... ...following offices, [SF, Palo Alto, Seattle]. #LI-REMOTE#LI...Temporary workPart timeWork at officeLocal areaRemote workRelocationRelocation package$227.87k
...equivalent experience.7+ years of industry experience.Strong software engineering and mathematical skills with knowledge of statistical methods.... ...distance from one of the following offices: San Francisco, Palo Alto, Seattle.Relocation Statement:This position is not eligible...Temporary workPart timeWork at officeLocal areaRelocationRelocation package- ...Python or Golang Software developerLocation: Palo Alto, CA Hybrid ( 3 to 4 days)Exp: 10+... ...CI/CD, terraform and AISenior Software Engineer - Enterprise AIEnterprise AI - Global Infrastructure... ...orchestration. Design, develop, and optimize distributed services and cloud-native...Part time
- ...We are looking for a Senior MLOps engineer to work closely with Data Scientists to build and deploy ML models on a modern MLOps stack. As Lead... ...for high-throughput, real-time inference as well as batch inference, ensuring optimal performance and reliability. Implement...Part time
- ...collaboration and a strong team culture, this role is expected to be in our Palo Alto office five days a week, unless otherwise specified. About the Role We're seeking an experienced LLM Inference Engineer to optimize our large language model (LLM) serving infrastructure. The ideal...Work at office
- ...Palo Alto, CA | Full-Time | On-site About Nace AI: Nace... ...: As a Senior MLOps Engineer, you will own the infrastructure... ...at the intersection of ML engineering, LLM inference infrastructure, and... ...scheduling, utilization, cost optimization) across cloud and on-prem...Full time
$193.3k - $261.5k
...intuitive ways of interacting with data, and we’re looking for top engineers to build them from the ground up.This is a hands-on position... ...paid time off, and parental leave. Learn more about our benefits at .USA, CA, East Palo Alto - 193,300.00 - 261,500.00 USD annually...Part timeInternshipLocal areaFlexible hours- ...About the Role We’re looking for an Applied ML Engineer to design, evaluate, and scale recommendation and ranking systems... ...skips, conversions) and sparse explicit feedback. Optimize systems for real-time inference, scalability, and robustness under non-stationary user...Full time
$154.1k - $267.15k
...Reference: 733479BRPosted: 2026-07-07Location: Palo Alto, CaliforniaSalary: $154,100 - $267,145... ...Senior Staff Software/System Integrator Engineer. The selected candidate will take on a... ...posting date in order to receive optimal consideration.At Lockheed Martin, we use...Full timeTemporary workPart timeWork experience placementWork at officeRemote workFlexible hours- ...Palo Alto, CASales /Full-time /Job Description: Sr. Sales EngineerEssential Functions:· Collaborate with our sales teams and partners... ...with Security Assessment Reports.· Work with Corporate PM and Engineering to capture latest product information and details to be utilized...Full timePart timeWork experience placement
- ...how businesses learn from and optimize in‑person customer... ...and deploy production‑grade ML systems with end‑to‑end ownership... ...model training, deployment, inference, and monitoring in production... ...professional experience in ML engineering. Strong programming skills in...Full time
$65 - $90 per hour
...DescriptionKforce's client, a leading automotive engineering organization in Palo Alto, CA is expanding its Hardware-in-the-Loop (HIL) validation capability and is actively seeking experienced HIL Engineers to support the buildout of multiple new systems.Summary:This role...Part timeImmediate startRemote work$145k - $165k
...network economics, AI and ML, online and real-world... ..., Growth, and Revenue optimization. Our mission is to... ...: Machine Learning Engineers (this role) who focus... ...three times per week in Palo Alto, California. In... ...recommendation systems or casual inference Familiarity with...Full timeWork experience placementCasual workWork at officeFlexible hours$170.5k - $315.49k
## Inference Optimization Engineer (local / edge runtime)Applylocations: US, California, Santa Clara: US, Oregon, Hillsboro: US, California, Folsom: US, Arizona, Phoenixtime type: Full timeposted on: Posted Yesterdayjob requisition id: JR0284871# **Job Details:**## Job...InternshipLocal areaImmediate startShift work$70 per hour
...Hiring: Palo Alto Firewall Engineer – Oil & Gas Domain (Onsite – USA) Location: Onsite – USA Job Type: Full-Time / Contract We are looking for an experienced Palo Alto Firewall Engineer with strong expertise in Palo Alto NGFW, Panorama, OT/IT security, and Oil...Full timeContract work$175k - $275k
...datasets, or full-cycle data engineering, Abaka AI provides the foundation... ...Abaka builds, trains, and optimizes multimodal AI systems. You... ...applied machine learning or ML engineering, with a demonstrated... ...scale distributed training and inference systems. ~ Familiarity with...Full timeImmediate startFlexible hours$208k - $244k
Join to apply for the Staff ML Engineer role at Grindr Join to apply for the Staff ML Engineer role at Grindr Get AI-powered advice on this... ...000.00/yr This is a hybrid role based in our San Francisco or Palo Alto offices (Palo Alto preferred) and will require you to be in...Full timeCasual workWork at officeImmediate startFlexible hours- ...world.Role OverviewAs our Staff Software Engineer, ML infra Engineer for Search & Discovery... ...Discovery organization is responsible for optimizing customers' navigation experience and... ...ML based ranking system and online ML inference servicesBuild strong cross-functional partnerships...Temporary workPart time
$177.19k - $364.8k
...here.We’re looking for a Staff Software Engineer to help build the next generation of Pinterest... ...partner closely with teams across data, ML/AI, analytics, and infrastructure to... ...scale.What you’ll do:Design, implement, and optimize Pinterest’s exabyte-scale data lake...Part timeWork at officeLocal areaRelocationRelocation package$229k - $343k
...addition to Bitmoji, Saturn, and other digital services.Snap Engineering teams build fun and technically sophisticated products that reach... ...eligible for equity in the form of RSUs.SummaryLocation: Santa Monica - 3100 Ocean Park Blvd; Palo Alto, CaliforniaType: Full time...Full timePart timeLive inWork at officeLocal area
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to ML Engineer, Inference & Optimization (Palo Alto). Be the first to apply!
- computer vision machine learning engineer Palo Alto, CA
- machine learning engineer Palo Alto, CA
- internship machine learning Palo Alto, CA
- machine learning research scientist Palo Alto, CA
- machine learning Palo Alto, CA
- machine learning remote Palo Alto, CA
- machine learning part time Palo Alto, CA
- machine learning intern Palo Alto, CA
- machine learning scientist Palo Alto, CA
- data engineer machine learning Palo Alto, CA














