Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

ML Engineer, Inference & Optimization (Palo Alto)

$250k - $350k
Part-time

Pika

About the RoleWe are seeking Senior/Staff level Inference Engineers to accelerate the performance of Pika's AI-driven products. In this highly technical role, you will operate at the intersection of cutting-edge inference acceleration, GPU parallelism, advanced model deployment, and video generation technologies. Your expertise will drive significant improvements to model speed and efficiency, ensuring our creative AI systems deliver industry-leading user experiences at scale.You will design and optimize inference pipelines, implement state-of-the-art acceleration techniques, and work closely with researchers and engineers across the team to push the boundaries of what’s possible in real-time AI deployment. Your efforts will play a foundational role in powering the next generation of Pika’s video and language models.What You’ll DoAccelerate Inference: Lead and implement advanced inference acceleration techniques, including attention optimization and quantization for efficient model serving.Maximize GPU Parallelism: Engineer and optimize GPU strategies across tensor, sequence, and pipeline parallelism (TP, SP, PP) for maximal efficiency and scalability.Programming for Performance: Develop and optimize high-performance computing kernels and distributed workloads using CUDA and NCCL.Advance AI Deployment: Collaborate with research and engineering teams to bring state-of-the-art videogen and large language models into production.Improve Training Efficiency: (Bonus) Contribute to improvements in model training speed, stability, and resource utilization as part of our deployment lifecycle.Technical Excellence: Drive rigorous code reviews, participate in technical discussions, and mentor fellow engineers on best practices in inference and GPU programming.What We’re Looking ForExperience: 5+ years engineering experience, with a strong track record in inference acceleration and model deployment at scale.Inference Mastery: Proven expertise in inference optimization, including quantization, attention acceleration, and deep learning compiler stacks.GPU & Parallelism: Deep knowledge of GPU programming (CUDA, NCCL) and experience with SP, TP, PP, and other forms of parallelism for distributed inference.AI Domain Knowledge: Familiarity with video generation (videogen) models and large language models (LLMs).Collaboration: Strong cross-discipline communication skills; able to drive shared goals across research and engineering functions.Ownership Mindset: Self-driven, solutions-oriented, and capable of managing ambiguity in a fast-paced startup environment.Bonus: Experience in enhancing training efficiency, stability, or resource optimization for large models.Nice to HaveExperience with high-throughput video or real-time streaming model deploymentFamiliarity with distributed training and optimization toolkitsContributions to open source projects in AI infrastructure or deep learning compilersStartup or rapid prototyping experienceWhat We OfferCompetitive salary in the AI industryEquity in a fast-growing startup shaping the future of AIComprehensive health benefits, monthly stipends, company retreatsA supportive and collaborative office culture—we’re all building and launching togetherAbout PikaAt Pika, we're crafting a future where video creation is seamless, intuitive, and universally accessible. Our mission is to empower creativity by breaking down technical barriers using the transformative power of AI. We’re a tight-knit, energetic team based in Palo Alto, CA, valuing efficiency, curiosity, and the ambition to make a meaningful impact on the world.We work from our Palo Alto office 3–5 days a week and welcome applicants who are eager to contribute onsite.Compensation Range: $250K - $350KLocationPalo Alto HQEmployment TypeFull timeLocation TypeOn-siteDepartmentResearchCompensationbase salary $250K – $350K

Vacancy posted 10 hours ago
Similar jobs that could be interesting for youBased on the ML Engineer, Inference & Optimization (Palo Alto) in Palo Alto, CA vacancy
  • $209k - $313k

     ...digital services.Snap Engineering teams build fun and technically...  ...causal impact, optimize decision-making, and...  ...understanding of causal inference and modern approaches...  ...and leveraging causal ML in production...  ...Francisco, California; Palo Alto, California; New York,... 
    Suggested
    Full time
    Part time
    Live in
    Work at office
    Local area

    Snap

    Palo Alto, CA
    10 hours ago
  • $230k - $260k

     ...a Principal Machine Learning Engineer, you will operate at the company...  ...the design of large-scale ML systems and shared platforms...  ...platforms (training, evaluation, inference, safety) used across multiple...  ...a hybrid role based in our Palo Alto HQ. We collaborate in-office... 
    Suggested
    Part time
    Work at office
    Immediate start
    3 days per week

    Typeface

    Palo Alto, CA
    10 hours ago
  • $222.72k - $389.75k

     ...building a new Programmatic Ads ML team to bring in exchange-...  ...We’re looking for a Staff ML engineer to develop core bidding and ranking systems that help us optimally buy and sell inventory across...  ...following offices: San Francisco, Palo Alto, Seattle.#LI-SM4At Pinterest we... 
    Suggested
    Part time
    Work at office
    Local area
    Relocation
    Relocation package

    Pinterest

    Palo Alto, CA
    15 hours ago
  • $190k - $234k

     ...What You’ll Do As a StaffMachine Learning Engineer/Applied Scientist, you will be...  ...fieldExperience with building and evolving ML Training and Inferencing systems at significant...  ...applicationsLocationThis is a hybrid role based in our Palo Alto HQ. We collaborate in-office 3 days a... 
    Suggested
    Part time
    Work at office
    Local area
    3 days per week

    Typeface

    Palo Alto, CA
    10 hours ago
  • $276k - $414k

     ...other digital services.Snap Engineering teams build fun and technically...  ...Engineer to join the Content ML team at Snap! We build large-...  ...design, train, deploy, and optimize state-of-the-art machine learning...  ...; San Francisco, California; Palo Alto, California; New York, New... 
    Suggested
    Full time
    Part time
    Live in
    Work at office
    Local area

    Snap

    Palo Alto, CA
    15 hours ago
  • $229k - $343k

     ...other digital services.Snap Engineering teams build fun and technically...  ...mentor engineers working on ML ranking systemsStay current with...  ..., embeddings, deep learning, optimization, evaluation, and...  ...form of RSUs.SummaryLocation: Palo Alto, California; Seattle, Washington... 
    Full time
    Part time
    Live in
    Work at office
    Local area

    Snap

    Palo Alto, CA
    10 hours ago
  • $185k - $225k

     ...building a data-driven decision engine that powers every aspect of...  ...feature success and optimizing user experiences, your work will...  ...Location: This role is based in Palo Alto, CA and involves a hybrid work...  ...testing, cohort analysis, causal inference).Excellent communication... 
    Part time
    Work at office
    Remote work

    Mudflap

    Palo Alto, CA
    10 hours ago
  •  ...democratize AI through high-performance, optimized, open-source and cutting-edge models,...  ...Mistral AI is seeking a Applied AI Engineer to facilitate the adoption of its products...  ...open source codebases for tasks such as inference and fine-tuning. • You’ll be involved... 
    Full time
    Work at office
    Visa sponsorship

    Mistral Ai

    Palo Alto, CA
    1 day ago
  • $222.72k - $389.75k

     ...for a Staff Machine Learning Engineer to lead the technical vision for...  ...of state-of-the-art applied ML projects for ads conversion. Design...  ...understand intention and infer interests from online activity...  ...following offices: San Francisco, Palo Alto, Seattle.#LI-HYBRID #LI-SM4At... 
    Part time
    Work at office
    Local area
    Relocation
    Relocation package

    Pinterest

    Palo Alto, CA
    15 hours ago
  •  ...Rubrik's Semantic AI Governance Engine, which is the first system...  ...lead even further.As an Applied ML Engineer on the SAGE team,...  ...supervised fine-tuning, preference optimization (DPO/RLAIF), and distillation...  ...Model Serving and Inference Infrastructure (25% of time)Designing... 
    Permanent employment
    Part time

    Rubrik

    Palo Alto, CA
    10 hours ago
  • $280.15k

     ...looking for a Distinguished Engineer to set the technical...  ..., raising the bar on ML and systems excellence...  ..., multi-objective optimization) that materially improve...  ..., and causal inference for discovery and generative...  ...following offices, [SF, Palo Alto, Seattle]. #LI-REMOTE#LI... 
    Temporary work
    Part time
    Work at office
    Local area
    Remote work
    Relocation
    Relocation package

    Pinterest

    Palo Alto, CA
    15 hours ago
  • $227.87k

     ...equivalent experience.7+ years of industry experience.Strong software engineering and mathematical skills with knowledge of statistical methods....  ...distance from one of the following offices: San Francisco, Palo Alto, Seattle.Relocation Statement:This position is not eligible... 
    Temporary work
    Part time
    Work at office
    Local area
    Relocation
    Relocation package

    Pinterest

    Palo Alto, CA
    15 hours ago
  •  ...Python or Golang Software developerLocation: Palo Alto, CA Hybrid ( 3 to 4 days)Exp: 10+...  ...CI/CD, terraform and AISenior Software Engineer - Enterprise AIEnterprise AI - Global Infrastructure...  ...orchestration. Design, develop, and optimize distributed services and cloud-native... 
    Part time

    HAN Staffing

    Palo Alto, CA
    10 hours ago
  •  ...We are looking for a Senior MLOps engineer to work closely with Data Scientists to build and deploy ML models on a modern MLOps stack. As Lead...  ...for high-throughput, real-time inference as well as batch inference, ensuring optimal performance and reliability. Implement... 
    Part time

    JP Morgan Chase

    Palo Alto, CA
    15 hours ago
  •  ...collaboration and a strong team culture, this role is expected to be in our Palo Alto office five days a week, unless otherwise specified. About the Role We're seeking an experienced LLM Inference Engineer to optimize our large language model (LLM) serving infrastructure. The ideal... 
    Work at office

    Hippocratic-Ai

    Palo Alto, CA
    3 days ago
  •  ...Palo Alto, CA | Full-Time | On-site About Nace AI: Nace...  ...: As a Senior MLOps Engineer, you will own the infrastructure...  ...at the intersection of ML engineering, LLM inference infrastructure, and...  ...scheduling, utilization, cost optimization) across cloud and on-prem... 
    Full time

    Nace AI

    Palo Alto, CA
    16 hours ago
  • $193.3k - $261.5k

     ...intuitive ways of interacting with data, and we’re looking for top engineers to build them from the ground up.This is a hands-on position...  ...paid time off, and parental leave. Learn more about our benefits at .USA, CA, East Palo Alto - 193,300.00 - 261,500.00 USD annually... 
    Part time
    Internship
    Local area
    Flexible hours

    AmazonWebServices

    Palo Alto, CA
    10 hours ago
  •  ...About the Role We’re looking for an Applied ML Engineer to design, evaluate, and scale recommendation and ranking systems...  ...skips, conversions) and sparse explicit feedback. Optimize systems for real-time inference, scalability, and robustness under non-stationary user... 
    Full time

    Darwin

    Palo Alto, CA
    1 day ago
  • $154.1k - $267.15k

     ...Reference: 733479BRPosted: 2026-07-07Location: Palo Alto, CaliforniaSalary: $154,100 - $267,145...  ...Senior Staff Software/System Integrator Engineer. The selected candidate will take on a...  ...posting date in order to receive optimal consideration.At Lockheed Martin, we use... 
    Full time
    Temporary work
    Part time
    Work experience placement
    Work at office
    Remote work
    Flexible hours

    Lockheed Martin

    Palo Alto, CA
    10 hours ago
  •  ...Palo Alto, CASales /Full-time /Job Description: Sr. Sales EngineerEssential Functions:· Collaborate with our sales teams and partners...  ...with Security Assessment Reports.· Work with Corporate PM and Engineering to capture latest product information and details to be utilized... 
    Full time
    Part time
    Work experience placement

    Confluera

    Palo Alto, CA
    10 hours ago
  •  ...how businesses learn from and optimize in‑person customer...  ...and deploy production‑grade ML systems with end‑to‑end ownership...  ...model training, deployment, inference, and monitoring in production...  ...professional experience in ML engineering. Strong programming skills in... 
    Full time

    Catalyst Labs, LLC

    Mountain View, CA
    1 day ago
  • $65 - $90 per hour

     ...DescriptionKforce's client, a leading automotive engineering organization in Palo Alto, CA is expanding its Hardware-in-the-Loop (HIL) validation capability and is actively seeking experienced HIL Engineers to support the buildout of multiple new systems.Summary:This role... 
    Part time
    Immediate start
    Remote work

    KForce

    Palo Alto, CA
    10 hours ago
  • $145k - $165k

     ...network economics, AI and ML, online and real-world...  ..., Growth, and Revenue optimization. Our mission is to...  ...: Machine Learning Engineers (this role) who focus...  ...three times per week in Palo Alto, California. In...  ...recommendation systems or casual inference Familiarity with... 
    Full time
    Work experience placement
    Casual work
    Work at office
    Flexible hours

    Match Group

    Palo Alto, CA
    5 days ago
  • $170.5k - $315.49k

    ## Inference Optimization Engineer (local / edge runtime)Applylocations: US, California, Santa Clara: US, Oregon, Hillsboro: US, California, Folsom: US, Arizona, Phoenixtime type: Full timeposted on: Posted Yesterdayjob requisition id: JR0284871# **Job Details:**## Job... 
    Internship
    Local area
    Immediate start
    Shift work

    Intel

    Santa Clara, CA
    4 days ago
  • $70 per hour

     ...Hiring: Palo Alto Firewall Engineer – Oil & Gas Domain (Onsite – USA) Location: Onsite – USA Job Type: Full-Time / Contract We are looking for an experienced Palo Alto Firewall Engineer with strong expertise in Palo Alto NGFW, Panorama, OT/IT security, and Oil... 
    Full time
    Contract work

    Gainwell technology

    Palo Alto, CA
    a month ago
  • $175k - $275k

     ...datasets, or full-cycle data engineering, Abaka AI provides the foundation...  ...Abaka builds, trains, and optimizes multimodal AI systems. You...  ...applied machine learning or ML engineering, with a demonstrated...  ...scale distributed training and inference systems. ~ Familiarity with... 
    Full time
    Immediate start
    Flexible hours

    Abaka Ai

    Palo Alto, CA
    1 day ago
  • $208k - $244k

    Join to apply for the Staff ML Engineer role at Grindr Join to apply for the Staff ML Engineer role at Grindr Get AI-powered advice on this...  ...000.00/yr This is a hybrid role based in our San Francisco or Palo Alto offices (Palo Alto preferred) and will require you to be in... 
    Full time
    Casual work
    Work at office
    Immediate start
    Flexible hours

    Grindr

    Palo Alto, CA
    4 days ago
  •  ...world.Role OverviewAs our Staff Software Engineer, ML infra Engineer for Search & Discovery...  ...Discovery organization is responsible for optimizing customers' navigation experience and...  ...ML based ranking system and online ML inference servicesBuild strong cross-functional partnerships... 
    Temporary work
    Part time

    Coupang

    Mountain View, CA
    10 hours ago
  • $177.19k - $364.8k

     ...here.We’re looking for a Staff Software Engineer to help build the next generation of Pinterest...  ...partner closely with teams across data, ML/AI, analytics, and infrastructure to...  ...scale.What you’ll do:Design, implement, and optimize Pinterest’s exabyte-scale data lake... 
    Part time
    Work at office
    Local area
    Relocation
    Relocation package

    Pinterest

    Palo Alto, CA
    10 hours ago
  • $229k - $343k

     ...addition to Bitmoji, Saturn, and other digital services.Snap Engineering teams build fun and technically sophisticated products that reach...  ...eligible for equity in the form of RSUs.SummaryLocation: Santa Monica - 3100 Ocean Park Blvd; Palo Alto, CaliforniaType: Full time... 
    Full time
    Part time
    Live in
    Work at office
    Local area

    Snap

    Palo Alto, CA
    10 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to ML Engineer, Inference & Optimization (Palo Alto). Be the first to apply!