Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Software Engineer, AI Infrastructure - LVM Inference & Evaluation

$168k - $205k
Full-time

Ambient.ai

Build a safer world with us, one incident at a time.

Ambient.ai is the category creator and leader in Agentic Physical Security. Powered by Ambient Pulsar, the first reasoning Vision-Language Model purpose-built for physical security, our platform seamlessly integrates with existing security cameras and physical access control systems to unify monitoring, access control, threat assessment, response, and investigations through an always-on reasoning layer that augments security operators with superhuman capabilities. The results: 95% fewer false alarms, investigations 20x faster, and 10x faster response.

The momentum speaks for itself: we doubled new ARR in FY26, and have delivered results for world-class customers including Cisco, ServiceNow, SentinelOne, TikTok, Bayer, and MoMA. That kind of momentum creates an environment where great people thrive, and it shows: we recently ranked #71 out of 500 on the Forbes best startup employers list.

Founded in 2017 and backed by Andreessen Horowitz, Y Combinator, and Allegion Ventures, Ambient.ai is on a fast-paced journey to fulfill our mission: prevent every security incident possible.

Ready to learn more? Connect with us on LinkedIn and YouTube

About the role:

Reporting to Raghu Nallamothu, you will design, build, and optimize the AI infrastructure that powers Ambient.ai’s real-time intelligence platform.

In this role, you will work on the systems required to run state-of-the-art deep learning models across many terabytes of video data in real time. You will help build and scale infrastructure for inference, evaluation, and continuous model improvement across computer vision models, large language models, large vision models, and multimodal AI systems.

This role is ideal for someone with a strong blend of infrastructure engineering, production ML systems, LLM/LVM inference, evaluation harnesses, and inference optimization experience. You will partner closely with research scientists and product engineering teams to bring the latest AI advancements into production for our customers.

What you'll do:

  • Design, build, and maintain cutting-edge AI infrastructure for real-time computer vision, LLM, LVM, and multimodal inference workloads.

  • Build scalable systems for running state-of-the-art models across large volumes of video and sensor data.

  • Optimize inference performance across latency, throughput, GPU utilization, reliability, and cost.

  • Develop robust evaluation harnesses and benchmarking systems to measure model quality, system performance, regressions, and production readiness.

  • Build infrastructure for continuous model evaluation, experimentation, and deployment.

  • Partner with research scientists to productionize the latest advances in computer vision, LLMs, LVMs, RAG, and multimodal AI.

  • Improve model-serving architecture, including batching, caching, routing, quantization, model parallelism, and hardware utilization.

  • Develop data engines and feedback loops for collecting training data, evaluating model behavior, and continuously improving AI performance.

  • Create reliable observability, monitoring, and debugging tools for production AI systems.

  • Help define best practices for deploying, evaluating, and operating AI systems in real-world enterprise environments.

What you'll bring:

  • 2+ years of industry experience building infrastructure, distributed systems, machine learning platforms, or production AI systems.

  • BS/MS in Computer Science or a related technical field, or equivalent practical experience.

  • Strong programming background, especially in Python, with solid software engineering fundamentals.

  • Experience designing and building scalable machine learning infrastructure for training, inference, evaluation, and deployment.

  • Hands-on experience running deep learning models in production, ideally including LLMs, LVMs, vision-language models, or multimodal models.

  • Strong understanding of inference optimization techniques, including batching, caching, quantization, parallelism, memory optimization, GPU utilization, and latency reduction.

  • Experience with model-serving frameworks or systems such as vLLM, Triton Inference Server or similar technologies.

  • Experience building evaluation frameworks, test harnesses, benchmarks, regression tests, or model-quality measurement systems.

  • Strong background in machine learning and deep learning; computer vision experience is a strong plus.

  • Experience designing data engines or pipelines for collecting, managing, and curating training and evaluation data.

  • Familiarity with integrating advanced AI systems such as LLMs, LVMs, RAG pipelines, embedding models, or multimodal models into production applications.

  • Experience with cloud infrastructure, containers, orchestration, distributed systems, and GPU-based workloads.

  • Strong collaboration and communication skills, with the ability to work effectively with research scientists, product teams, infrastructure teams, and stakeholders.

  • Proactive problem-solving ability, a strong ownership mindset, and adaptability to incorporate new AI technologies and methodologies.

Nice to Have

  • Experience operating large-scale GPU infrastructure or distributed inference systems.

  • Experience with CUDA, NCCL, PyTorch, TensorRT, ONNX, or similar ML systems technologies.

  • Experience with video understanding, real-time computer vision, multimodal AI, or physical-world AI systems.

  • Experience with model compression, speculative decoding, distillation, pruning, or low-latency serving techniques.

  • Experience with prompt evaluation, model regression testing, human-in-the-loop evaluation, or automated quality gates.

  • Familiarity with retrieval-augmented generation, vector databases, embedding models, re-rankers, or search infrastructure.

  • Experience building internal ML platforms or tools used by researchers and applied ML teams.

What Success Looks Like

You will be successful in this role if you can build practical, scalable infrastructure that helps Ambient.ai deploy better AI models faster and more reliably. You should be comfortable working across the full stack of production AI systems, from model behavior and evaluation to serving architecture, GPU performance, observability, and customer-facing reliability.

This is a hands-on engineering role for someone excited to help bring the next generation of AI, computer vision, LLMs, and LVMs into real-world production environments.

Why join us:

  • We are creating an entirely new category within a 180+ billion-dollar physical security industry and looking for team members who are also passionate about our mission to prevent every security incident possible

  • We partner with an incredible customer roster of F500 companies, including Adobe, TikTok, Gap and SentinelOne

  • Regular Full-time employees receive stock options for the opportunity to share ownership in the success of our company

  • Comprehensive health + welfare package (Medical, Dental, Vision, Life, EAP, Legal Services, 401k plan)

  • We offer flexible time off to rest and recharge, including Winter Break (time off between Christmas and New Year’s for most roles, depending on customer demand)

  • The latest tech and awesome swag will be delivered to your door

  • Enjoy a full range of opportunities to connect with your awesome co-workers

  • We love to hike , are foodies, and love music! Check out our most recent Ambient Spotify Playlist

We’ve found that in-person time meaningfully supports collaboration, creativity, and team alignment. Our talent, engineering, product, design, and marketing teams work from our Redwood City office three days a week. All other Bay Area employees join on Fridays to stay connected and close out the week together.

Ready to learn more? Connect with us on LinkedIn | YouTube

#LI-Hybrid

Ambient.ai is proud to be an Equal Opportunity Employer. Ambient does not unlawfully discriminate on the basis of race, color, religion, sex (including pregnancy, childbirth, breastfeeding, or related medical conditions), gender identity, gender expression, national origin, ancestry citizenship, age, physical or mental disability, legally protected medical condition, family care status, military or veteran status, marital status, registered domestic partner status, sexual orientation, genetic information, or any other basis protected by local, state, or federal laws. Ambient is an E-Verify participant.

Vacancy posted 5 days ago
Similar jobs that could be interesting for youBased on the Software Engineer, AI Infrastructure - LVM Inference & Evaluation in Redwood City, CA vacancy
  • $135k - $260k

     ...About Beacon AI We’re a fast-moving team of aviators, engineers, and operators building an AI platform...  ...skilled Cloud and ML Infrastructure Engineers to lead the...  ...and/or SageMaker; evaluate when to use ECS/EKS, Lambda, or Batch for inference jobs. Build and maintain... 
    Suggested
    Permanent employment
    Full time
    Local area
    Remote work
    3 days per week

    Beacon AI

    San Carlos, CA
    7 days ago
  • $157k - $235k

     ...digital services.Snap Engineering teams build fun...  ...re looking for a Software Engineer to join...  .... We are an AI native team which...  ...Design and optimize infrastructure systems for machine...  ...high-performance inference systems to ensure...  ...model training, evaluation, and inference in... 
    Suggested
    Full time
    Live in
    Work at office
    Local area

    Snap

    Palo Alto, CA
    3 days ago
  • $193.93k - $352.29k

     ...opportunity for AI to drive positive...  ...fungible is the infrastructure that decides whether...  ...can be trusted — evaluation, verification,...  ...inside Nuro's own engineering organization, under...  ...You3+ years of software engineering experience...  ...the hood at inference. Attention and KV... 
    Suggested
    Immediate start
    Flexible hours

    Nuro

    Mountain View, CA
    4 days ago
  • $145k - $200k

     ...world’s leading software for data-driven decisions...  ...are a software engineering team with...  ...production. We deploy AI models to run in...  ...full stack, from inference engines, GPU...  ...intersection of infrastructure and machine learning...  ...ability to quickly evaluate and integrate new... 
    Suggested
    Full time
    Work experience placement
    Work at office
    Remote work
    Work from home
    Relocation package

    Palantir Technologies

    Palo Alto, CA
    3 days ago
  • $160.36k - $240.54k

     ...immediate and profound opportunity for AI to drive positive change in the physical...  ...leading investors.About the RoleThe ML Infrastructure team is responsible for building & improving...  ...road validation.Maintain an in-house ML inference platform to serve large language models... 
    Suggested
    Immediate start
    Flexible hours

    Nuro

    Mountain View, CA
    4 days ago
  • C3 AI (NYSE: AI), is the Enterprise AI application software company. C3 AI delivers a family of fully...  ...for a senior software engineer to help build the C3 Agentic...  ...accuracy: build the evaluations, grounding, and...  ...throughput, scalability, and inference cost across the... 
    Work experience placement

    C3 IoT

    Redwood City, CA
    5 hours ago
  •  ...models get served — integrating new architectures into the inference engine, scaling deployments across thousands of machines, and keeping...  ...engine.Collaborate across research, engineering, and infrastructure to optimize model efficiency and deployments.Build internal... 

    Luma AI

    Redwood City, CA
    3 days ago
  • $175k - $287k

     ...business needs of the team. LinkedIn’s Core AI is building the Evaluation Operating System (EOS), a foundational Agent...  ...This platform also is responsible for tracing infrastructure for all LinkedIn AI Agents.As a Staff Engineer, you will own the end-to-end technical... 
    For contractors
    Work experience placement
    Work at office
    Flexible hours

    Linkedin

    Mountain View, CA
    2 days ago
  • $180k - $250k

     ...despite using far less compute at inference time, substantially reducing...  ..., Microsoft, Amazon, and AI visionaries like Geoff...  ...both data research and data engineering necessary to solve this incredibly...  ...for an experienced Cloud Infrastructure Engineer to join our core... 
    Work at office
    Relocation package

    datologyai

    Redwood City, CA
    17 hours ago
  • $200k - $300k

     ...Company Overview At Skild AI, we are building the world's first general...  ...Skild AI, Inc. seeks a Senior Software Engineer, AI Training & Infrastructure in San Mateo, CA. You will be...  ...preparation, training orchestration, evaluation, and deployment—for real-world robotics... 
    Full time

    Skild AI

    San Mateo, CA
    1 day ago
  • $160.36k - $240.54k

     ...immediate and profound opportunity for AI to drive positive change in the physical...  ...investors. About the Role The ML Infrastructure team is responsible for building & improving...  ...validation. Maintain an in-house ML inference platform to serve large language models... 
    Immediate start
    Flexible hours

    Nuro

    Mountain View, CA
    12 days ago
  • $180k - $215k

     ...an artificial intelligence (AI) powered technology stack...  ...its commercial self-driving software to develop, test and deploy...  ...looking for a Senior Software Engineer to build infrastructure, tools, and systems that...  ...Planning team to develop, debug, evaluate, and deploy autonomy... 
    Visa sponsorship

    Kodiak Robotics

    Mountain View, CA
    2 days ago
  • $160.36k - $240.54k

     ...profound opportunity for AI to drive positive...  ...investors.About the RoleOur software team is growing, and...  ...looking for talented engineers to join us and be...  ...Simulation, and Technical Infrastructure.Data Platform: The Data...  ...supports the autonomy evaluation infrastructure by... 
    Immediate start
    Flexible hours

    Nuro

    Mountain View, CA
    2 days ago
  • $152k - $228k

     ...profound opportunity for AI to drive positive...  ...bench-top systems to evaluate and regression test different...  ...aspects of the software and hardware integration...  ...road. You will own the infrastructure that makes this possible...  ..., and much much more.Engineers across the company... 
    Temporary work
    Immediate start
    Flexible hours

    Nuro

    Mountain View, CA
    5 hours ago
  • $198k - $326k

     ...us in building the AI Governance...  ...model validation and evaluation workflows, and apply...  ..., and enforceable engineering capabilities.As a Sr. Staff Software Engineer, you will...  ...of AI Governance infrastructure. You will solve complex...  ..., validation, inference, or deployment.Preferred... 
    For contractors
    Work at office
    Flexible hours

    Linkedin

    Mountain View, CA
    5 hours ago
  •  ...Software Engineer, Infrastructure Scaled Cognition is the world's only model lab dedicated exclusively to...  ...world CX workflows. Founded by serial AI entrepreneurs, former Microsoft...  ...Cognition you will: Design and improve inference infrastructure that powers our AI... 

    Scaled Cognition

    Mountain View, CA
    2 days ago
  •  ...Software EngineerAnyscale is looking for a Software Engineer to join the Infrastructure team. Anyscale aims to provide the next generation of tools...  ...and running distributed AI applications in the cloud as...  ...Opportunity Employer. Candidates are evaluated without regard to age, race,... 

    Anyscale

    Palo Alto, CA
    3 days ago
  • $180k - $300k

     ...despite using far less compute at inference time, substantially reducing...  ..., Microsoft, Amazon, and AI visionaries like Geoff...  ...both data research and data engineering necessary to solve this incredibly...  ...looking for an experienced Infrastructure Engineer to join as a member... 
    Work at office
    Work from home
    Relocation package

    DatologyAI

    San Mateo, CA
    4 days ago
  •  ...shift in how people discover, evaluate, and purchase products....  ..., we're building the AI-native social operating system...  ...Founded by ex-Meta product and engineering leaders, we've raised over...  ...'re looking for a Senior Software Engineer, Infrastructure to own the reliability,... 
    Work at office
    Remote work
    Flexible hours
    Shift work

    Nectar Social

    Palo Alto, CA
    3 days ago
  • $160.36k - $240.54k

     ...profound opportunity for AI to drive positive...  ...diversity of its training and evaluation data. The team plays a...  ...and reliable data infrastructure. This infrastructure...  ...collaborates closely with system engineers to thoroughly validate...  ...across broader software organizationA bachelor... 
    Immediate start
    Flexible hours

    Nuro

    Mountain View, CA
    2 days ago
  • $193.93k - $291.15k

     ...and profound opportunity for AI to drive positive change in...  ...autonomous driving, and the ML Infrastructure team builds and operates the...  ...- data-to-training-to-evaluation pipelines that are introspectable...  ...Science, Electrical Engineering, or a closely related field,... 
    Work experience placement
    Immediate start
    Flexible hours

    Nuro

    Mountain View, CA
    4 days ago
  • $150k - $200k

     ...design, make limited use of software, and are difficult and stressful...  ..., beautifully designed, AI optimized home heating and cooling...  ...C++ on custom devices, cloud infrastructure, and mobile apps. Right now,...  ...We're looking for a software engineer to own our developer... 
    Full time
    Work at office
    Local area
    Relocation package
    Shift work
    3 days per week

    Quilt

    Redwood City, CA
    12 days ago
  • $148.2k - $222.2k

    Staff Engineer, Infrastructure Platforms Position SummaryThe Staff Engineer, Infrastructure...  ...Research, Bioinformatics, Software Engineering, Information...  ...and operational excellence.Evaluate emerging technologies and...  ..., genomics, bioinformatics, AI/ML, or other scientific computing... 
    Full time
    Work from home
    Monday to Friday

    Pacific Biosciences

    Menlo Park, CA
    1 day ago
  • $180k - $250k

     ...We're looking for an experienced Data Platform Engineer to join as a member of our core Datology AI team. As one of our early senior hires, you will partner...  ...technical lead of a Data Engineering / Platform / Infrastructure Team. Experience building ML/DL systems and/... 
    Work at office
    Visa sponsorship
    Relocation package

    datologyai

    Redwood City, CA
    4 days ago
  • $200k - $420k

     ...At River AI, our mission is to create personal AI owned and shaped by each...  ...: personal hardware for local inference, bespoke training infrastructure, next-generation UIs, and frontier...  ...Who we are We are scientists, engineers, and builders from the industry's top... 
    Full time
    Local area
    Visa sponsorship
    Relocation package

    River AI Inc.

    Palo Alto, CA
    3 days ago
  • $1,000 - $2,030 per month

     ...ll do Work on the Machine Learning and AI Infrastructure team, which is focused on providing infrastructure for data scientists and other engineers to use ML and AI technologies effectively. Build and maintain software and tools for classical ML stack which... 
    Full time
    Temporary work
    Work at office
    Flexible hours

    Cloudkitchens

    Mountain View, CA
    1 day ago
  •  ...profound opportunity for AI to drive positive...  ...About the Role As a software engineering intern, you will work...  ...Onboard Systems, ML Infrastructure, Simulation, or...  ...supports the autonomy evaluation infrastructure by providing...  ...training and onboard inference. Our solutions... 
    Internship
    Immediate start
    Flexible hours

    Nuro

    Mountain View, CA
    12 days ago
  •  ...with the ultimate goal of enabling human life on Mars.SOFTWARE ENGINEER, INFERENCE (AI DATA ENGINEERING)The application software team is the central...  ...systems end-to-end, owning everything from distributed infrastructure to deep low-level optimizations. You will work on... 
    Permanent employment
    Temporary work
    Remote work
    Worldwide
    Weekend work

    SpaceX

    Palo Alto, CA
    1 day ago
  • $180k - $276k

     ..., CA / Seattle, WASoftware - Software Systems /Full-time /HybridThe...  ...the team that owns the build infrastructure making that possible: maintaining...  ...monorepo used by hundreds of engineers.As a key contributor, you...  ...use artificial intelligence (AI) tools to support parts of the... 
    Full time
    Temporary work
    Remote work
    Relocation package

    Zoox

    Foster, CA
    5 hours ago
  • $188.5k - $282.7k

     ...innovation and solving complex engineering problems for our customers....  ...you’ll do:Responsibilities:Software Development: Understand and...  ...businesses with their data when infrastructure is attacked.Rubrik provides...  ...Accelerating the World's AI TransformationRubrik (RBRK),... 
    Local area

    Rubrik

    Palo Alto, CA
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Software Engineer, AI Infrastructure - LVM Inference & Evaluation. Be the first to apply!