Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior Software Engineer, AI Infrastructure - LVM Inference & Evaluation

Full-time

Ambient.ai

Build a safer world with us, one incident at a time.

Ambient.ai is the category creator and leader in Agentic Physical Security. Powered by Ambient Pulsar, the first reasoning Vision-Language Model purpose-built for physical security, our platform seamlessly integrates with existing security cameras and physical access control systems to unify monitoring, access control, threat assessment, response, and investigations through an always-on reasoning layer that augments security operators with superhuman capabilities. The results: 95% fewer false alarms, investigations 20x faster, and 10x faster response.

The momentum speaks for itself: we doubled new ARR in FY26, and have delivered results for world-class customers including Cisco, ServiceNow, SentinelOne, TikTok, Bayer, and MoMA. That kind of momentum creates an environment where great people thrive, and it shows: we recently ranked #71 out of 500 on the Forbes best startup employers list .

Founded in 2017 and backed by Andreessen Horowitz, Y Combinator, and Allegion Ventures, Ambient.ai is on a fast-paced journey to fulfill our mission: prevent every security incident possible.

Ready to learn more? Connect with us on LinkedIn and YouTube

About the role:

Reporting to Raghu Nallamothu, you will design, build, and optimize the AI infrastructure that powers Ambient.ai ’s real-time intelligence platform.

In this role, you will work on the systems required to run state-of-the-art deep learning models across many terabytes of video data in real time. You will help build and scale infrastructure for inference, evaluation, and continuous model improvement across computer vision models, large language models, large vision models, and multimodal AI systems.

This role is ideal for someone with a strong blend of infrastructure engineering, production ML systems, LLM/LVM inference, evaluation harnesses, and inference optimization experience. You will partner closely with research scientists and product engineering teams to bring the latest AI advancements into production for our customers.

What you'll do:

  • Design, build, and maintain cutting-edge AI infrastructure for real-time computer vision, LLM, LVM, and multimodal inference workloads.

  • Build scalable systems for running state-of-the-art models across large volumes of video and sensor data.

  • Optimize inference performance across latency, throughput, GPU utilization, reliability, and cost.

  • Develop robust evaluation harnesses and benchmarking systems to measure model quality, system performance, regressions, and production readiness.

  • Build infrastructure for continuous model evaluation, experimentation, and deployment.

  • Partner with research scientists to productionize the latest advances in computer vision, LLMs, LVMs, RAG, and multimodal AI.

  • Improve model-serving architecture, including batching, caching, routing, quantization, model parallelism, and hardware utilization.

  • Develop data engines and feedback loops for collecting training data, evaluating model behavior, and continuously improving AI performance.

  • Create reliable observability, monitoring, and debugging tools for production AI systems.

  • Help define best practices for deploying, evaluating, and operating AI systems in real-world enterprise environments.

What you'll bring:

  • 4+ years of industry experience building infrastructure, distributed systems, machine learning platforms, or production AI systems.

  • BS/MS in Computer Science or a related technical field, or equivalent practical experience.

  • Strong programming background, especially in Python, with solid software engineering fundamentals.

  • Experience designing and building scalable machine learning infrastructure for training, inference, evaluation, and deployment.

  • Hands-on experience running deep learning models in production, ideally including LLMs, LVMs, vision-language models, or multimodal models.

  • Strong understanding of inference optimization techniques, including batching, caching, quantization, parallelism, memory optimization, GPU utilization, and latency reduction.

  • Experience with model-serving frameworks or systems such as vLLM, Triton Inference Server or similar technologies.

  • Experience building evaluation frameworks, test harnesses, benchmarks, regression tests, or model-quality measurement systems.

  • Strong background in machine learning and deep learning; computer vision experience is a strong plus.

  • Experience designing data engines or pipelines for collecting, managing, and curating training and evaluation data.

  • Familiarity with integrating advanced AI systems such as LLMs, LVMs, RAG pipelines, embedding models, or multimodal models into production applications.

  • Experience with cloud infrastructure, containers, orchestration, distributed systems, and GPU-based workloads.

  • Strong collaboration and communication skills, with the ability to work effectively with research scientists, product teams, infrastructure teams, and stakeholders.

  • Proactive problem-solving ability, a strong ownership mindset, and adaptability to incorporate new AI technologies and methodologies.

Nice to Have

  • Experience operating large-scale GPU infrastructure or distributed inference systems.

  • Experience with CUDA, NCCL, PyTorch, TensorRT, ONNX, or similar ML systems technologies.

  • Experience with video understanding, real-time computer vision, multimodal AI, or physical-world AI systems.

  • Experience with model compression, speculative decoding, distillation, pruning, or low-latency serving techniques.

  • Experience with prompt evaluation, model regression testing, human-in-the-loop evaluation, or automated quality gates.

  • Familiarity with retrieval-augmented generation, vector databases, embedding models, re-rankers, or search infrastructure.

  • Experience building internal ML platforms or tools used by researchers and applied ML teams.

What Success Looks Like

You will be successful in this role if you can build practical, scalable infrastructure that helps Ambient.ai deploy better AI models faster and more reliably. You should be comfortable working across the full stack of production AI systems, from model behavior and evaluation to serving architecture, GPU performance, observability, and customer-facing reliability.

This is a hands-on engineering role for someone excited to help bring the next generation of AI, computer vision, LLMs, and LVMs into real-world production environments.

Why join us:

  • We are creating an entirely new category within a 180+ billion-dollar physical security industry and looking for team members who are also passionate about our mission to prevent every security incident possible 

  • We partner with an incredible customer roster of F500 companies, including Adobe, TikTok, Gap and SentinelOne

  • Regular Full-time employees receive stock options for the opportunity to share ownership in the success of our company 

  • Comprehensive health + welfare package (Medical, Dental, Vision, Life, EAP, Legal Services, 401k plan)

  • We offer flexible time off to rest and recharge, including Winter Break (time off between Christmas and New Year’s for most roles, depending on customer demand)

  • The latest tech and awesome swag will be delivered to your door

  • Enjoy a full range of opportunities to connect with your awesome co-workers

  • We love to hike , are foodies, and love music! Check out our most recent Ambient Spotify Playlist

We’ve found that in-person time meaningfully supports collaboration, creativity, and team alignment. Our talent, engineering, product, design, and marketing teams work from our Redwood City office three days a week. All other Bay Area employees join on Fridays to stay connected and close out the week together.

Ready to learn more? Connect with us on LinkedIn | YouTube

#LI-Hybrid

Ambient.ai is proud to be an Equal Opportunity Employer. Ambient does not unlawfully discriminate on the basis of race, color, religion, sex (including pregnancy, childbirth, breastfeeding, or related medical conditions), gender identity, gender expression, national origin, ancestry citizenship, age, physical or mental disability, legally protected medical condition, family care status, military or veteran status, marital status, registered domestic partner status, sexual orientation, genetic information, or any other basis protected by local, state, or federal laws. Ambient is an E-Verify participant.

Vacancy posted 21 hours ago
Similar jobs that could be interesting for youBased on the Senior Software Engineer, AI Infrastructure - LVM Inference & Evaluation in Remote vacancy
  •  ...a time. Ambient.ai is the category creator...  ...optimize the AI infrastructure that powers...  ...infrastructure for inference, evaluation, and continuous...  ...of infrastructure engineering, production ML systems, LLM/LVM inference, evaluation...  ..., with solid software engineering fundamentals... 
    Senior
    Full time
    Work at office
    Local area
    Flexible hours
    3 days per week

    Ambient.ai

    Remote
    2 days ago
  •  ...building the next-generation AI search and shopping...  ...ranking to multi-agent LLM engines, post-training infrastructure, and personalized memory....  ...Contribute to Training and Inference Infrastructure: Collaborate...  ...partner with algorithm teams to evaluate agent and LLM innovations,... 
    Senior

    TikTok

    Seattle, WA
    2 days ago
  •  ...builds the shared infrastructure that helps...  ...throughput batch inference, and fine-...  ...for Generative AI at DoorDash, leading...  ...and inference engines, fine-tuning...  ...ideal for a senior engineer who...  ...experience in software engineeringDeep...  ...team by evaluating job related qualifications... 
    Senior
    Hourly pay
    Work at office
    Local area
    Remote work
    Flexible hours

    Doordash

    Seattle, WA
    2 days ago
  • $176k - $220k

     ...started Handshake AI and built the...  ...researchers to create evaluations, publish...  ...Work together with engineers, scientists, operators...  ...data is the core infrastructure to AI advancement...  ...re looking for a Senior Software Engineer to join...  ...evaluation, and inference across Handshake.... 
    Senior
    Full time
    Work at office
    Remote work
    Flexible hours

    Handshake

    San Francisco, CA
    6 days ago
  • $127k - $223k

     ...Description Waabi, founded by AI visionary Raquel...  ...more visit: The Evaluation Algorithms team is responsible...  ...-loop simulation engine built with the latest...  ...Develop the tooling, infrastructure, and pipelines to...  ...programming and strong software engineering fundamentals... 
    Senior
    Full time
    Work at office
    Work from home
    Flexible hours

    Waabi

    San Francisco, CA
    11 days ago
  •  ...Superintelligence Cloud, is a leader in AI cloud infrastructure serving tens of thousands...  ...AI training and inference, raw GPU and CPU...  ...The Lambda Infrastructure Engineering organization forges the foundation...  ...for an experienced Senior Software Engineer to join our storage... 
    Senior
    Work at office
    Local area
    Work from home
    Flexible hours

    Lambda Labs

    San Jose, CA
    4 days ago
  • $155k - $250k

    Senior Software Engineer, InfrastructureAcuityMD is a software and data platform...  .... We're a high-growth AI and Data company scaling rapidly...  ...other learn and grow.As an Infrastructure Engineer, you’ll be...  ...professional experience when evaluating candidates.Team focused mindset... 
    Senior
    For contractors
    Work at office
    Remote work
    Work from home
    Home office
    Flexible hours

    AcuityMD

    Boston, MA
    21 hours ago
  • $200k - $275k

     ...Senior Software Engineer, ML Infrastructure The era of pervasive AI has arrived. In this era, organizations will use generative AI to unlock hidden value in their...  ..., building, and operating the production-grade inference infrastructure that powers SambaNova's serving... 
    Senior
    Remote work

    SambaNova Systems

    United States
    4 days ago
  • $184k - $287.5k

     ...forefront of the generative AI revolution, building the software and systems that power...  ...We are looking for a Senior Software Engineer to lead the bring-up,...  ...training and inference workloads across NVIDIA...  ...large-scale AI clusters, infrastructure, and end-to-end workloads... 
    Senior
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    21 hours ago
  • $160.9k - $260.7k

     ...tools that define how software gets built and delivered. As AI agents redefine...  ..., and secure infrastructure that make autonomous...  ...supports hundreds of engineers and carries high-...  ...'re looking for a senior engineer who can own...  ...and consistent evaluations. Recordings are optional... 
    Senior
    Full time
    Remote work
    Work from home
    Home office
    Visa sponsorship
    Shift work

    Docker

    Seattle, WA
    4 days ago
  •  ...in how people discover, evaluate, and purchase products. The...  ..., we're building the AI-native social operating system...  ...by ex-Meta product and engineering leaders, we've raised over...  ...We're looking for a Senior Software Engineer, Infrastructure to own the reliability, scalability... 
    Senior
    Work at office
    Remote work
    Flexible hours
    Shift work

    Nectar Social

    Palo Alto, CA
    1 day ago
  •  ...Docker Infrastructure Engineering Role Docker has been one of the most loved brands in...  ...building the tools that define how software gets built and delivered. As AI agents redefine software...  ...secure, reliable routing. Evaluate and adopt improvements with a bias... 
    Senior
    Remote work
    Home office
    Visa sponsorship
    Shift work
    Afternoon shift

    Docker

    United States
    1 day ago
  •  ...Opportunity Deepgram is looking for a Senior Software Engineer - Model Evaluation & AI Systems to join the team...  ...methodology and build the infrastructure that measures model quality at scale...  ...alongside Research, model training, inference, and product teams to provide... 
    Senior
    Full time

    Deepgram

    Remote
    29 days ago
  • $152k - $241.5k

     ...recently, GPU deep learning ignited modern AI — the next era of computing — with...  ...for an AI & Deep Learning Compiler Engineer. NVIDIA is hiring software engineers for its Deep Learning & AI...  ...has been the backbone of NVIDIA’s inference engine, spanning across data centers... 
    Senior
    Full time
    Remote work

    Nvidia

    Austin, TX
    21 hours ago
  • $119.8k - $234.7k

     ...25%Profession: Software EngineeringDiscipline...  ...define how AI runs on Windows...  ...challenges.As a Senior Software Engineer, you will design...  ...new knowledge, evaluating new trends, technical...  ...learning inference, graph compilation...  ...machine learning infrastructure, or performance-... 
    Senior
    Ongoing contract
    Local area
    Remote work
    Flexible hours
    3 days per week

    Microsoft

    Redmond, WA
    1 day ago
  • $224k - $356.5k

     ...autonomous driving, and evaluation is how we know the...  ...organization develops AI drivers!We are looking for a senior engineer to own the engine that...  ...behavior planning, and infrastructure — setting expectations,...  ...experience).12+ years building software, with significant time... 
    Senior
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    2 days ago
  • $180k - $215k

     ...artificial intelligence (AI) powered technology stack...  ...commercial self-driving software to develop, test and deploy...  ...We are looking for a Senior Software Engineer to build infrastructure, tools, and systems that...  ...team to develop, debug, evaluate, and deploy autonomy software... 
    Senior
    Full time
    Visa sponsorship

    Kodiak

    Remote
    2 days ago
  • $160k - $240k

    Senior Software Engineer - AI Inference Location New York Business Area Engineering and CTO Ref # 10050779 Description & Requirements...  ...Our team: Join the team that is building the core infrastructure for AI at Bloomberg. The Bloomberg AI Inference... 
    Senior
    Temporary work
    For contractors
    Work experience placement

    Bloomberg

    New York, NY
    21 hours ago
  • $152k - $241.5k

    NVIDIA seeks a Senior Software Engineer specializing in Deep Learning Inference for our growing team. As a key contributor, you will help design, build, and optimize...  ...accelerated software that powers today’s most sophisticated AI applications. Our team is responsible for... 
    Senior
    Full time
    Remote work

    Nvidia

    Texas
    1 day ago
  •  .... About the Organization The Evaluation team builds and evolves the evaluation...  ...into clear feedback for engineering and leadership, and help...  ...introspect autonomous driving software performance at subsystem interfaces...  .... Experience leveraging AI-assisted development and... 
    Senior
    Full time
    Local area
    Work from home

    General Motors

    Sunnyvale, TX
    1 day ago
  • $193.3k - $261.5k

     ...Amazon Neuron, the software development kit used...  ...unparalleled ML inference and training performance...  ...boundary, our engineers build systematic infrastructure, innovate new methods...  ...what's possible in AI acceleration.As part...  ...mentorship. Our senior members enjoy one-on... 
    Senior
    Work experience placement
    Internship
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    18 hours ago
  •  ...leading developer of Embodied AI technology. Our advanced AI software and foundation models...  ...role As a software engineer for Wayve’s Simulation Technology...  ...is used to develop and evaluate Wayve’s driving...  ...large-scale machine learning inference systems running in cloud... 
    Senior
    Full time
    Work at office
    Work from home

    Wayve

    United Kingdom
    6 days ago
  • $180k - $250k

     ...Lightning AI is the company...  ...developer-first software with cost-efficient...  ...production inference, with security...  ...customer and engineering team depends on...  ...looking for a Senior Software Engineer...  ...with product, infrastructure, and AI...  ...capabilities. Evaluate and improve technical... 
    Senior
    Work at office
    Work from home
    Flexible hours
    2 days per week

    Lightning AI

    New York, NY
    4 days ago
  •  ...usher in this new era, we seek AI-native thinkers across every...  ...challenges including infrastructure optimizations, orchestration...  ...collaboratively and proactively with senior architects, PMs, and team...  ...in serving LLMs using inference engines like vLLM, TensorRT-LLM, TEI... 
    Senior

    Snowflake

    Remote
    2 days ago
  • $202.5k - $247.5k

     ...localhost or running AI workloads in...  ...delivery, AI inference, device fleets...  ...! We like software that’s serious...  ...systems ngrok engineers rely on to...  ...We think about infrastructure the way software...  ...Job Title Senior Software Engineer...  ...will be evaluated based on factors... 
    Senior
    Permanent employment
    Full time
    Work at office
    Local area
    Remote work
    Home office
    Flexible hours

    Ngrok

    Remote
    2 days ago
  • $204k - $259k

     ...dynamics, and state-of-the-art Generative AI to create a training ground for the Waymo Driver. The Simulator Evaluation team faces the ultimate data challenge:...  ...world is "real"? We are looking for a Senior Software Engineer to build the metrics and systems that grade... 
    Senior
    Full time
    Remote work

    Waymo

    San Francisco, CA
    2 days ago
  •  ...Role Overview Design and evaluate high-quality datasets and evaluations that advance...  ...corrections across multiple languages, assess AI-generated implementations for...  ...measure model capabilities across the software engineering lifecycle. Key Responsibilities Curate... 
    Senior
    For contractors
    Remote work
    10 hours per week
    Flexible hours

    SaidGig

    Canada
    more than 2 months ago
  •  ...development of high-quality datasets and evaluation pipelines that improve and benchmark large...  ...language models for code generation and software engineering tasks. You will curate and author reference code, evaluate and refine AI-generated solutions across multiple programming... 
    Senior
    Full time
    For contractors
    Remote work
    10 hours per week
    Flexible hours

    SaidGig

    United States
    more than 2 months ago
  • $200k - $250k

     ...About Wizard AI At Wizard AI, we’re building...  ...and we’re looking for a Senior MLOps Engineer to help us run them reliably...  ...— across a custom-built inference platform powering a live...  ...years of experience in software, ML, platform, or infrastructure engineering, with hands-... 
    Senior
    Full time
    Remote work
    Flexible hours

    Wizard

    United States
    2 days ago
  •  ...Description We're building a dataset to evaluate AI coding agents - how well a model...  ...plan, you create a company: codebase, infrastructure, context (conversations, documentation...  ...involvement. Qualifications ~Strong software engineering experience: 8+ years in any modern... 
    Senior
    Temporary work
    Freelance
    Remote work
    Flexible hours

    Toloka

    Remote
    20 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior Software Engineer, AI Infrastructure - LVM Inference & Evaluation. Be the first to apply!