Senior Software Engineer, AI Infrastructure - LVM Inference & Evaluation
Ambient.ai
Build a safer world with us, one incident at a time.
Ambient.ai is the category creator and leader in Agentic Physical Security. Powered by Ambient Pulsar, the first reasoning Vision-Language Model purpose-built for physical security, our platform seamlessly integrates with existing security cameras and physical access control systems to unify monitoring, access control, threat assessment, response, and investigations through an always-on reasoning layer that augments security operators with superhuman capabilities. The results: 95% fewer false alarms, investigations 20x faster, and 10x faster response.
The momentum speaks for itself: we doubled new ARR in FY26, and have delivered results for world-class customers including Cisco, ServiceNow, SentinelOne, TikTok, Bayer, and MoMA. That kind of momentum creates an environment where great people thrive, and it shows: we recently ranked #71 out of 500 on the Forbes best startup employers list .
Founded in 2017 and backed by Andreessen Horowitz, Y Combinator, and Allegion Ventures, Ambient.ai is on a fast-paced journey to fulfill our mission: prevent every security incident possible.
Ready to learn more? Connect with us on LinkedIn and YouTube
About the role:
Reporting to Raghu Nallamothu, you will design, build, and optimize the AI infrastructure that powers Ambient.ai ’s real-time intelligence platform.
In this role, you will work on the systems required to run state-of-the-art deep learning models across many terabytes of video data in real time. You will help build and scale infrastructure for inference, evaluation, and continuous model improvement across computer vision models, large language models, large vision models, and multimodal AI systems.
This role is ideal for someone with a strong blend of infrastructure engineering, production ML systems, LLM/LVM inference, evaluation harnesses, and inference optimization experience. You will partner closely with research scientists and product engineering teams to bring the latest AI advancements into production for our customers.
What you'll do:
Design, build, and maintain cutting-edge AI infrastructure for real-time computer vision, LLM, LVM, and multimodal inference workloads.
Build scalable systems for running state-of-the-art models across large volumes of video and sensor data.
Optimize inference performance across latency, throughput, GPU utilization, reliability, and cost.
Develop robust evaluation harnesses and benchmarking systems to measure model quality, system performance, regressions, and production readiness.
Build infrastructure for continuous model evaluation, experimentation, and deployment.
Partner with research scientists to productionize the latest advances in computer vision, LLMs, LVMs, RAG, and multimodal AI.
Improve model-serving architecture, including batching, caching, routing, quantization, model parallelism, and hardware utilization.
Develop data engines and feedback loops for collecting training data, evaluating model behavior, and continuously improving AI performance.
Create reliable observability, monitoring, and debugging tools for production AI systems.
Help define best practices for deploying, evaluating, and operating AI systems in real-world enterprise environments.
What you'll bring:
4+ years of industry experience building infrastructure, distributed systems, machine learning platforms, or production AI systems.
BS/MS in Computer Science or a related technical field, or equivalent practical experience.
Strong programming background, especially in Python, with solid software engineering fundamentals.
Experience designing and building scalable machine learning infrastructure for training, inference, evaluation, and deployment.
Hands-on experience running deep learning models in production, ideally including LLMs, LVMs, vision-language models, or multimodal models.
Strong understanding of inference optimization techniques, including batching, caching, quantization, parallelism, memory optimization, GPU utilization, and latency reduction.
Experience with model-serving frameworks or systems such as vLLM, Triton Inference Server or similar technologies.
Experience building evaluation frameworks, test harnesses, benchmarks, regression tests, or model-quality measurement systems.
Strong background in machine learning and deep learning; computer vision experience is a strong plus.
Experience designing data engines or pipelines for collecting, managing, and curating training and evaluation data.
Familiarity with integrating advanced AI systems such as LLMs, LVMs, RAG pipelines, embedding models, or multimodal models into production applications.
Experience with cloud infrastructure, containers, orchestration, distributed systems, and GPU-based workloads.
Strong collaboration and communication skills, with the ability to work effectively with research scientists, product teams, infrastructure teams, and stakeholders.
Proactive problem-solving ability, a strong ownership mindset, and adaptability to incorporate new AI technologies and methodologies.
Nice to Have
Experience operating large-scale GPU infrastructure or distributed inference systems.
Experience with CUDA, NCCL, PyTorch, TensorRT, ONNX, or similar ML systems technologies.
Experience with video understanding, real-time computer vision, multimodal AI, or physical-world AI systems.
Experience with model compression, speculative decoding, distillation, pruning, or low-latency serving techniques.
Experience with prompt evaluation, model regression testing, human-in-the-loop evaluation, or automated quality gates.
Familiarity with retrieval-augmented generation, vector databases, embedding models, re-rankers, or search infrastructure.
Experience building internal ML platforms or tools used by researchers and applied ML teams.
What Success Looks Like
You will be successful in this role if you can build practical, scalable infrastructure that helps Ambient.ai deploy better AI models faster and more reliably. You should be comfortable working across the full stack of production AI systems, from model behavior and evaluation to serving architecture, GPU performance, observability, and customer-facing reliability.
This is a hands-on engineering role for someone excited to help bring the next generation of AI, computer vision, LLMs, and LVMs into real-world production environments.
Why join us:
We are creating an entirely new category within a 180+ billion-dollar physical security industry and looking for team members who are also passionate about our mission to prevent every security incident possible
We partner with an incredible customer roster of F500 companies, including Adobe, TikTok, Gap and SentinelOne
Regular Full-time employees receive stock options for the opportunity to share ownership in the success of our company
Comprehensive health + welfare package (Medical, Dental, Vision, Life, EAP, Legal Services, 401k plan)
We offer flexible time off to rest and recharge, including Winter Break (time off between Christmas and New Year’s for most roles, depending on customer demand)
The latest tech and awesome swag will be delivered to your door
Enjoy a full range of opportunities to connect with your awesome co-workers
We love to hike , are foodies, and love music! Check out our most recent Ambient Spotify Playlist
We’ve found that in-person time meaningfully supports collaboration, creativity, and team alignment. Our talent, engineering, product, design, and marketing teams work from our Redwood City office three days a week. All other Bay Area employees join on Fridays to stay connected and close out the week together.
Ready to learn more? Connect with us on LinkedIn | YouTube
#LI-Hybrid
Ambient.ai is proud to be an Equal Opportunity Employer. Ambient does not unlawfully discriminate on the basis of race, color, religion, sex (including pregnancy, childbirth, breastfeeding, or related medical conditions), gender identity, gender expression, national origin, ancestry citizenship, age, physical or mental disability, legally protected medical condition, family care status, military or veteran status, marital status, registered domestic partner status, sexual orientation, genetic information, or any other basis protected by local, state, or federal laws. Ambient is an E-Verify participant.
- ...a time. Ambient.ai is the category creator... ...optimize the AI infrastructure that powers... ...infrastructure for inference, evaluation, and continuous... ...of infrastructure engineering, production ML systems, LLM/LVM inference, evaluation... ..., with solid software engineering fundamentals...SeniorFull timeWork at officeLocal areaFlexible hours3 days per week
- ...building the next-generation AI search and shopping... ...ranking to multi-agent LLM engines, post-training infrastructure, and personalized memory.... ...Contribute to Training and Inference Infrastructure: Collaborate... ...partner with algorithm teams to evaluate agent and LLM innovations,...Senior
- ...builds the shared infrastructure that helps... ...throughput batch inference, and fine-... ...for Generative AI at DoorDash, leading... ...and inference engines, fine-tuning... ...ideal for a senior engineer who... ...experience in software engineeringDeep... ...team by evaluating job related qualifications...SeniorHourly payWork at officeLocal areaRemote workFlexible hours
$176k - $220k
...started Handshake AI and built the... ...researchers to create evaluations, publish... ...Work together with engineers, scientists, operators... ...data is the core infrastructure to AI advancement... ...re looking for a Senior Software Engineer to join... ...evaluation, and inference across Handshake....SeniorFull timeWork at officeRemote workFlexible hours$127k - $223k
...Description Waabi, founded by AI visionary Raquel... ...more visit: The Evaluation Algorithms team is responsible... ...-loop simulation engine built with the latest... ...Develop the tooling, infrastructure, and pipelines to... ...programming and strong software engineering fundamentals...SeniorFull timeWork at officeWork from homeFlexible hours- ...Superintelligence Cloud, is a leader in AI cloud infrastructure serving tens of thousands... ...AI training and inference, raw GPU and CPU... ...The Lambda Infrastructure Engineering organization forges the foundation... ...for an experienced Senior Software Engineer to join our storage...SeniorWork at officeLocal areaWork from homeFlexible hours
$155k - $250k
Senior Software Engineer, InfrastructureAcuityMD is a software and data platform... .... We're a high-growth AI and Data company scaling rapidly... ...other learn and grow.As an Infrastructure Engineer, you’ll be... ...professional experience when evaluating candidates.Team focused mindset...SeniorFor contractorsWork at officeRemote workWork from homeHome officeFlexible hours$200k - $275k
...Senior Software Engineer, ML Infrastructure The era of pervasive AI has arrived. In this era, organizations will use generative AI to unlock hidden value in their... ..., building, and operating the production-grade inference infrastructure that powers SambaNova's serving...SeniorRemote work$184k - $287.5k
...forefront of the generative AI revolution, building the software and systems that power... ...We are looking for a Senior Software Engineer to lead the bring-up,... ...training and inference workloads across NVIDIA... ...large-scale AI clusters, infrastructure, and end-to-end workloads...SeniorFull timeRemote work$160.9k - $260.7k
...tools that define how software gets built and delivered. As AI agents redefine... ..., and secure infrastructure that make autonomous... ...supports hundreds of engineers and carries high-... ...'re looking for a senior engineer who can own... ...and consistent evaluations. Recordings are optional...SeniorFull timeRemote workWork from homeHome officeVisa sponsorshipShift work- ...in how people discover, evaluate, and purchase products. The... ..., we're building the AI-native social operating system... ...by ex-Meta product and engineering leaders, we've raised over... ...We're looking for a Senior Software Engineer, Infrastructure to own the reliability, scalability...SeniorWork at officeRemote workFlexible hoursShift work
- ...Docker Infrastructure Engineering Role Docker has been one of the most loved brands in... ...building the tools that define how software gets built and delivered. As AI agents redefine software... ...secure, reliable routing. Evaluate and adopt improvements with a bias...SeniorRemote workHome officeVisa sponsorshipShift workAfternoon shift
- ...Opportunity Deepgram is looking for a Senior Software Engineer - Model Evaluation & AI Systems to join the team... ...methodology and build the infrastructure that measures model quality at scale... ...alongside Research, model training, inference, and product teams to provide...SeniorFull time
$152k - $241.5k
...recently, GPU deep learning ignited modern AI — the next era of computing — with... ...for an AI & Deep Learning Compiler Engineer. NVIDIA is hiring software engineers for its Deep Learning & AI... ...has been the backbone of NVIDIA’s inference engine, spanning across data centers...SeniorFull timeRemote work$119.8k - $234.7k
...25%Profession: Software EngineeringDiscipline... ...define how AI runs on Windows... ...challenges.As a Senior Software Engineer, you will design... ...new knowledge, evaluating new trends, technical... ...learning inference, graph compilation... ...machine learning infrastructure, or performance-...SeniorOngoing contractLocal areaRemote workFlexible hours3 days per week$224k - $356.5k
...autonomous driving, and evaluation is how we know the... ...organization develops AI drivers!We are looking for a senior engineer to own the engine that... ...behavior planning, and infrastructure — setting expectations,... ...experience).12+ years building software, with significant time...SeniorFull timeRemote work$180k - $215k
...artificial intelligence (AI) powered technology stack... ...commercial self-driving software to develop, test and deploy... ...We are looking for a Senior Software Engineer to build infrastructure, tools, and systems that... ...team to develop, debug, evaluate, and deploy autonomy software...SeniorFull timeVisa sponsorship$160k - $240k
Senior Software Engineer - AI Inference Location New York Business Area Engineering and CTO Ref # 10050779 Description & Requirements... ...Our team: Join the team that is building the core infrastructure for AI at Bloomberg. The Bloomberg AI Inference...SeniorTemporary workFor contractorsWork experience placement$152k - $241.5k
NVIDIA seeks a Senior Software Engineer specializing in Deep Learning Inference for our growing team. As a key contributor, you will help design, build, and optimize... ...accelerated software that powers today’s most sophisticated AI applications. Our team is responsible for...SeniorFull timeRemote work- .... About the Organization The Evaluation team builds and evolves the evaluation... ...into clear feedback for engineering and leadership, and help... ...introspect autonomous driving software performance at subsystem interfaces... .... Experience leveraging AI-assisted development and...SeniorFull timeLocal areaWork from home
$193.3k - $261.5k
...Amazon Neuron, the software development kit used... ...unparalleled ML inference and training performance... ...boundary, our engineers build systematic infrastructure, innovate new methods... ...what's possible in AI acceleration.As part... ...mentorship. Our senior members enjoy one-on...SeniorWork experience placementInternshipLocal areaFlexible hours- ...leading developer of Embodied AI technology. Our advanced AI software and foundation models... ...role As a software engineer for Wayve’s Simulation Technology... ...is used to develop and evaluate Wayve’s driving... ...large-scale machine learning inference systems running in cloud...SeniorFull timeWork at officeWork from home
$180k - $250k
...Lightning AI is the company... ...developer-first software with cost-efficient... ...production inference, with security... ...customer and engineering team depends on... ...looking for a Senior Software Engineer... ...with product, infrastructure, and AI... ...capabilities. Evaluate and improve technical...SeniorWork at officeWork from homeFlexible hours2 days per week- ...usher in this new era, we seek AI-native thinkers across every... ...challenges including infrastructure optimizations, orchestration... ...collaboratively and proactively with senior architects, PMs, and team... ...in serving LLMs using inference engines like vLLM, TensorRT-LLM, TEI...Senior
$202.5k - $247.5k
...localhost or running AI workloads in... ...delivery, AI inference, device fleets... ...! We like software that’s serious... ...systems ngrok engineers rely on to... ...We think about infrastructure the way software... ...Job Title Senior Software Engineer... ...will be evaluated based on factors...SeniorPermanent employmentFull timeWork at officeLocal areaRemote workHome officeFlexible hours$204k - $259k
...dynamics, and state-of-the-art Generative AI to create a training ground for the Waymo Driver. The Simulator Evaluation team faces the ultimate data challenge:... ...world is "real"? We are looking for a Senior Software Engineer to build the metrics and systems that grade...SeniorFull timeRemote work- ...Role Overview Design and evaluate high-quality datasets and evaluations that advance... ...corrections across multiple languages, assess AI-generated implementations for... ...measure model capabilities across the software engineering lifecycle. Key Responsibilities Curate...SeniorFor contractorsRemote work10 hours per weekFlexible hours
- ...development of high-quality datasets and evaluation pipelines that improve and benchmark large... ...language models for code generation and software engineering tasks. You will curate and author reference code, evaluate and refine AI-generated solutions across multiple programming...SeniorFull timeFor contractorsRemote work10 hours per weekFlexible hours
$200k - $250k
...About Wizard AI At Wizard AI, we’re building... ...and we’re looking for a Senior MLOps Engineer to help us run them reliably... ...— across a custom-built inference platform powering a live... ...years of experience in software, ML, platform, or infrastructure engineering, with hands-...SeniorFull timeRemote workFlexible hours- ...Description We're building a dataset to evaluate AI coding agents - how well a model... ...plan, you create a company: codebase, infrastructure, context (conversations, documentation... ...involvement. Qualifications ~Strong software engineering experience: 8+ years in any modern...SeniorTemporary workFreelanceRemote workFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Software Engineer, AI Infrastructure - LVM Inference & Evaluation. Be the first to apply!
- software engineer internship Remote
- software development engineer aws Remote
- software developer internship no experience Remote
- real time software engineer Remote
- financial software developer Remote
- oracle software engineer Remote
- part time software developer Remote
- graduate software developer Remote
- software engineer travel Remote
- experienced software developer Remote


