Principal On-Device AI Engineer for Real-Time Game Inference
Unity
Unity is seeking a Principal Engineer for On-Device AI Inference & Systems to advance multi-modal models (transformers and diffusion networks) running entirely within our game engine runtime. You will own the inference and integration stack end-to-end—from exporting and optimizing trained checkpoints to kernel-level tuning and shipping features that run at interactive frame rates within strict memory and power budgets. #J-18808-Ljbffr Unity
$190k - $250k
A leading financial technology firm in California seeks an AI Inference engineer to join its team. The role involves developing APIs for AI inference, improving system reliability, and optimizing LLM performance. Required qualifications include experience with ML systems...Suggested- ...seeking a Senior Embedded Machine Learning Engineer to own end-to-end deployment of... ...build the C++ infrastructure that hosts inference on devices. You’ll work closely with CVML, embedded... ..., and thermal budgets while ensuring real-time performance in the field. #J-18808-...Suggested
- Illumio is hiring a Senior Software Engineer in Sunnyvale, California, to architect high-scale distributed systems, focusing on data processing and real-time analytics. The role requires expertise in backend engineering with Java, Python, or Go, and strong knowledge in...Suggested
$180k
...firm based in Palo Alto is seeking to hire an experienced engineer to work on multimodal AI systems. The ideal candidate will have hands-on... ...developing data pipelines, and advancing capabilities in real-time interactions. The role offers competitive compensation ranging...Suggested- ...observability, and RCA infrastructure that makes Sage Care’s AI assistant Role Overview Own and build the full... ...Sage Care’s AI assistant trustworthy and debuggable —in real time and post-call. This engineer builds the visibility layer across telephony, transcription...SuggestedImmediate start
- Sanas is building real-time speech and language models deployed on-premise... ...delivering low-latency, high-throughput AI for multi-node GPU workloads. As a Senior Engineer, you will shape core... ...performance optimizations, and own the inference engine to scale research and...
$228.7k - $309.4k
We are looking for a Principal Applied Scientist to drive the research and development of real-time multimodal conversational AI. You will operate across two focus areas: advancing... ...roadmap, and work closely with inference engineers to ensure your models are designed...PrincipalLocal areaFlexible hours- ...computing experiences—from AI and data centers, to PCs, gaming and embedded systems.... ...collaboration, we believe real progress comes from bold ideas... ...customers. Workload Performance Engineering: Lead the profiling,... ...operating systems (OS) and device driver development is a...PrincipalGames
- ...individual can thrive.Job DescriptionThe AI Inference Engineer plays a critical role in the AI... ...centers to resource-constrained edge devices—with a strong emphasis on maximizing throughput... ...-scaling architectures for online (real-time) and batch inference pipelines,...Full timeLocal areaImmediate start
- DataVisor is hiring a Software Engineer, Artificial Intelligence to architect the Intelligence Layer and Data Consortium. You will design... ...and operate distributed, production-grade services ingesting real-time signals from millions of users and enable agentic flows with...
$148.7k - $297.3k
...products in diagnostics, medical devices, nutritionals and branded generic... ...Overview:Abbott Vascular is seeking a Principal AI/ML Engineer to develop advanced machine... ...clinical use cases.Optimize models for real-time or near-real-time inference in embedded or product...Principal- ...centers and AI compute is rapidly... ...of inference compute. The... ...deliver up to 100 times the energy efficiency... ...seeking a Principal Platform... ...measured boot, device attestation,... ...subsystems on real silicon, can... ...and firmware engineers, and knows the... ...hardware. Work on game-changing...PrincipalGamesFull time
$117.7k - $221.4k
...data from large-scale real-world sensor streams. The... ...efficient for embodied AI systems. We believe the... ...reflects how Cola engineers think: build durable intermediate... ..., featurization, and inference foundations that power... ...to the location three times a week {or other...Full timeLocal areaRemote workWork from homeRelocation packageFlexible hours- ...builds the world's largest AI chip, 56 times larger than GPUs. This architecture... ...-leading training and inference speeds; over 10 times... ...AI applications, unlocking real-time iteration and increasing... ...About The RoleWe're hiring a Principal Engineer for our Inference Cloud...Principal
$248k - $379.5k
...Leader for their GeForce NOW cloud team in California. You will oversee the development of advanced diagnostic and AI/ML solutions necessary for real-time analytics, leading and mentoring teams. The ideal candidate will possess extensive experience in AI/ML, along with...Principal- ...designing, developing, and maintaining advanced LabVIEW-based systems, particularly within the medical device domain. This role involves extensive work with LabVIEW FPGA, real-time systems, and hardware integration. The position also requires strong qualifications in database...
- Cerebras Systems, Inc. is seeking a Software Engineer in Sunnyvale, California to enhance high-performance, low-latency inference infrastructure. This role involves deploying... ...and Python. Join us in building revolutionary AI technology! #J-18808-Ljbffr CEREBRAS SYSTEMS...
$180k
...California, is seeking a talented individual to join the Omni team and develop next-generation AI experiences that go beyond text. You will be responsible for advancing real-time multimodal intelligence across audio, video, and world modeling if you have a proven track...- Inworld AI, based in Mountain View, California, seeks a Product Lead to drive the development... ...will collaborate with AI researchers and engineers to turn cutting-edge capabilities into... ..., shaping how the world builds real-time AI applications. #J-18808-Ljbffr Inworld...
- Bellota Labs in Redwood City, CA is seeking a Senior/Principal Software Engineer to shape the architecture and build the WPTHome platform. You’ll... ...the Unity/C# client and backend services, delivering real‑time, low‑latency gameplay for a live, real‑money product. You’...Games
$190k - $250k
Location San Francisco Employment Type Full time Location Type Hybrid Department AI We are looking for an AI Inference engineer to join our growing team. Our current stack is... ...scale deployment of machine learning models for real-time inference. Responsibilities Develop APIs...Full time- Bellota Labs is hiring a Senior/Principal Software Engineer to help shape WPTHome with architecture, design, and scalable development. You will work on real-time, low-latency game services spanning Unity/C# clients and backend systems, including matchmaking, wallet, and...Games
$151.8k - $332.2k
What you can expect We are looking for an AI Inference Engineer with a solid background in speech recognition and model inference. In this role... ...integrating frameworks such as PyTorch and TensorFlow for real-time inference.Salary Range or On Target Earnings:Minimum:$151,8...Full timeWork at officeRemote work$193.93k - $291.15k
...the most immediate and profound opportunity for AI to drive positive change in the physical world. Safer streets, more time for what matters, and easier access to the... ...logistics fleets to personal vehicles.With years of real-world deployment experience and a flexible,...Immediate startRemote workFlexible hours- ...with the ultimate goal of enabling human life on Mars.SOFTWARE ENGINEER, INFERENCE (AI DATA ENGINEERING)The application software team is the... ...batching Optimize latency and throughput of model inference under real production workloads, including low-level GPU kernel work,...Permanent employmentTemporary workRemote workWorldwideWeekend work
- ...is searching for a hands-on AI Compiler Engineer who thrives at the convergence... ...of our silicon for real-world impact. Innovation here... ...skillsSolid understanding of AI inference workloads (CNNs, transformers... ...Austin (Oakhill, Office); San Jose (Holger Way)Type: Full timePrincipalFull timeWork at officeLocal area
$182.5k - $260.5k
...networking for the cloud and AI era. We secure and... ...cloud, data, and AI in real time, everywhere. Thousands... ...platform, its Zero Trust Engine, and the powerful... ...Scientist, you own the inference and optimization layer... ...cpp, or MLX/CoreML). On-device or edge inference experience...Principal$229.9k - $262.4k
...Sr. Lead AI Engineer (FM Hosting, LLM Inference) Overview: At Capital One, we are creating responsible and reliable AI systems, changing banking... ...an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in...Full timePart timeLocal area$229.9k - $262.4k
Sr. Lead AI Engineer (Inference Optimization, FM hosting, AI Platform) Overview: At Capital One, we are creating responsible and reliable AI... ...been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in...Full timePart timeLocal area- NVIDIA Corporation is seeking a System Software Engineer for Vision AI to craft high‑performance pipelines that process video, image, and 3D data in real‑time across edge and cloud environments. You will collaborate with perception, simulation, and platform teams to translate...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Principal On-Device AI Engineer for Real-Time Game Inference. Be the first to apply!
- principal network engineer Mountain View, CA
- principal engineer Mountain View, CA
- principal infrastructure engineer Mountain View, CA
- senior civil engineer project manager Mountain View, CA
- senior chief engineer Mountain View, CA
- engineering director Mountain View, CA
- senior director engineering Mountain View, CA
- director systems engineering Mountain View, CA
- principal developer Mountain View, CA
- general engineer Mountain View, CA

