Perception Deployment Engineer - Model Deployment & Optimization
$199k - $270kZoox Inc.
The Perception team is pioneering the development of a multi-modality foundation model to drive the next generation of autonomous system intelligence.
As a Perception Deployment Engineer, you will focus on bringing highly efficient, production-ready large-scale models to our on-vehicle stack. We are looking for experts with hands-on experience in compressing, accelerating, and deploying complex computer vision or foundation models for power- and thermal-constrained vehicle SOCs. You will optimize the ML models, write custom CUDA kernels, and build highly concurrent inference code to ensure real-time, deterministic execution on edge devices.
\n In this role, you will:-
Design and develop production-level, low latency, and memory-safe C++ and CUDA code for real-time perception algorithms on vehicle systems.
-
Optimize large-scale models (Multi-Modal Sensor Fusion models, LLMs, VLMs) using advanced quantization (PTQ, QAT), pruning, mixed-precision inference frameworks.
-
Architect and implement model conversion and compilation pipelines using TensorRT for edge deployment.
-
Perform rigorous parity checking, accuracy recovery, and latency benchmarking between PyTorch frameworks and compiled edge binaries.
-
Develop and optimize custom ML OPs and TensorRT Plugins with efficient CUDA kernels to minimize latency and maximize memory bandwidth on AI accelerators.
-
Production-level C++ (14/17/20) and Python programming skills, with experience developing concurrent, memory-safe, real-time inference code for edge devices.
-
Deep expertise in model compression technologies (e.g., model quantization such as PTQ and QAT) and mixed-precision inference frameworks (INT8, FP8, BF16/FP16).
-
Proven experience optimizing large-scale models (Multi-Modal Sensor Fusion models, LLMs, VLMs/VLAs) utilizing Efficient Attention mechanisms (e.g., FlashAttention, Linear Attention), KV-cache optimization (e.g., PagedAttention.
-
Extensive experience with model conversion/compilation pipelines (e.g., ONNX, TensorRT, torch.compile) and performing rigorous latency benchmark and model quality parity valuation.
-
Proficiency in low-level programming for AI accelerators, specifically developing and optimizing custom ML OPs and TensorRT Plugins with efficient CUDA kernel implementations.
-
Familiarity with SOTA autonomous driving perception algorithms (temporal 3D object detection, BEV, 3D Occupancy Networks) and multi-modal sensor processing (Vision, LiDAR, Radar).
-
Experience with end-to-end autonomous driving paradigms (VLM/VLA models, Foundation models) and edge deployment technologies (e.g., TensorRT-LLM).
$199,000 - $270,000 a year
Base Salary Range
There are three major components to compensation for this position: salary, Amazon Restricted Stock Units (RSUs), and Zoox Stock Appreciation Rights. A sign-on bonus may be offered as part of the compensation package. The listed range applies only to the base salary. Compensation will vary based on geographic location and level. Leveling, as well as positioning within a level, is determined by a range of factors, including, but not limited to, a candidate's relevant years of experience, domain knowledge, and interview performance. The salary range listed in this posting is representative of the range of levels Zoox is considering for this position.
Zoox also offers a comprehensive package of benefits, including paid time off (e.g. sick leave, vacation, bereavement), unpaid time off, Zoox Stock Appreciation Rights, Amazon RSUs, health insurance, long-term care insurance, long-term and short-term disability insurance, and life insurance.
\nAbout Zoox
Zoox is developing the first ground-up, fully autonomous vehicle fleet and the supporting ecosystem required to bring this technology to market. Sitting at the intersection of robotics, machine learning, and design, Zoox aims to provide the next generation of mobility-as-a-service in urban environments. We’re looking for top talent that shares our passion and wants to be part of a fast-moving and highly execution-oriented team.
Follow us on LinkedIn
Accommodations
If you need an accommodation to participate in the application or interview process please reach out to View email address on aiapply.co or your assigned recruiter.
A Final Note:
You do not need to match every listed expectation to apply for this position. Here at Zoox, we know that diverse perspectives foster the innovation we need to be successful, and we are committed to building a team that encompasses a variety of backgrounds, experiences, and skills.
$295.25k - $345.04k
...mission to connect a billion people with optimism and civility, and looking for... ...inferences per day across Discovery, Safety, Engine, and much more. As a Model Optimization engineer on ML Platform... ...tooling for model optimization and deployment. Collaborate with cross-...SuggestedFull timeWork experience placementH1bWork at officeLocal areaVisa sponsorshipMonday to Friday- ...Job Description Job Description The Perception team is pioneering the development of a multi-modality foundation model to drive the next generation of... ...system intelligence. As a Model Optimization & Deployment Engineer, you will focus on bringing highly efficient...SuggestedTemporary workRelocation package
$195.2k - $262.2k
...developers and enterprises from data and model training through to production deployment, without the cost and complexity... ...ML infrastructure. Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across...SuggestedFull timeTemporary workImmediate startRemote work$124k
Tesla AI Foundation Model Architect At Tesla AI, we'... ...scale. In addition, we deploy these models to edge hardware... ...-making. You will optimize inference latency,... ...AI compiler, inference engine, and silicon teams to ensure... ...Collaborate across perception, planning, robotics, digital...SuggestedHourly payFull timeTemporary workImmediate startFlexible hours$150k
...highly motivated, and focused on engineering excellence. This... ...You will join the Grok Voice Model team to help build the world... ...from prototype to global-scale deployment for stable, low-latency, delightful... ...reinforcement learning, and optimizations for accuracy, factuality,...SuggestedTemporary work- ...Description Job Description As the Manager of Model Validation & Verification (VnV) for Behavior Autonomy, you will lead an engineering and data science team responsible for... ...dataset and evaluation pipelines, and optimize runtime and compute costs. Executive Communication...Temporary workRelocation package
$195.2k - $262.2k
...developers and enterprises from data and model training through to production deployment, without the cost and complexity... ...ML infrastructure. Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across...Full timeTemporary workImmediate startRemote work$260k - $350k
...this leader will own the finance partnership with AiDASH’s R&D (Engineering, Product, Data Science, and AI Data Ops), and Professional... ...with procurement strategy, and owning the corporate operating model. They will help determine where the company invests, how products...Full timeWork at officeImmediate startFlexible hours2 days per week3 days per week- ...environment. Roles and Responsibilities : Evaluates new desktop software technology and management tools. Recommends solutions for deployment in the desktop environment. Designs, builds, and deploys OS images, applications, and desktop environment modifications using...Relocation
- Genentech, a member of the Roche group, is seeking a Patient Experience Operating Model Designer to join Public Affairs & Access in South San Francisco. The role focuses on defining people, process, governance, and capability models to enable scalable delivery of innovative...
- Genentech is seeking a Patient Experience Operating Model Designer to activate innovative patient experiences designed by our Experience Design team. You will define the people, process, governance, and capability models required for scalable delivery, partnering with Solution...
$168k - $268.4k
...never been made means doing what's never been done. If you're an engineer, scientist, or builder who thrives on problems no one has... ...Together, we are building purpose-built foundation and frontier AI models trained on Lilly data at scale, tightening the feedback loop between...Full timeRemote workFlexible hours2 days per week$180k
...small, highly motivated, and focused on engineering excellence. This organization is for individuals... ...As a multimodal engineer on the Imagine Model Team, you will develop cutting-edge AI... ...learning systems. Ability to deliver optimal end-to-end user experiences. Hands-on...Temporary work$18 - $20 per hour
...support for Google Workspace, MS Office, Box, Logitech video systems, and Zoom Perform basic mobile device management including deployment, collection, inventory, and reset/wipe procedures Customer Service Deliver prompt and courteous user support via: Wolken ticketing...Hourly payRemote workMonday to Friday- US Tech Solutions is a global staff augmentation firm providing a wide-range of talent on-demand and total workforce solutions. To know more about US Tech Solutions, please visit our website We are constantly on the lookout for professionals to fulfill the staffing needs...Work experience placementWork at office
- ...for our clients. That includes installing, diagnosing, repairing, maintaining, and upgrading hardware and equipment while ensuring optimal workstation performance. You will also troubleshoot problem areas (in person and/or remote) in a timely and accurate fashion and provide...Work at officeRemote work
$55 - $65 per hour
.... About the role: We are seeking an experienced Desktop Engineer to manage and support the organization’s endpoint and workplace... ...including procurement coordination, enrollment, provisioning, deployment, refresh, reassignment, and secure retirement. Monitor...Hourly payContract workFor contractorsCasual workInternshipWork at officeLocal areaRemote work$145.59k
...10x better than anything that exists today. Our scientists, engineers, sales executives, and visionaries are united by an unwavering... ...and patient information. Design, document, and continuously optimize IT support processes to enhance service efficiency and customer...Worldwide- Job Post Requirements / Qualifications For more information on Requirements/Qualifications, please contact the employer. Comments and Other Information For more information on Comments and Other Information, please contact the employer. San Mateo High School...Weekend work
- Job Summary Breakfast service = 1.5 hours in morning (approximately 7:30am-9am) Lunch service = 2 hours in afternoon (approximately 11:00am-1pm) Requirements / Qualifications Edjoin Classified Application including three references. Comments and Other Information...Day shift
$180k
...knowledge. Our team is small, highly motivated, and focused on engineering excellence. This organization is for individuals who... ...teammates. ABOUT THE ROLE: You will work on the most critical modeling challenges at any given time. You will get clarity on your...Temporary work$35 - $40 per hour
...issues. Record all changes to all hardware assets in the Asset Management System (AMS). Configure, test for quality assurance, deploy and support computers, smartphones, printers and other hardware provided by the client. Support telecoms and voicemail moves,...Contract workTemporary workWork at officeLocal areaRemote workShift work- Administrative Secretary Responsibilities: Under direction, to serve as secretary to an administrative official usually at the coordinator level, relieving him/her of clerical and routine administrative details; to perform work of above-average difficulty requiring ...Work at office
- Redwood City School District is seeking a Clerical Substitute. The district takes pride in its dedicated workforce and strong connection with students, families, and the community. The School Office Assistant I classification performs a variety of clerical assistance duties...Work at officeLocal area
- Job Description Job Description REDWOOD CITY SCHOOL DISTRICT CLASS TITLE: SCHOOL OFFICE ASSISTANT I RCSD takes pride in our dedicated workforce and strong connection with our students, families, and community. Our highly dedicated and skilled team of professionals...ApprenticeshipWork at officeLocal area
- ...Support Technician for a TOP community hospital! As the Desktop Support Tech you will work alongside the Service Desk and Desktop Engineering teams with hardware, software, and other services to deliver technology infrastructure services to their esteemed community. This...
$35 per hour
Schedule: Monday-Friday, 8:00 AM-5:00 PM Contract: 3+ Month Contract with Likely Extension Pay: Up to $35/hour We are seeking an experienced Desktop Support Technician to provide hands-on, onsite IT support in Palo Alto, CA. This is a 3+ month contract with a strong ...Contract workTemporary workMonday to Friday- Cafeteria Helper Under the supervision of the Food Service Production and Inventory Management Supervisor, the Cafeteria Helper assists in the collection of cafeteria monies, completes required reports, and assists in the kitchen as needed. This is an on-call, as-needed...
- ...Support Technician – L2 Location: South San Francisco, CA Summary: We are looking to hire a skilled L2 Technician (Hardware Deployment and Repair) to assist our clients with computer hardware and software issues. You will be required to work on-site or via remote...Work at officeRemote workFlexible hours
- ...results and SLA measurements. • Respond to assignments usually involving the support of the field service technicians repairing or deploying hardware and peripherals at cyberCSI's customer repair facility. • Prioritize assignments is order of importance. • Update the...Local areaFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Perception Deployment Engineer - Model Deployment & Optimization. Be the first to apply!




