Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Perception Deployment Engineer - Model Deployment & Optimization

$199k - $270k
Full-time

Zoox Inc.

The Perception team is pioneering the development of a multi-modality foundation model to drive the next generation of autonomous system intelligence.

As a Perception Deployment Engineer, you will focus on bringing highly efficient, production-ready large-scale models to our on-vehicle stack. We are looking for experts with hands-on experience in compressing, accelerating, and deploying complex computer vision or foundation models for power- and thermal-constrained vehicle SOCs. You will optimize the ML models, write custom CUDA kernels, and build highly concurrent inference code to ensure real-time, deterministic execution on edge devices.

\n In this role, you will:
  • Design and develop production-level, low latency, and memory-safe C++ and CUDA code for real-time perception algorithms on vehicle systems.

  • Optimize large-scale models (Multi-Modal Sensor Fusion models, LLMs, VLMs) using advanced quantization (PTQ, QAT), pruning, mixed-precision inference frameworks.

  • Architect and implement model conversion and compilation pipelines using TensorRT for edge deployment.

  • Perform rigorous parity checking, accuracy recovery, and latency benchmarking between PyTorch frameworks and compiled edge binaries.

  • Develop and optimize custom ML OPs and TensorRT Plugins with efficient CUDA kernels to minimize latency and maximize memory bandwidth on AI accelerators.

Qualifications:
  • Production-level C++ (14/17/20) and Python programming skills, with experience developing concurrent, memory-safe, real-time inference code for edge devices.

  • Deep expertise in model compression technologies (e.g., model quantization such as PTQ and QAT) and mixed-precision inference frameworks (INT8, FP8, BF16/FP16).

  • Proven experience optimizing large-scale models (Multi-Modal Sensor Fusion models, LLMs, VLMs/VLAs) utilizing Efficient Attention mechanisms (e.g., FlashAttention, Linear Attention), KV-cache optimization (e.g., PagedAttention.

  • Extensive experience with model conversion/compilation pipelines (e.g., ONNX, TensorRT, torch.compile) and performing rigorous latency benchmark and model quality parity valuation.

  • Proficiency in low-level programming for AI accelerators, specifically developing and optimizing custom ML OPs and TensorRT Plugins with efficient CUDA kernel implementations.

Bonus Qualifications:
  • Familiarity with SOTA autonomous driving perception algorithms (temporal 3D object detection, BEV, 3D Occupancy Networks) and multi-modal sensor processing (Vision, LiDAR, Radar).

  • Experience with end-to-end autonomous driving paradigms (VLM/VLA models, Foundation models) and edge deployment technologies (e.g., TensorRT-LLM).

\n

$199,000 - $270,000 a year

Base Salary Range

There are three major components to compensation for this position: salary, Amazon Restricted Stock Units (RSUs), and Zoox Stock Appreciation Rights. A sign-on bonus may be offered as part of the compensation package. The listed range applies only to the base salary. Compensation will vary based on geographic location and level. Leveling, as well as positioning within a level, is determined by a range of factors, including, but not limited to, a candidate's relevant years of experience, domain knowledge, and interview performance. The salary range listed in this posting is representative of the range of levels Zoox is considering for this position.

Zoox also offers a comprehensive package of benefits, including paid time off (e.g. sick leave, vacation, bereavement), unpaid time off, Zoox Stock Appreciation Rights, Amazon RSUs, health insurance, long-term care insurance, long-term and short-term disability insurance, and life insurance.

\n

About Zoox

Zoox is developing the first ground-up, fully autonomous vehicle fleet and the supporting ecosystem required to bring this technology to market. Sitting at the intersection of robotics, machine learning, and design, Zoox aims to provide the next generation of mobility-as-a-service in urban environments. We’re looking for top talent that shares our passion and wants to be part of a fast-moving and highly execution-oriented team.

Follow us on LinkedIn

Accommodations

If you need an accommodation to participate in the application or interview process please reach out to View email address on aiapply.co or your assigned recruiter.

A Final Note:

You do not need to match every listed expectation to apply for this position. Here at Zoox, we know that diverse perspectives foster the innovation we need to be successful, and we are committed to building a team that encompasses a variety of backgrounds, experiences, and skills.

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Perception Deployment Engineer - Model Deployment & Optimization in Foster, CA vacancy
  • $295.25k - $345.04k

     ...mission to connect a billion people with optimism and civility, and looking for...  ...inferences per day across Discovery, Safety, Engine, and much more. As a Model Optimization engineer on ML Platform...  ...tooling for model optimization and deployment. Collaborate with cross-... 
    Suggested
    Full time
    Work experience placement
    H1b
    Work at office
    Local area
    Visa sponsorship
    Monday to Friday

    Roblox

    San Mateo, CA
    3 days ago
  •  ...Job Description Job Description The Perception team is pioneering the development of a multi-modality foundation model to drive the next generation of...  ...system intelligence. As a Model Optimization & Deployment Engineer, you will focus on bringing highly efficient... 
    Suggested
    Temporary work
    Relocation package

    Zoox

    Foster, CA
    25 days ago
  • $195.2k - $262.2k

     ...developers and enterprises from data and model training through to production deployment, without the cost and complexity...  ...ML infrastructure. Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across... 
    Suggested
    Full time
    Temporary work
    Immediate start
    Remote work

    Nebius

    Palo Alto, CA
    2 days ago
  • $124k

    Tesla AI Foundation Model Architect At Tesla AI, we'...  ...scale. In addition, we deploy these models to edge hardware...  ...-making. You will optimize inference latency,...  ...AI compiler, inference engine, and silicon teams to ensure...  ...Collaborate across perception, planning, robotics, digital... 
    Suggested
    Hourly pay
    Full time
    Temporary work
    Immediate start
    Flexible hours

    Tesla

    Palo Alto, CA
    4 days ago
  • $150k

     ...highly motivated, and focused on engineering excellence. This...  ...You will join the Grok Voice Model team to help build the world...  ...from prototype to global-scale deployment for stable, low-latency, delightful...  ...reinforcement learning, and optimizations for accuracy, factuality,... 
    Suggested
    Temporary work

    SpaceXAI

    Palo Alto, CA
    29 days ago
  •  ...Description Job Description As the Manager of Model Validation & Verification (VnV) for Behavior Autonomy, you will lead an engineering and data science team responsible for...  ...dataset and evaluation pipelines, and optimize runtime and compute costs. Executive Communication... 
    Temporary work
    Relocation package

    Zoox

    Foster, CA
    18 days ago
  • $195.2k - $262.2k

     ...developers and enterprises from data and model training through to production deployment, without the cost and complexity...  ...ML infrastructure. Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across... 
    Full time
    Temporary work
    Immediate start
    Remote work

    Nebius

    Palo Alto, CA
    2 days ago
  • $260k - $350k

     ...this leader will own the finance partnership with AiDASH’s R&D (Engineering, Product, Data Science, and AI Data Ops), and Professional...  ...with procurement strategy, and owning the corporate operating model. They will help determine where the company invests, how products... 
    Full time
    Work at office
    Immediate start
    Flexible hours
    2 days per week
    3 days per week

    AiDash

    Palo Alto, CA
    4 days ago
  •  ...environment. Roles and Responsibilities : Evaluates new desktop software technology and management tools. Recommends solutions for deployment in the desktop environment. Designs, builds, and deploys OS images, applications, and desktop environment modifications using... 
    Relocation

    Boston Private

    San Mateo, CA
    2 days ago
  • Genentech, a member of the Roche group, is seeking a Patient Experience Operating Model Designer to join Public Affairs & Access in South San Francisco. The role focuses on defining people, process, governance, and capability models to enable scalable delivery of innovative... 

    F. Hoffmann-La Roche AG

    South San Francisco, CA
    4 days ago
  • Genentech is seeking a Patient Experience Operating Model Designer to activate innovative patient experiences designed by our Experience Design team. You will define the people, process, governance, and capability models required for scalable delivery, partnering with Solution... 

    Genentech

    South San Francisco, CA
    3 days ago
  • $168k - $268.4k

     ...never been made means doing what's never been done. If you're an engineer, scientist, or builder who thrives on problems no one has...  ...Together, we are building purpose-built foundation and frontier AI models trained on Lilly data at scale, tightening the feedback loop between... 
    Full time
    Remote work
    Flexible hours
    2 days per week

    Eli Lilly

    South San Francisco, CA
    3 days ago
  • $180k

     ...small, highly motivated, and focused on engineering excellence. This organization is for individuals...  ...As a multimodal engineer on the Imagine Model Team, you will develop cutting-edge AI...  ...learning systems. Ability to deliver optimal end-to-end user experiences. Hands-on... 
    Temporary work

    SpaceXAI

    Palo Alto, CA
    12 days ago
  • $18 - $20 per hour

     ...support for Google Workspace, MS Office, Box, Logitech video systems, and Zoom Perform basic mobile device management including deployment, collection, inventory, and reset/wipe procedures Customer Service Deliver prompt and courteous user support via: Wolken ticketing... 
    Hourly pay
    Remote work
    Monday to Friday

    GDR Group

    Palo Alto, CA
    19 hours ago
  • US Tech Solutions is a global staff augmentation firm providing a wide-range of talent on-demand and total workforce solutions. To know more about US Tech Solutions, please visit our website We are constantly on the lookout for professionals to fulfill the staffing needs...
    Work experience placement
    Work at office

    US Tech Solutions

    San Mateo, CA
    4 days ago
  •  ...for our clients. That includes installing, diagnosing, repairing, maintaining, and upgrading hardware and equipment while ensuring optimal workstation performance. You will also troubleshoot problem areas (in person and/or remote) in a timely and accurate fashion and provide... 
    Work at office
    Remote work

    The Rockridge Group

    South San Francisco, CA
    a month ago
  • $55 - $65 per hour

     .... About the role: We are seeking an experienced Desktop Engineer to manage and support the organization’s endpoint and workplace...  ...including procurement coordination, enrollment, provisioning, deployment, refresh, reassignment, and secure retirement. Monitor... 
    Hourly pay
    Contract work
    For contractors
    Casual work
    Internship
    Work at office
    Local area
    Remote work

    Allogene Therapeutics

    South San Francisco, CA
    10 days ago
  • $145.59k

     ...10x better than anything that exists today. Our scientists, engineers, sales executives, and visionaries are united by an unwavering...  ...and patient information. Design, document, and continuously optimize IT support processes to enhance service efficiency and customer... 
    Worldwide

    BillionToOne

    Menlo Park, CA
    1 day ago
  • Job Post Requirements / Qualifications For more information on Requirements/Qualifications, please contact the employer. Comments and Other Information For more information on Comments and Other Information, please contact the employer. San Mateo High School...
    Weekend work

    San Mateo High School

    San Mateo, CA
    4 days ago
  • Job Summary Breakfast service = 1.5 hours in morning (approximately 7:30am-9am) Lunch service = 2 hours in afternoon (approximately 11:00am-1pm) Requirements / Qualifications Edjoin Classified Application including three references. Comments and Other Information...
    Day shift

    San Carlos School District

    San Carlos, CA
    1 day ago
  • $180k

     ...knowledge. Our team is small, highly motivated, and focused on engineering excellence. This organization is for individuals who...  ...teammates. ABOUT THE ROLE: You will work on the most critical modeling challenges at any given time. You will get clarity on your... 
    Temporary work

    SpaceXAI

    Palo Alto, CA
    7 days ago
  • $35 - $40 per hour

     ...issues. Record all changes to all hardware assets in the Asset Management System (AMS). Configure, test for quality assurance, deploy and support computers, smartphones, printers and other hardware provided by the client. Support telecoms and voicemail moves,... 
    Contract work
    Temporary work
    Work at office
    Local area
    Remote work
    Shift work

    TEKsystems

    Redwood City, CA
    1 day ago
  • Administrative Secretary Responsibilities: Under direction, to serve as secretary to an administrative official usually at the coordinator level, relieving him/her of clerical and routine administrative details; to perform work of above-average difficulty requiring ...
    Work at office

    Palo Alto Unified School District

    Palo Alto, CA
    19 hours ago
  • Redwood City School District is seeking a Clerical Substitute. The district takes pride in its dedicated workforce and strong connection with students, families, and the community. The School Office Assistant I classification performs a variety of clerical assistance duties...
    Work at office
    Local area

    Redwood City School District

    Redwood City, CA
    19 hours ago
  • Job Description Job Description REDWOOD CITY SCHOOL DISTRICT CLASS TITLE: SCHOOL OFFICE ASSISTANT I RCSD takes pride in our dedicated workforce and strong connection with our students, families, and community. Our highly dedicated and skilled team of professionals...
    Apprenticeship
    Work at office
    Local area

    Redwood City School District

    Redwood City, CA
    a month ago
  •  ...Support Technician for a TOP community hospital! As the Desktop Support Tech you will work alongside the Service Desk and Desktop Engineering teams with hardware, software, and other services to deliver technology infrastructure services to their esteemed community. This... 

    Option 1 Staffing Services, Inc.

    Palo Alto, CA
    2 days ago
  • $35 per hour

    Schedule: Monday-Friday, 8:00 AM-5:00 PM Contract: 3+ Month Contract with Likely Extension Pay: Up to $35/hour We are seeking an experienced Desktop Support Technician to provide hands-on, onsite IT support in Palo Alto, CA. This is a 3+ month contract with a strong ...
    Contract work
    Temporary work
    Monday to Friday

    Pomeroy

    Palo Alto, CA
    1 day ago
  • Cafeteria Helper Under the supervision of the Food Service Production and Inventory Management Supervisor, the Cafeteria Helper assists in the collection of cafeteria monies, completes required reports, and assists in the kitchen as needed. This is an on-call, as-needed...

    San Bruno Park School District

    San Bruno, CA
    8 hours ago
  •  ...Support Technician – L2 Location: South San Francisco, CA Summary: We are looking to hire a skilled L2 Technician (Hardware Deployment and Repair) to assist our clients with computer hardware and software issues. You will be required to work on-site or via remote... 
    Work at office
    Remote work
    Flexible hours

    The Rockridge Group

    South San Francisco, CA
    9 days ago
  •  ...results and SLA measurements. • Respond to assignments usually involving the support of the field service technicians repairing or deploying hardware and peripherals at cyberCSI's customer repair facility. • Prioritize assignments is order of importance. • Update the... 
    Local area
    Flexible hours

    The Rockridge Group

    South San Francisco, CA
    25 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Perception Deployment Engineer - Model Deployment & Optimization. Be the first to apply!