Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

AI Infrastructure Engineer

NIO USA, INC

Job Description

Job Description

About NIO

NIO Inc. is a pioneer and a leading company in the global smart electric vehicle market. Founded in November 2014, NIO aspires to shape a sustainable and brighter future with the mission of “Blue Sky Coming”.

NIO envisions itself as a user enterprise where innovative technology meets experience excellence. NIO designs, develops, manufactures and sells smart electric vehicles, driving innovations in next-generation core technologies. NIO distinguishes itself through continuous technological breakthroughs and innovations, exceptional products and services, and a community for shared growth.

NIO provides premium smart electric vehicles under the NIO brand, premium smart electric vehicles for families through the ONVO brand, and small smart high-end electric cars with the FIREFLY brand.

About the Position

We are looking for a senior AI Inference Infrastructure Software Engineer with strong hands-on experience building, optimizing, and deploying high-performance, scalable inference systems. This position is focused on designing, implementing, and delivering production-grade software that powers real-world applications of Large Language Models (LLMs) and Vision-Language Models (VLMs).

This is an exciting opportunity for an engineer who thrives at the intersection of AI systems, hardware acceleration, and large-scale robust deployment, and who wants to see their contributions ship in production, at scale.

In this role, you will directly shape the architecture, roadmap and performance of AI capabilities of our AIOS platform, driving innovations that make LLM/VLM systems fast, efficient, and scalable across cloud, edge, and hybrid edge-cloud environments. You will work closely with system, hardware, and product teams to deliver high-performance inference kernels for hardware accelerators, design scalable inference serving systems, and integrate optimizations such tensor parallelism and custom kernels into production pipelines. Your work will have immediate impact, powering intelligent automotive systems in the next generation of electric vehicles.

Roles and Responsibilities

  • Design and implement high-performance, scalable inference systems for LLMs and VLMs across cloud, edge, and edge-cloud hybrid platforms.

  • Develop and optimize custom kernels and operators for specific hardware accelerators (GPU, NPU, DSP, etc.), improving throughput, latency, and memory efficiency.

  • Integrate advanced optimization techniques such as KV-cache management, tensor/model parallelism, quantization, and memory-efficient execution into production inference systems.

  • Partner with system and hardware teams to ensure tight hardware-software integration and optimal performance across diverse compute environments.

  • Translate architectural requirements into robust, maintainable, production-ready software that meets performance, safety, and reliability standards.

  • Define and drive the evolution roadmap for LLM/VLM inference in the AIOS stack, ensuring scalability and adaptability to new workloads.

  • Stay ahead of industry trends and competitor solutions, applying best practices from both AI and large-scale systems engineering.

Must Qualifications

  • 5+ years of hands-on software development experience in building and optimizing AI inference systems at scale.

  • Direct experience in LLM/VLM model internals, including Transformer-based architectures, inference bottlenecks, and optimization techniques.

  • Strong expertise in performance engineering: kernel development, parallelism strategies, memory optimization, and distributed inference systems.

  • Proficiency with GPU/NPU programming (CUDA, or vendor-specific SDKs), compiler toolchains, and deep learning frameworks (PyTorch, or TensorFlow).

  • Strong programming skills in C/C++, with a track record of delivering high-performance, production-grade software.

  • Solid foundation in computer architecture, systems programming (CPU/GPU pipelines, memory hierarchy, scheduling), and embedded systems.

  • BS/MS in Computer Science, Computer Engineering, or related technical field.

  • Excellent communication and collaboration skills, with the ability to work across cross-functional teams.

Preferred Qualifications

  • Master’s or PhD degree in Computer Science, Electrical/Computer Engineering, or related fields, plus 5 years industry experience

  • Experience building inference serving systems for large models, including batching, scheduling, caching, and load balancing.

  • Expertise in hardware-aware model optimization (e.g., kernel fusion, mixed precision, quantization, pruning).

  • Familiarity with edge and embedded AI, including real-time constraints and limited-resource optimization.

  • Contributions to widely used AI frameworks, libraries, or performance-critical software (open source or proprietary).

Compensation

  • Within the range, individual pay is determined by work location and additional factors, including job-related skills, experience, and relevant education or training.
  • Please note that the compensation details listed in US role postings reflect the base salary only. It does not include discretionary bonus, equity, or benefits.

Benefits:

Along with competitive pay, as a full-time NIO employee, you are eligible for the following benefits on the first day you join NIO:

  • Anthem Blue Cross, HSA, and Kaiser HMO medical plans with $0 for Employee Only Coverage.  
  • Dental (including orthodontic coverage) and vision plan.  Both provide options with a $0 paycheck contribution covering you and your eligible dependents.
  • Company Paid HSA (Health Savings Account) Contribution when enrolled in the High Deductible Anthem Blue Cross medical plan
  • Healthcare and Dependent Care Flexible Spending Accounts (FSA)
  • 401(k) with Brokerage Link option
  • Company paid Basic Life, AD&D, short-term and long-term disability insurance
  • Employee Assistance Program
  • Sick and Vacation time
  • 13 Paid Holidays a year
  • Paid Parental Leave for first 8 weeks at full pay (eligible after 90 days of employment with NIO)
  • Paid Disability Leave for first 6 weeks at full pay (eligible after 90 days of employment with NIO)
  • Voluntary benefits including: Voluntary Life and AD&D options for you, your spouse/domestic partner and dependent child(ren), pet insurance
  • Commuter benefits
  • Mobile Cell Phone Credit
  • Free lunch and snacks
  • Onsite gym
  • Employee discounts and perks program

We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.

Vacancy posted 8 days ago
Similar jobs that could be interesting for youBased on the AI Infrastructure Engineer in San Jose, CA vacancy
  •  ...next-generation computing experiences—from AI and data centers, to PCs, gaming and embedded...  ...THE ROLE: We are seeking a DevOps / Platform Engineer to join our team building and operating large-scale GPU compute infrastructure that powers AI and ML workloads. THE PERSON:... 
    Suggested

    AMD

    San Jose, CA
    4 days ago
  • $180k - $240k

     ...facilitating effortless integration into customers' logistics operations. About the role We are seeking a Senior AI Infrastructure Engineer to design, build, and scale the high-performance AI platform powering our autonomous driving models. While researchers focus... 
    Suggested
    Odd job
    Work at office

    Gatik AI

    Santa Clara, CA
    4 days ago
  • $184k - $287.5k

     ...leading cloud product that powers innovative AI research and developers. We focus on...  ...workloads, as well as developing scalable AI infrastructure services globally. We are seeking an AI infrastructure software engineer to join our team. You'll be instrumental in... 
    Suggested
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    2 days ago
  •  ...using it to some extent and at TransPerfect, we are no exception. We’re building an advanced voice processing platform that leverages AI for natural voice recognition, synthesis, and interaction. As part of the team, you’ll help design and develop scalable solutions... 
    Suggested
    Full time

    TransPerfect

    San Jose, CA
    2 days ago
  • $178k - $321k

     ...Audit function has an early but working AI-native capability: a multi-agent platform...  ...setting, and the governed data and AI infrastructure everything else depends on. We hire on demonstrated...  ...just implement it. This is a two-person engineering team: you deploy, debug, and hotfix your... 
    Suggested

    OKX

    San Jose, CA
    4 days ago
  • $184k - $287.5k

     ...tapping into the unlimited potential of AI to define the next era of computing. An...  ...Group (SCG) is seeking Senior AI Platform Engineers. They will set the technical direction...  ...foundational platforms at the intersection of ML infrastructure and large-scale systems, this is your... 
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $139.23k - $163.8k

     ...excel at—all from Day One.Job DescriptionJob SummaryThe Lead Engineer (Generative AI) is a senior technical role responsible for designing,...  ...high-throughput, low-latency AI workloadsLeverage modern infrastructure practices:Containerization (Docker)Orchestration (... 
    Work experience placement
    Local area
    3 days per week

    US Bank

    Cupertino, CA
    14 hours ago
  •  ...are builders, makers, and innovators helping enterprises move beyond AI experimentation and into real-world impact through enterprise AI platforms, products, and services. We combine deep engineering expertise with AI innovation to help clients modernize, build intelligent... 

    Publicis Media

    San Jose, CA
    4 days ago
  • $183.6k - $297k

     ...Execution, Integrity, and Inclusion. We weave AI into the fabric of everything we do and...  ...of integrating AI into cybersecurity infrastructure — building intelligent systems that...  ...incidents at scale. As a Principal Software Engineer, you will own the technical vision for... 
    Full time
    Work at office

    Palo Alto Networks

    Santa Clara, CA
    2 days ago
  • $274k - $304k

     ...Saviynt's AI-powered identity platform manages and governs human and non-human access...  ..., please visit AI Platform Engineer - Training & Inference Saviynt's AI...  ...SLMs and cloud LLMs • Build RL training infrastructure: define Flyte workflows for RL pipelines... 

    Saviynt

    Milpitas, CA
    1 day ago
  •  ...next-generation computing experiences—from AI and data centers, to PCs, gaming and...  ...THE ROLEWe are hiring AI / ML Platform Engineers to build the platform layer that makes AI...  ...reproducible. This role focuses on the infrastructure and platform systems that support large-... 

    AMD

    Santa Clara, CA
    3 days ago
  • $150k - $200k

     ...Business System which makes everything possible.The Backend & AI Platform Engineer is responsible for creating the backend platform that...  ...services and APIs that connect laboratory instruments, cloud infrastructure, and AI agents into a unified platform.Design scalable... 
    Full time
    Part time

    Danaher Corporation

    San Jose, CA
    3 days ago
  •  ...maintains a close and long-term relationship with our direct client. In support of their needs, we are looking for an  AI Devops Infrastructure Engineer/GPU Infrastructure Engineer. Job Title:  AI Devops Infrastructure Engineer/GPU Infrastructure Engineer Job... 
    Contract work

    Maxonic

    San Jose, CA
    2 days ago
  • $229.9k - $262.4k

    Sr. Lead AI Engineer (Gen AI Platform Services) Overview: At Capital One, we are creating responsible and reliable AI systems, changing...  ...customer experiences. Our investments in technology infrastructure and world-class talent — along with our deep experience in machine... 
    Full time
    Part time
    Local area

    Capital One Financial Corp

    San Jose, CA
    1 day ago
  • $229.9k - $262.4k

    Sr. Lead AI Engineer (GenAI Platform) Overview: At Capital One, we are creating responsible and reliable AI systems, changing banking...  ...customer experiences. Our investments in technology infrastructure and world-class talent — along with our deep experience in machine... 
    Full time
    Part time
    Local area

    Capital One

    San Jose, CA
    5 days ago
  • $215.2k - $245.6k

    AI Engineer 4 (Gen AI Platform Services) At Capital One, we are creating responsible and reliable AI systems, changing banking for good...  ...customer experiences. Our investments in technology infrastructure and world-class talent — along with our deep experience in machine... 
    Full time
    Part time
    Local area

    Capital One

    San Jose, CA
    1 day ago
  •  ...Job Title: AI Infrastructure Systems Engineer Location: San Jose, CA Contract JD: Highly skilled AI Infrastructure Systems Engineer to lead the deployment, provisioning, and optimization of our next-generation NVIDIA GB300 NVL72 AI supercomputer... 
    Contract work

    VDart Inc

    San Jose, CA
    2 days ago
  • $152k - $241.5k

     ...weight models are foundational to American AI leadership and cybersecurity, and that...  ...scrutiny. Our AI Safety & Security Engineering team builds and evaluates AI-powered tooling...  ...and maintain the agent harness.Evaluation infrastructure: Build the systems we use to run and... 
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    3 days ago
  • $165k - $247k

     ...offering product solutions in the automotive, industrial, infrastructure and IoT markets. Our robust product portfolio includes world...  ...and the world. In this role, you will be part of the AI & Cloud Engineering (ACE) Division and MLOps team. We are developing a comprehensive... 
    Temporary work
    Local area
    Remote work
    Worldwide
    Relocation
    Flexible hours

    Renesas Electronics

    San Jose, CA
    3 days ago
  • $150k - $350k

     ...Figure is an AI robotics company developing autonomous general-purpose humanoid robots...  ...level intelligence. Its robots are engineered to perform a variety of tasks in the home...  ...is looking for an experienced Training Infrastructure Engineer, to take our infrastructure to... 
    Full time

    Figure

    San Jose, CA
    more than 2 months ago
  • $240k - $260k

     ...SAVIYNT Saviynt is a leader in identity security, delivering an AI-powered platform that governs and secures access to applications...  ...is governed across the AI Platform. You define the standards ML engineers and scientists build on, and ensure every training signal is... 

    Saviynt

    Milpitas, CA
    a month ago
  • $140k - $165k

     ...most advanced electronic devices and IT infrastructure, enabling enhanced performance and user...  ...Why Join Us? Build foundational AI infrastructure that powers next-gen enterprise...  ...Role: We are seeking a hands-on AI Engineer to design, deploy, and maintain on-prem... 
    Full time

    SK Hynix Memory Solutions America Inc.

    San Jose, CA
    more than 2 months ago
  • $191k - $315k

     ...Overview:  The Network Growth and Relationship AI team is at the forefront of creating...  ...close collaboration with the product, engineering and data science team and has a very...  ...Prior experience with large scale ML data infrastructure ~ Experience with developing and designing... 
    Full time
    For contractors
    Work at office
    Flexible hours

    Linkedin

    Sunnyvale, CA
    more than 2 months ago
  • $200k - $400k

     ...Figure is an AI Robotics company developing a general purpose humanoid. Our humanoid...  ...'re looking for a senior-level backend engineer who has scaled high-throughput, low-latency...  ...and has strong instincts around cloud infrastructure and real-time streaming pipelines. You'... 
    Full time
    Work at office

    Figure

    San Jose, CA
    more than 2 months ago
  • $190k - $260k

     ...has developed an artificial intelligence (AI) powered technology stack purpose-built...  ...large-scale world models - depends on infrastructure that turns thousands of hours of multimodal...  ...training throughput. We are looking for engineers who make model training fast: streaming... 
    Temporary work
    Work at office
    Visa sponsorship

    Kodiak Robotics

    Mountain View, CA
    2 days ago
  • $120k - $220k

     ...news and information powered by advanced AI, recommendation systems, and adtech....  ...team to fulfill our mission: building the infrastructure layer for content intelligence. If you...  ...hiring our  first dedicated Agent Platform engineer to own this layer end-to-end. You'll... 
    Full time
    Local area
    Work from home

    News Break

    Mountain View, CA
    24 days ago
  •  ...Bitdeer is a world-leading technology company for AI and Bitcoin mining infrastructure. Bitdeer is committed to providing comprehensive Bitcoin mining...  ...We are seeking a Senior AI Storage Infrastructure Engineer to build the critical data-delivery fabric of our AI-native... 
    Remote job
    Full time
    Local area

    Bitdeer Technologies Group

    San Jose, CA
    24 days ago
  • $228k - $279k

     ...Senior Platform Engineer Mountain View, US About EarnIn As one of the first pioneers...  ...shifting from manual configuration to AI-augmented orchestration, where the...  ...and operate the agentic systems and core infrastructure that power our platform. You will work across... 
    Full time
    Work at office
    Shift work
    2 days per week

    Earnin

    Mountain View, CA
    3 days ago
  •  ...issues to a successful resolution. Perform testing in local and remote labs. Drive appropriate automated test execution to test engineers at various global locations. Provide training and guidance to test teams both onshore and offshore to properly set up, execute,... 
    Local area
    Remote work

    Net2Source

    San Jose, CA
    2 days ago
  •  ...responsibilities Design and build agentic AI systems, including autonomous agents,...  ...’s degree in Computer Science, AI, Engineering, or related field. ~3+ years of software...  ...containerization (Docker, Kubernetes), and infrastructure-as-code tools. About us Dear... 
    Full time

    Flat Rock Technology

    San Jose, CA
    more than 2 months ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to AI Infrastructure Engineer. Be the first to apply!