AI Infrastructure Engineer
NIO USA, INC
Job Description
Job Description
About NIO
NIO Inc. is a pioneer and a leading company in the global smart electric vehicle market. Founded in November 2014, NIO aspires to shape a sustainable and brighter future with the mission of “Blue Sky Coming”.
NIO envisions itself as a user enterprise where innovative technology meets experience excellence. NIO designs, develops, manufactures and sells smart electric vehicles, driving innovations in next-generation core technologies. NIO distinguishes itself through continuous technological breakthroughs and innovations, exceptional products and services, and a community for shared growth.
NIO provides premium smart electric vehicles under the NIO brand, premium smart electric vehicles for families through the ONVO brand, and small smart high-end electric cars with the FIREFLY brand.
About the Position
We are looking for a senior AI Inference Infrastructure Software Engineer with strong hands-on experience building, optimizing, and deploying high-performance, scalable inference systems. This position is focused on designing, implementing, and delivering production-grade software that powers real-world applications of Large Language Models (LLMs) and Vision-Language Models (VLMs).
This is an exciting opportunity for an engineer who thrives at the intersection of AI systems, hardware acceleration, and large-scale robust deployment, and who wants to see their contributions ship in production, at scale.
In this role, you will directly shape the architecture, roadmap and performance of AI capabilities of our AIOS platform, driving innovations that make LLM/VLM systems fast, efficient, and scalable across cloud, edge, and hybrid edge-cloud environments. You will work closely with system, hardware, and product teams to deliver high-performance inference kernels for hardware accelerators, design scalable inference serving systems, and integrate optimizations such tensor parallelism and custom kernels into production pipelines. Your work will have immediate impact, powering intelligent automotive systems in the next generation of electric vehicles.
Roles and ResponsibilitiesDesign and implement high-performance, scalable inference systems for LLMs and VLMs across cloud, edge, and edge-cloud hybrid platforms.
Develop and optimize custom kernels and operators for specific hardware accelerators (GPU, NPU, DSP, etc.), improving throughput, latency, and memory efficiency.
Integrate advanced optimization techniques such as KV-cache management, tensor/model parallelism, quantization, and memory-efficient execution into production inference systems.
Partner with system and hardware teams to ensure tight hardware-software integration and optimal performance across diverse compute environments.
Translate architectural requirements into robust, maintainable, production-ready software that meets performance, safety, and reliability standards.
Define and drive the evolution roadmap for LLM/VLM inference in the AIOS stack, ensuring scalability and adaptability to new workloads.
Stay ahead of industry trends and competitor solutions, applying best practices from both AI and large-scale systems engineering.
5+ years of hands-on software development experience in building and optimizing AI inference systems at scale.
Direct experience in LLM/VLM model internals, including Transformer-based architectures, inference bottlenecks, and optimization techniques.
Strong expertise in performance engineering: kernel development, parallelism strategies, memory optimization, and distributed inference systems.
Proficiency with GPU/NPU programming (CUDA, or vendor-specific SDKs), compiler toolchains, and deep learning frameworks (PyTorch, or TensorFlow).
Strong programming skills in C/C++, with a track record of delivering high-performance, production-grade software.
Solid foundation in computer architecture, systems programming (CPU/GPU pipelines, memory hierarchy, scheduling), and embedded systems.
BS/MS in Computer Science, Computer Engineering, or related technical field.
Excellent communication and collaboration skills, with the ability to work across cross-functional teams.
Master’s or PhD degree in Computer Science, Electrical/Computer Engineering, or related fields, plus 5 years industry experience
Experience building inference serving systems for large models, including batching, scheduling, caching, and load balancing.
Expertise in hardware-aware model optimization (e.g., kernel fusion, mixed precision, quantization, pruning).
Familiarity with edge and embedded AI, including real-time constraints and limited-resource optimization.
Contributions to widely used AI frameworks, libraries, or performance-critical software (open source or proprietary).
- Within the range, individual pay is determined by work location and additional factors, including job-related skills, experience, and relevant education or training.
- Please note that the compensation details listed in US role postings reflect the base salary only. It does not include discretionary bonus, equity, or benefits.
Benefits:
Along with competitive pay, as a full-time NIO employee, you are eligible for the following benefits on the first day you join NIO:
- Anthem Blue Cross, HSA, and Kaiser HMO medical plans with $0 for Employee Only Coverage.
- Dental (including orthodontic coverage) and vision plan. Both provide options with a $0 paycheck contribution covering you and your eligible dependents.
- Company Paid HSA (Health Savings Account) Contribution when enrolled in the High Deductible Anthem Blue Cross medical plan
- Healthcare and Dependent Care Flexible Spending Accounts (FSA)
- 401(k) with Brokerage Link option
- Company paid Basic Life, AD&D, short-term and long-term disability insurance
- Employee Assistance Program
- Sick and Vacation time
- 13 Paid Holidays a year
- Paid Parental Leave for first 8 weeks at full pay (eligible after 90 days of employment with NIO)
- Paid Disability Leave for first 6 weeks at full pay (eligible after 90 days of employment with NIO)
- Voluntary benefits including: Voluntary Life and AD&D options for you, your spouse/domestic partner and dependent child(ren), pet insurance
- Commuter benefits
- Mobile Cell Phone Credit
- Free lunch and snacks
- Onsite gym
- Employee discounts and perks program
We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.
$229.9k - $262.4k
Senior Lead AI Engineer (GenAI Platform Services) Overview: At Capital One, we are creating responsible and reliable AI systems... ...customer experiences. Our investments in technology infrastructure and world-class talent — along with our deep experience in...SuggestedFull timePart timeLocal area$180k - $240k
...facilitating effortless integration into customers' logistics operations. About the role We are seeking a Senior AI Infrastructure Engineer to design, build, and scale the high-performance AI platform powering our autonomous driving models. While researchers focus...SuggestedOdd jobWork at office$229.9k - $262.4k
Senior Lead AI Engineer (Gen AI Platform Services, Agentic AI) Overview: At Capital One, we are creating responsible and reliable AI... ...customer experiences. Our investments in technology infrastructure and world-class talent — along with our deep experience in machine...SuggestedFull timePart timeLocal area- ...next-generation computing experiences-from AI and data centers, to PCs, gaming and... ...ROLE We are hiring AI / ML Platform Engineers to build the platform layer that makes AI... ...reproducible. This role focuses on the infrastructure and platform systems that support large-...Suggested
$197.3k - $225.1k
Lead AI/ML Engineer (Platform, kubeflow) Overview At Capital One, we are creating responsible and reliable AI systems, changing banking... ...customer experiences. Our investments in technology infrastructure and world-class talent — along with our deep experience in machine...SuggestedFull timePart timeLocal area$133.2k - $192.8k
Job Details: Job Description: AI Engineer - Agentic AI Systems Why This Role Matters AI is shifting from models to autonomous systems. At Altera, we are building agentic AI platforms tightly coupled with FPGA acceleration, enabling intelligent systems that can reason...Full timeInternshipLocal areaShift work$151.8k - $265.35k
...-class team in San Jose, CA, where your engineering skills will flourish! In this role, you’... ...future of Adobe’s next-generation agentic AI strategy by building the agentic harness... ..., agentic systems, or developer-facing infrastructure for agent evaluation, orchestration,...Full timeTemporary workLocal areaWorldwide$200k - $322k
...recently, GPU deep learning ignited modern AI — the next era of computing. NVIDIA is a... ...choice to join us today.Design-for-X Engineering at NVIDIA works on groundbreaking innovations... ...deployment cycles as part of the AI Infrastructure requirements at an org-wide level.For...Full time$197.3k - $225.1k
Lead AI Engineer At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital... ...personalized customer experiences. Our investments in technology infrastructure and world-class talent — along with our deep experience in...Full timePart timeLocal area- ...CoreWeave, the AI hyperscaler, is hiring software engineers to help build and scale Weave, a platform for AI agents and GenAI applications. In a foundational role on a growing team, you will design core functionality, gather feedback from users, and iterate on the product...
- ...intelligence that brings it all to life. As an AI-powered, robot-agnostic orchestration... ...The Role As an Applied AI Software Engineer, you sit at the intersection of Machine... ...or TensorRT. ● Data Pipelines: Build the infrastructure to collect, clean, and version-control...
$105.5k - $213.5k
Cloud and AI Ops EngineerThis role has been designed as ‘’Onsite’ with an expectation... ...Edge (SASE) business. The Cloud and AI Engineer builds from the ground up to meet the... ...Job Responsibilities):Deployment of Cloud infrastructure using Docker Containerization.Automate Cloud...Full timeWork experience placementWork at officeLocal areaImmediate start2 days per week$184k - $287.5k
Joining NVIDIA's DGX Cloud AI Efficiency Team means contributing to the infrastructure that powers our innovative AI research. This team focuses on developing tools... .... We are seeking an AI infrastructure software engineer to join our team. You'll be instrumental in...Full timeRemote work$95.6k - $150.7k
...network is built out Focus entirely on connecting the two infrastructures securely while protecting corporate data from new vulnerabilities... ...in quality, and driving operational outcomes The Team AI & Engineering leverages cutting-edge engineering capabilities to build,...- ...Figure is an AI robotics company developing autonomous general-purpose humanoid robots. Our goal is to build embodied AI systems that... ...that power humanoid autonomy. We are looking for a Helix AI Engineer, Pretraining to build large-scale foundation models that learn...Full timeWork at office
- Figure is an AI robotics company developing autonomous general-purpose humanoid robots. Our goal is to build embodied AI systems that... ...that power humanoid autonomy. We are looking for a Helix AI Engineer, Video Pretraining to lead the development of large-scale video...Full timeWork at office
- Figure is an AI robotics company developing autonomous general-purpose humanoid robots. Our goal is to build embodied AI systems that... ...that power humanoid autonomy. We are looking for a Helix AI Engineer, Modeling to design and advance the core model architectures and...Full timeWork at office
$160k - $180k
...value candidates who are actively using AI tools to enhance productivity, automate... ...equivalent experience in Computer Science, Engineering, Machine Learning, or other relevant... ...Specifics like Redshift or Spark are needed. Infrastructure engineering and/or devops knowledge on...Full time- ...Helix AI Engineer, Robot Learning Figure is an AI robotics company developing autonomous general-purpose humanoid robots. The goal of the company is to ship humanoid robots with human level intelligence. Its robots are engineered to perform a variety of tasks in the...Full time
- Figure is an AI robotics company developing autonomous general-purpose humanoid robots... .... We are looking for a Helix AI Engineer, Generative AI to build and scale generative... ...Work closely with data, training infrastructure, and agent teams to integrate generative...Full timeWork at office
- Figure is an AI robotics company developing autonomous general-purpose humanoid robots... .... We are looking for a Helix AI Engineer, Reinforcement Learning to develop learning... ...including distributed rollouts, simulation infrastructure, and experiment management Design...Full timeWork at office
$151.8k - $265.35k
The OpportunityAdobe empowers individuals and organizations to create exceptional content effortlessly. The AI for Engineering team builds a scalable, production-grade AI platform that powers creativity across design, imaging, motion, and personalization.We are seeking...Full timeTemporary workLocal areaWorldwide$178k - $321k
...Audit function has an early but working AI-native capability: a multi-agent platform... ...setting, and the governed data and AI infrastructure everything else depends on. We hire on demonstrated... ...just implement it. This is a two-person engineering team: you deploy, debug, and hotfix your...- ...that accelerate next-generation computing experiences—from AI and data centers, to PCs, gaming and embedded systems. Grounded... ...career. THE ROLE:We are hiring a AI Research Scientist - Infrastructure Engineer, Reinforcement Learning, to own reinforcement learning infrastructure...
- ...using it to some extent and at TransPerfect, we are no exception. We’re building an advanced voice processing platform that leverages AI for natural voice recognition, synthesis, and interaction. As part of the team, you’ll help design and develop scalable solutions...Full time
$184k - $287.5k
...tapping into the unlimited potential of AI to define the next era of computing. An... ...Group (SCG) is seeking Senior AI Platform Engineers. They will set the technical direction... ...foundational platforms at the intersection of ML infrastructure and large-scale systems, this is your...Full time$240k - $260k
...Job Description Job Description AI Platform Engineer – Training & Inference Saviynt's AI-powered identity platform manages and... ...between self-hosted SLMs and cloud LLMs • Build RL training infrastructure: define Flyte workflows for RL pipelines (rollout, reward...$175.5k - $263.8k
...provide the financial insight that our executives need to keep Apple as the preeminent company in the world. Description As an AI Solutions Engineer on the Corporate FP&A Data Analytics team, you'll pair hands‑on vibe‑coding and AI‑assisted development with the work that...Relocation$177.1k - $387.5k
...ll design, build, and own the platform that powers Zoom AI Services, enabling AI capabilities to be delivered as... ...’ll work across API design, distributed systems, cloud infrastructure, and AI platform engineering to build reliable, high-performance services that power...Full timeWork at officeRemote work$147k - $237.5k
...Execution, Integrity, and Inclusion. We weave AI into the fabric of everything we do and... ...of integrating AI into cybersecurity infrastructure — building intelligent systems that... ...incidents at scale. As a Principal Software Engineer, you will own the technical vision for...Full timeWork at office
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to AI Infrastructure Engineer. Be the first to apply!
- ai developer San Jose, CA
- ai prompt engineer San Jose, CA
- ai engineer San Jose, CA
- senior ai engineer San Jose, CA
- senior infrastructure engineer San Jose, CA
- infrastructure engineer San Jose, CA
- infrastructure developer San Jose, CA
- remote infrastructure engineer San Jose, CA
- gen ai developer
- ai automation engineer



