AI Infrastructure Engineer
$192.1k - $249.6kNIO
Senior AI Inference Infrastructure Software EngineerNIO is a pioneer and a leading company in the premium smart electric vehicle market. Founded in November 2014, NIO's mission is to shape a joyful lifestyle. NIO aims to build a community starting with smart electric vehicles to share joy and grow together with users.NIO designs, develops, jointly manufactures and sells premium smart electric vehicles, driving innovations in next-generation technologies in autonomous driving, digital technologies, electric powertrains and batteries. NIO differentiates itself through its continuous technological breakthroughs and innovations, such as its industry-leading battery swapping technologies, Battery as a Service, or BaaS, as well as its proprietary autonomous driving technologies and Autonomous Driving as a Service, or ADaaS.NIO's product portfolio consists of the ES8, a six-seater smart electric flagship SUV, the ES7 (or the EL7), a mid-large five-seater smart electric SUV, the ES6, a five-seater all-round smart electric SUV, the EC7, a five-seater smart electric flagship coupe SUV, the EC6, a five-seater smart electric coupe SUV, the ET7, a smart electric flagship sedan, and the ET5, a mid-size smart electric sedan.We are looking for a senior AI Inference Infrastructure Software Engineer with strong hands-on experience building, optimizing, and deploying high-performance, scalable inference systems. This position is focused on designing, implementing, and delivering production-grade software that powers real-world applications of Large Language Models (LLMs) and Vision-Language Models (VLMs). This is an exciting opportunity for an engineer who thrives at the intersection of AI systems, hardware acceleration, and large-scale robust deployment, and who wants to see their contributions ship in production, at scale.In this role, you will directly shape the architecture, roadmap and performance of AI capabilities of our AIOS platform, driving innovations that make LLM/VLM systems fast, efficient, and scalable across cloud, edge, and hybrid edge-cloud environments. You will work closely with system, hardware, and product teams to deliver high-performance inference kernels for hardware accelerators, design scalable inference serving systems, and integrate optimizations such tensor parallelism and custom kernels into production pipelines. Your work will have immediate impact, powering intelligent automotive systems in the next generation of electric vehicles.Roles and Responsibilities:Design and implement high-performance, scalable inference systems for LLMs and VLMs across cloud, edge, and edge-cloud hybrid platforms.Develop and optimize custom kernels and operators for specific hardware accelerators (GPU, NPU, DSP, etc.), improving throughput, latency, and memory efficiency.Integrate advanced optimization techniques such as KV-cache management, tensor/model parallelism, quantization, and memory-efficient execution into production inference systems.Partner with system and hardware teams to ensure tight hardware-software integration and optimal performance across diverse compute environments.Translate architectural requirements into robust, maintainable, production-ready software that meets performance, safety, and reliability standards.Define and drive the evolution roadmap for LLM/VLM inference in the AIOS stack, ensuring scalability and adaptability to new workloads.Stay ahead of industry trends and competitor solutions, applying best practices from both AI and large-scale systems engineering.Must Qualifications:5+ years of hands-on software development experience in building and optimizing AI inference systems at scale.Direct experience in LLM/VLM model internals, including Transformer-based architectures, inference bottlenecks, and optimization techniques.Strong expertise in performance engineering: kernel development, parallelism strategies, memory optimization, and distributed inference systems.Proficiency with GPU/NPU programming (CUDA, or vendor-specific SDKs), compiler toolchains, and deep learning frameworks (PyTorch, or TensorFlow).Strong programming skills in C/C++, with a track record of delivering high-performance, production-grade software.Solid foundation in computer architecture, systems programming (CPU/GPU pipelines, memory hierarchy, scheduling), and embedded systems.BS/MS in Computer Science, Computer Engineering, or related technical field.Excellent communication and collaboration skills, with the ability to work across cross-functional teams.Compensation:The US base salary range for this full-time position is $192,100.00 - $249,600.00.Within the range, individual pay is determined by work location and additional factors, including job-related skills, experience, and relevant education or training.Please note that the compensation details listed in US role postings reflect the base salary only. It does not include discretionary bonus, equity, or benefits.Benefits:Anthem Blue Cross, HSA, and Kaiser HMO medical plans with $0 for Employee Only Coverage.Dental (including orthodontic coverage) and vision plan. Both provide options with a $0 paycheck contribution covering you and your eligible dependents.Company Paid HSA (Health Savings Account) Contribution when enrolled in the High Deductible Anthem Blue Cross medical planHealthcare and Dependent Care Flexible Spending Accounts (FSA)401(k) with Brokerage Link optionCompany paid Basic Life, AD&D, short-term and long-term disability insuranceEmployee Assistance ProgramSick and Vacation time13 Paid Holidays a yearPaid Parental Leave for first 8 weeks at full pay (eligible after 90 days of employment with NIO)Paid Disability Leave for first 6 weeks at full pay (eligible after 90 days of employment with NIO)Voluntary benefits including: Voluntary Life and AD&D options for you, your spouse/domestic partner and dependent child(ren), pet insuranceCommuter benefitsMobile Cell Phone CreditFree lunch and snacksOnsite gymEmployee discounts and perks program
$170.5k - $315.49k
Job Details:Job Description: We are looking for a performance-obsessed AI Infrastructure Engineer to push LLM inference to its absolute limits on Intel's next-generation GPU architectures.In this role, you will dive deep into the inference stack and redefine peak performance...SuggestedFull timeLocal areaImmediate startShift work$180k - $240k
...Senior AI Infrastructure EngineerSanta Clara, CAAbout the roleWe are seeking a Senior AI Infrastructure Engineer to design, build, and scale the high-performance AI platform powering our autonomous driving models. While researchers focus on developing perception, planning...SuggestedWork at office$184k - $287.5k
Joining NVIDIA's DGX Cloud AI Efficiency Team means contributing to the infrastructure that powers our innovative AI research. This team focuses on developing tools... .... We are seeking an AI infrastructure software engineer to join our team. You'll be instrumental in...SuggestedFull timeRemote work$184k - $287.5k
...leading cloud product that powers innovative AI research and developers. We focus on... ...workloads, as well as developing scalable AI infrastructure services globally. We are seeking an AI infrastructure software engineer to join our team. You'll be instrumental in...SuggestedFull timeRemote work$151.8k - $265.35k
The OpportunityAdobe empowers individuals and organizations to create exceptional content effortlessly. The AI for Engineering team builds a scalable, production-grade AI platform that powers creativity across design, imaging, motion, and personalization.We are seeking...SuggestedFull timeTemporary workLocal areaWorldwide- ...using it to some extent and at TransPerfect, we are no exception. We’re building an advanced voice processing platform that leverages AI for natural voice recognition, synthesis, and interaction. As part of the team, you’ll help design and develop scalable solutions...Full time
- ...that accelerate next-generation computing experiences—from AI and data centers, to PCs, gaming and embedded systems. Grounded... ...career. THE ROLE:We are hiring a AI Research Scientist - Infrastructure Engineer, Reinforcement Learning, to own reinforcement learning infrastructure...
$184k - $287.5k
...tapping into the unlimited potential of AI to define the next era of computing. An... ...Group (SCG) is seeking Senior AI Platform Engineers. They will set the technical direction... ...foundational platforms at the intersection of ML infrastructure and large-scale systems, this is your...Full time$178k - $321k
...Audit function has an early but working AI-native capability: a multi-agent platform... ...setting, and the governed data and AI infrastructure everything else depends on. We hire on demonstrated... ...just implement it. This is a two-person engineering team: you deploy, debug, and hotfix your...- Bitdeer is a world-leading technology company for AI and Bitcoin mining infrastructure. Bitdeer is committed to providing comprehensive Bitcoin mining... ...We are seeking a Senior AI Storage Infrastructure Engineer to build the critical data-delivery fabric of our AI-native...Local area
$147k - $237.5k
...Execution, Integrity, and Inclusion. We weave AI into the fabric of everything we do and... ...of integrating AI into cybersecurity infrastructure — building intelligent systems that... ...incidents at scale. As a Principal Software Engineer, you will own the technical vision for...Full timeWork at office$250.8k - $286.2k
Senior Lead AI Engineer (GenAI Platform Services, Agentic Platform) Overview: At Capital One, we are creating responsible and reliable... ...customer experiences. Our investments in technology infrastructure and world-class talent — along with our deep experience in machine...Full timePart timeLocal area$177.1k - $387.5k
...ll design, build, and own the platform that powers Zoom AI Services, enabling AI capabilities to be delivered as... ...’ll work across API design, distributed systems, cloud infrastructure, and AI platform engineering to build reliable, high-performance services that power...Full timeWork at officeRemote work$145.6k - $246.4k
...forefront of innovation, integrating advanced AI and autonomous driving technologies into... ...connectivity.As a core member of our AI Infrastructure team, you will be responsible for... ...or higher in Computer Science, Software Engineering, Artificial Intelligence, or related fields...Full timeOverseas- ...next-generation computing experiences—from AI and data centers, to PCs, gaming and... ...THE ROLEWe are hiring AI / ML Platform Engineers to build the platform layer that makes AI... ...reproducible. This role focuses on the infrastructure and platform systems that support large-...
$150k - $200k
...System which makes everything possible. The Backend & AI Platform Engineer is responsible for creating the backend platform that powers... ...and APIs that connect laboratory instruments, cloud infrastructure, and AI agents into a unified platform. Design scalable...Full timePart time- ...products that accelerate next-generation computing experiences—from AI and data centers, to PCs, gaming and embedded systems. Grounded... ..., we advance your career. THE ROLEWe are seeking a Senior Data Engineer to design, build, and optimize our next-generation data platform...
$269.1k - $307.2k
...Distinguished AI Engineer - Agentic AI Platform (Remote Eligible) At Capital One, we are creating responsible and reliable AI systems... ...customer experiences. Our investments in technology infrastructure and world-class talent — along with our deep experience in machine...Full timePart timeWork at officeRemote work- NVIDIA Corporation is seeking a Senior Customer Success Engineer for the DGX Cloud team in the US. You will embed with internal customers... ..., work with Kubernetes and DevOps tooling, and help drive infrastructure strategy with engineering, product, and finance teams. #J-18...Remote job
$152k - $241.5k
...weight models are foundational to American AI leadership and cybersecurity, and that... ...scrutiny. Our AI Safety & Security Engineering team builds and evaluates AI-powered tooling... ...and maintain the agent harness.Evaluation infrastructure: Build the systems we use to run and...Full timeRemote work- Bitdeer Technologies Group is seeking an L1 NOC/US Data Center operator to support NeoCloud's GPU DCs during 8AM-8PM PST shifts. You will monitor GPU clusters, networks, and storage, respond to alerts, and execute runbooks for common incidents across shore-to-APAC handoffs...Shift workNight shift
$224k - $356.5k
NVIDIA in Santa Clara is seeking a System Software Engineer for Vision AI to develop high-performance vision systems that process massive data streams. The ideal candidate will have over 12 years of experience in software development with expertise in C++, Python, and...- Yoh is seeking a Principal Technical Support Engineer to lead customer deployments and platform bring-up for next‑gen AI networking. You will optimize performance on high-speed Ethernet fabrics and GPU clusters, working across SONiC-based systems, switch ASICs, and related...
$147.9k - $220k
NetApp in San Jose, CA is looking for a skilled performance engineer to design, develop, and optimize cloud software for storage services... ...performance analysis, strong programming skills in C, and experience with AI/ML techniques. This hybrid position emphasizes collaboration...- Astera Labs is looking for a Principal Product Applications Engineer to join their Aries PCIe Retimer team in San Jose, CA. This role bridges the gap between customers and engineering, engaging directly with hyperscalers and OEMs to drive product adoption while solving...
$125.5k - $261.6k
...a better working world. The opportunity We are seeking AI Systems Engineers to build and operate the foundational substrate that powers EY’s AI-native platform. This role owns the infrastructure and cloud-native platform layers of the Hybrid AI Multi-Environment...Full timeContract workSummer holidayFlexible hours$286.2k - $326.7k
...Senior Director, AI Engineering -Agentic AI PlatformAt Capital One, we are creating responsible and reliable AI systems, changing banking... ...customer experiences. Our investments in technology infrastructure and world-class talent — along with our deep experience in machine...Full timePart timeRemote work- Adobe seeks a Senior Software Engineer to advance the Meta Factory agent harness, shaping runtimes, context handling, and tool integration... ..., and drive scalable, secure architectures for enterprise-scale AI-enabled workflows. Ideal candidates bring 12+ years in software...
- NVIDIA is seeking senior engineers to advance its AI platform, focusing on performance optimizations in deep learning frameworks using JAX. The role emphasizes modular, fast, and coordinated tooling to handle data, training and analysis for diverse DL solutions. Strong...
- ByteDance's Volcano Ark training platform is seeking engineers to advance end-to-end large model post-training (SFT, RL) on a serverless... ...stack. Join the Data AML team driving scalable, secure ML infrastructure. You will design elastic, multi-tenant training across data...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to AI Infrastructure Engineer. Be the first to apply!
- machine learning ai engineer San Jose, CA
- ai developer San Jose, CA
- senior ai engineer San Jose, CA
- ai engineer San Jose, CA
- ai ml engineer San Jose, CA
- ai prompt engineer San Jose, CA
- ai engineer remote San Jose, CA
- data infrastructure engineer San Jose, CA
- infrastructure engineering manager San Jose, CA
- senior infrastructure engineer San Jose, CA

