Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

AI Infrastructure Engineer

$170.5k - $315.49k

Intel

Job Details:Job Description: We are looking for a performance-obsessed AI Infrastructure Engineer to push LLM inference to its absolute limits on Intel's next-generation GPU architectures.In this role, you will dive deep into the inference stack and redefine peak performance. You will work end-to-end across the stack: profiling bottlenecks, writing custom GPU kernels, and upstreaming your optimizations directly into industry-standard serving frameworks like vLLM and SGLang. Your optimizations will be instrumental in unlocking the full potential of Intel hardware for state-of-the-art generative AI workloads.What You Will Do• Drive Inference Performance: Own the end-to-end optimization pipeline for running state-of-the-art LLMs on Intel GPUs.• Deep Stack Optimization: Profile, diagnose, and resolve cross-stack performance bottlenecks.• Kernel Development and Integration: Design, write, and optimize custom high-performance kernels for critical attention mechanisms, MoE, quantization, and operator fusions.• Open Source Leadership: Upstream your architectural improvements and hardware backends directly into open-source repositories like vLLM, SGLang, and PyTorch, acting as a bridge between the hardware teams and the open-source community.• Shape the Hardware Roadmap: Apply roofline analysis and systematic profiling to decompose bottlenecks. You will partner with our architecture and compiler teams to shape future GPU roadmaps based on real-world GenAI workload data.• Show passion about AI infrastructure and performance optimization.Qualifications:Minimum Qualifications• Bachelors Degree in Computer Science, Software Engineering, Artificial Intelligence/Machine Learning, or related field and 4+ years experience, Masters Degree and 3+ years, OR PhD.• 3+ years of relevant software engineering experience in GPU computing, AI systems, or high-performance computing (HPC).• Proficiency in modern C++ and Python. You are comfortable reading and modifying complex systems-level code.Preferred Qualifications • Understanding of CPU/GPU architecture.• Understanding of modern LLM architectures and inference paradigms: attention mechanisms, KV caching, continuous batching, speculative decoding, and prefill-decode disaggregation.• Prior open-source contributions to inference engines (vLLM, SGLang, PyTorch, llama.cpp).• Hands-on experience writing and optimizing custom GPU kernels using Triton, SYCL, CUDA/CUTLASS, or other DSLs.• Experience with scale-out inference orchestration across multi-node topologies.• You leverage AI coding agents daily to accelerate your own workflow and benchmark generation.Your expertise will play a vital role in advancing Intel's AI technology. We invite you to bring your skills, experience, and passion for AI to make an impact-apply today.Job Type:Experienced HireShift:Shift 1 (United States of America)Primary Location: US, California, Santa ClaraAdditional Locations:US, California, Folsom, US, Oregon, Hillsboro, US, Texas, AustinPosting Statement:All qualified applicants will receive consideration for employment without regard to race, color, religion, religious creed, sex, national origin, ancestry, age, physical or mental disability, medical condition, genetic information, military and veteran status, marital status, pregnancy, gender, gender expression, gender identity, sexual orientation, or any other characteristic protected by local law, regulation, or ordinance.Position of TrustN/ABenefitsWe offer a total compensation package that ranks among the best in the industry. It consists of competitive pay, stock bonuses, and benefit programs which include health, retirement, and vacation. Find out more about the benefits of working at Intel. Annual Salary Range for jobs which could be performed in the US: $170,500.00-315,490.00 USDThe range displayed on this job posting reflects the minimum and maximum target compensation for the position across all US locations. Within the range, individual pay is determined by work location and additional factors, including job-related skills, experience, and relevant education or training. Your recruiter can share more about the specific compensation range for your preferred location during the hiring process.Work Model for this RoleThis role will be eligible for our hybrid work model which allows employees to split their time between working on-site at their assigned Intel site and off-site. * Job posting details (such as work model, location or time type) are subject to change.*ADDITIONAL INFORMATION: Intel is committed to Responsible Business Alliance (RBA) compliance and ethical hiring practices. We do not charge any fees during our hiring process. Candidates should never be required to pay recruitment fees, medical examination fees, or any other charges as a condition of employment. If you are asked to pay any fees during our hiring process, please report this immediately to your recruiter.SummaryLocation: US, California, Santa Clara; US, Oregon, Hillsboro; US, California, Folsom; US, Texas, AustinType: Full time

Vacancy posted 18 hours ago
Similar jobs that could be interesting for youBased on the AI Infrastructure Engineer in Santa Clara, CA vacancy
  • $200k - $322k

     ...recently, GPU deep learning ignited modern AI — the next era of computing. NVIDIA is a...  ...choice to join us today.Design-for-X Engineering at NVIDIA works on groundbreaking innovations...  ...deployment cycles as part of the AI Infrastructure requirements at an org-wide level.For... 
    Suggested
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $184k - $287.5k

     ...leading cloud product that powers innovative AI research and developers. We focus on...  ...workloads, as well as developing scalable AI infrastructure services globally. We are seeking an AI infrastructure software engineer to join our team. You'll be instrumental in... 
    Suggested
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    1 day ago
  • $184k - $287.5k

    Joining NVIDIA's DGX Cloud AI Efficiency Team means contributing to the infrastructure that powers our innovative AI research. This team focuses on developing tools...  .... We are seeking an AI infrastructure software engineer to join our team. You'll be instrumental in... 
    Suggested
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    22 hours ago
  • $180k - $240k

     ...facilitating effortless integration into customers' logistics operations. About the role We are seeking a Senior AI Infrastructure Engineer to design, build, and scale the high-performance AI platform powering our autonomous driving models. While researchers focus... 
    Suggested
    Odd job
    Work at office

    Gatik AI

    Santa Clara, CA
    2 days ago
  • $192.1k - $249.6k

     ...Senior AI Inference Infrastructure Software Engineer NIO is a pioneer and a leading company in the premium smart electric vehicle market. Founded in November 2014, NIO's mission is to shape a joyful lifestyle. NIO aims to build a community starting with smart electric... 
    Suggested
    Full time
    Temporary work
    Immediate start
    Flexible hours

    NIO

    San Jose, CA
    1 day ago
  • $215.2k - $245.6k

    Lead AI Engineer (Gen AI Platform Services) Overview At Capital One, we are creating responsible and reliable...  ...customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine... 
    Full time
    Part time
    Local area

    Capital One Financial Corporation

    San Jose, CA
    22 hours ago
  • $178k - $321k

     ...Audit function has an early but working AI-native capability: a multi-agent platform...  ...setting, and the governed data and AI infrastructure everything else depends on. We hire on demonstrated...  ...just implement it. This is a two-person engineering team: you deploy, debug, and hotfix your... 

    OKX

    San Jose, CA
    2 days ago
  • $151.8k - $265.35k

    The OpportunityAdobe empowers individuals and organizations to create exceptional content effortlessly. The AI for Engineering team builds a scalable, production-grade AI platform that powers creativity across design, imaging, motion, and personalization.We are seeking... 
    Full time
    Temporary work
    Local area
    Worldwide

    Adobe Systems

    San Jose, CA
    22 hours ago
  • $184k - $287.5k

     ...tapping into the unlimited potential of AI to define the next era of computing. An...  ...Group (SCG) is seeking Senior AI Platform Engineers. They will set the technical direction...  ...foundational platforms at the intersection of ML infrastructure and large-scale systems, this is your... 
    Full time

    Nvidia

    Santa Clara, CA
    22 hours ago
  •  ...using it to some extent and at TransPerfect, we are no exception. We’re building an advanced voice processing platform that leverages AI for natural voice recognition, synthesis, and interaction. As part of the team, you’ll help design and develop scalable solutions... 
    Full time

    TransPerfect

    San Jose, CA
    22 hours ago
  •  ...that accelerate next-generation computing experiences—from AI and data centers, to PCs, gaming and embedded systems. Grounded...  ...career. THE ROLE:We are hiring a AI Research Scientist - Infrastructure Engineer, Reinforcement Learning, to own reinforcement learning infrastructure... 

    AMD

    Santa Clara, CA
    3 days ago
  • $177.1k - $387.5k

     ...ll design, build, and own the platform that powers Zoom AI Services, enabling AI capabilities to be delivered as...  ...’ll work across API design, distributed systems, cloud infrastructure, and AI platform engineering to build reliable, high-performance services that power... 
    Full time
    Work at office
    Remote work

    Zoom

    San Jose, CA
    2 days ago
  • $147k - $237.5k

     ...Execution, Integrity, and Inclusion. We weave AI into the fabric of everything we do and...  ...of integrating AI into cybersecurity infrastructure — building intelligent systems that...  ...incidents at scale. As a Principal Software Engineer, you will own the technical vision for... 
    Full time
    Work at office

    Palo Alto Networks

    Santa Clara, CA
    22 hours ago
  • $174.72k - $295.68k

     ...forefront of innovation, integrating advanced AI and autonomous driving technologies into...  ...connectivity.As a core member of our AI Infrastructure team, you will be responsible for...  ...or higher in Computer Science, Software Engineering, Artificial Intelligence, or related fields... 
    Full time
    Overseas

    XPENG Motors

    Santa Clara, CA
    22 hours ago
  •  ...next-generation computing experiences—from AI and data centers, to PCs, gaming and...  ...THE ROLEWe are hiring AI / ML Platform Engineers to build the platform layer that makes AI...  ...reproducible. This role focuses on the infrastructure and platform systems that support large-... 

    AMD

    Santa Clara, CA
    1 day ago
  • $152k - $241.5k

     ...weight models are foundational to American AI leadership and cybersecurity, and that...  ...scrutiny. Our AI Safety & Security Engineering team builds and evaluates AI-powered tooling...  ...and maintain the agent harness.Evaluation infrastructure: Build the systems we use to run and... 
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    1 day ago
  • $229.9k - $262.4k

    Sr. Lead AI Engineer (GenAI Platform) Overview: At Capital One, we are creating responsible and reliable AI systems, changing banking...  ...customer experiences. Our investments in technology infrastructure and world-class talent — along with our deep experience in machine... 
    Full time
    Part time
    Local area

    Capital One

    San Jose, CA
    3 days ago
  • $229.9k - $262.4k

    Senior Lead AI Engineer (GenAI Platform, Agentic Infrastructure) Overview: At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time,... 
    Full time
    Part time
    Local area

    Capital One

    San Jose, CA
    2 days ago
  •  ...Title: Prinicipal AI Engineer Location: Sunnyvale, California. Duration: 6 to 12+ Months Job Description:...  ...Proficient PostgreSQL, embeddings/vector search Cloud/Infrastructure Proficient GCP (BigQuery, GCS), Kubernetes, CCM Integration... 
    Contract work

    Redolent

    Sunnyvale, CA
    1 day ago
  • $229.9k - $262.4k

    Senior Lead AI Engineer (Gen AI Platform Services, Agentic AI) Overview: At Capital One, we are creating responsible and reliable AI...  ...customer experiences. Our investments in technology infrastructure and world-class talent — along with our deep experience in machine... 
    Full time
    Part time
    Local area

    Capital One

    San Jose, CA
    3 days ago
  •  ...leading technology company is seeking a Senior Software Engineer specializing in AI and ML in Sunnyvale, California. The ideal candidate will...  ...benefits, targeting mid-senior level professionals looking to drive innovations in AI infrastructure. #J-18808-Ljbffr Google
    Full time

    Google

    Sunnyvale, CA
    22 hours ago
  • $296.3k - $374.8k

     ...number of applications are received.Meet the TeamThe AI Platforms and Enablement team is the engine behind Cisco's internal AI transformation. Our...  ...operate at the highest levels of Cisco’s internal AI infrastructure, managing significant platform scale, high-volume inference... 
    Full time
    Temporary work
    Local area
    Flexible hours

    CISCO Systems

    San Jose, CA
    2 days ago
  • $245k - $325k

     ...SambaNova in San Jose, California is seeking a Director of Software Engineering to lead a high-performing engineering team in delivering cutting-edge AI inference platforms. The role involves overseeing team development, driving key engineering initiatives, and directly... 

    Jobleads-US

    San Jose, CA
    1 day ago
  • $200k - $322k

     ...technical Senior full-stack developer to build the next generation AI platforms and products, focusing on RAG, agentic AI, and...  ...leverage LLMs for reasoning and tool orchestration, and mentor engineers. Base salary ranges from 200,000 USD to 322,000 USD, with equity... 

    NVIDIA

    Santa Clara, CA
    4 days ago
  • $181.1k - $318.4k

    A leading technology company in Cupertino is seeking a Machine Learning Engineer to build infrastructure for product-focused machine learning projects. The ideal candidate will have a strong background in backend systems development and solid knowledge of machine learning... 

    Apple

    Cupertino, CA
    22 hours ago
  • Apple is seeking a Software Engineer for the AiDP data services team to build cloud-based data engineering solutions and streaming platforms powering critical business use cases. You will influence platform tools, APIs, and architecture while collaborating with internal... 

    Socket

    Sunnyvale, CA
    2 days ago
  • Lendistry, LLC. is seeking a Senior AI Engineer to lead the delivery of AI solutions, including document intelligence and risk assessment tools. In this role, you will be responsible for mentoring junior engineers and shaping AI-driven workflows, improving the borrower... 

    Lendistry, LLC.

    Santa Clara, CA
    4 days ago
  • $190k - $260k

     ...has developed an artificial intelligence (AI) powered technology stack purpose-built...  ...large-scale world models - depends on infrastructure that turns thousands of hours of multimodal...  ...training throughput. We are looking for engineers who make model training fast: streaming... 
    Temporary work
    Work at office
    Visa sponsorship

    Kodiak Robotics

    Mountain View, CA
    22 hours ago
  •  ...GRAIL is seeking a visionary Director of Software Engineering to lead our Data Platform engineering organization in Sunnyvale, CA. You will drive technical strategy, architecture, and delivery of a cloud-native data platform that handles petabyte-scale datasets. You... 

    Jobleads-US

    Sunnyvale, CA
    3 days ago
  • NVIDIA is seeking a Technical Marketing Engineer (TME) to communicate the latest AI platform software advances to developers and researchers. You will craft technical blog posts, guides, demonstrations, and benchmarks that help users run and optimize multi-GPU workloads... 

    NVIDIA AI

    Santa Clara, CA
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to AI Infrastructure Engineer. Be the first to apply!