Get new jobs by email
  •  ...part of an inclusive, adaptable, and forward-thinking organization, apply now.We are currently seeking a On-Premise LLM Inference & GPU Systems Engineer to join our team in Charlotte, North Carolina (US-NC), United States (US).Role Overview We are seeking an AI... 
    Suggested
    Work at office
    Remote work
    Flexible hours

    NTT DATA

    Charlotte, NC
    5 days ago
  •  ...comprehensive LLM serving platform with a focus on end-to-end inference research. You will work with the research lead to pick high‑impact...  .... You will collaborate with customers and Forward Deployed Engineers to deploy and tune models, and you will push frontier... 
    Suggested

    modal

    New York, NY
    2 days ago
  • $184k - $287.5k

    We are seeking highly skilled and motivated software engineers to join us and build AI inference systems that serve large-scale models with extreme efficiency. You’ll architect and implement high-performance inference stacks, optimize GPU kernels and compilers, drive industry... 
    Suggested
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  • $200k - $420k

     ...rewriting the entire stack from scratch: personal hardware for local inference, bespoke training infrastructure, next-generation UIs, and...  ...deep learning research. Who we are We are scientists, engineers, and builders from the industry's top tech companies and AI... 
    Suggested
    Full time
    Local area
    Visa sponsorship
    Relocation package

    River AI Inc.

    Palo Alto, CA
    4 days ago
  •  ...Title:  Applied AI Engineer — Inference & Agent Systems Location: United States What We're Building Arcana is building AI agents that synthesize information across heterogeneous sources and deliver structured, reasoned answers in real time. The product only... 
    Suggested
    Full time

    Arcana Analytics

    United States
    1 day ago
  • $190k - $260k

     ...they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft...  ...), especially how they influence latency and throughput of inference.Strong understanding or working experience with distributed systems... 
    Suggested
    Full time
    Work experience placement
    Work at office
    Local area
    Remote work
    Home office

    Cohere

    New York, NY
    3 days ago
  •  ...The AMD AI Group is looking for a Senior Software Development Engineer to own the end-to-end model execution stack on AMD Instinct GPUs...  ...training infrastructure at scale and high-performance inference serving. THE PERSON: This role demands someone who has shipped... 
    Suggested

    AMD

    Markham, IL
    4 days ago
  •  ...customers, better. And it means we prioritize a diverse F5 community where each individual can thrive.Job DescriptionThe AI Inference Engineer plays a critical role in the AI lifecycle by bridging the gap between high-performance model development and optimized deployment... 
    Suggested
    Full time
    Local area
    Immediate start

    F5 Networks

    San Jose, CA
    4 days ago
  • $184k - $287.5k

    We're now looking for a Sr. Inference Engineer, for GPU Kernel Optimization! What does it take to push every LLM inference operation to its performance ceiling? Our LLM Inference Performance Analysis and Optimization team builds the answer from the ground up. We develop... 
    Suggested
    Full time

    Nvidia

    Austin, TX
    5 days ago
  • $193.3k - $261.5k

    We are looking for a Senior Inference Engineer to own inference for real-time multimodalconversational AI. This is a full-stack inference role: you will work across the entire path a modeltakes from research to production — shaping model architecture so it is servable,... 
    Suggested
    Internship
    Local area
    Flexible hours

    Amazon

    Sunnyvale, CA
    4 days ago
  • Cloudflare is seeking a high-agency, systems engineer to design and build AI inference infrastructure across its global network. You will tackle sub-second model cold starts, multi-accelerator workload scheduling, and efficient cache management alongside AI/ML engineers... 
    Suggested

    Webhosting

    Austin, TX
    4 days ago
  • $170k - $245k

     ...to be backed by Andreessen Horowitz, NEA, and Addition with $250+ million raised to date.About the roleAs a Distributed LLM Inference Engineer, you will help systems and optimizations that push the boundaries of performance for inference at large scale. This is an incredibly... 
    Suggested
    Work at office

    Anyscale

    San Francisco, CA
    1 day ago
  •  ...us as we shape the future of AI and beyond. Together, we advance your career. THE ROLE:We are looking for a Senior GPU Inference Performance Engineer to own end-to-end performance analysis of GPU-accelerated AI inference workloads. You will profile, diagnose, and explain... 
    Suggested

    AMD

    Santa Clara, CA
    3 days ago
  • $167k - $209k

     ...profound difference for the dreamers and builders in the world. We are seeking a Senior Engineer II to implement and contribute to the design and optimization of our Serverless Inference infrastructure and APIs. In this role, you will tackle the challenges of large-scale... 
    Suggested
    Full time
    Local area
    Worldwide
    Flexible hours

    DigitalOcean

    Seattle, WA
    5 days ago
  • $184k - $287.5k

    We are now looking for a Senior DL Algorithms Engineer! NVIDIA is seeking senior engineers who are mindful of performance analysis and...  ...you will be doing:Implement language and multimodal model inference as part of NVIDIA Inference Microservices (NIMs).Contribute new... 
    Suggested
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  •  ...LLM Inference Engineer San Francisco or Remote Locations: San Francisco or Remote About The Role The NEAR AI team is building decentralized and confidential machine learning infrastructure to enable user-owned AI. Our mission is to build highly scalable and efficient... 
    Remote work

    NEAR.AI

    United States
    3 days ago
  •  ...Role Mission As HAI's LLM Inference Engineer, you will own the serving infrastructure that determines whether our breakthrough healthcare AI reaches patients efficiently and reliably. You'll optimize the systems that translate raw model capability into sub-100ms responses... 
    Work at office

    Hippocratic AI

    Menlo Park, CA
    3 days ago
  •  ...ultimate goal of enabling human life on Mars.AUTOMATION AND CONTROLS ENGINEER, AI SATELLITES (STARMIND) SpaceX is leveraging its experience...  ..., power distribution, and processing needed for orbital AI inference. The Starmind team designs, builds, and operates the orbital... 
    Permanent employment
    Weekend work

    SpaceX

    Bastrop, TX
    3 days ago
  •  ...the ultimate goal of enabling human life on Mars.OPTICAL TEST ENGINEER, AI SATELLITES (STARMIND) SpaceX is leveraging its experience...  ...control, power distribution, and processing needed for orbital AI inference. As we continue to upgrade and expand the constellation, we’re... 
    Permanent employment
    Work experience placement
    Internship
    Weekend work

    SpaceX

    Bastrop, TX
    3 days ago
  •  ...architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud...  ...high-speed inference.About The RoleWe're hiring a Principal Engineer for our Inference Cloud Platform. This team owns the cloud... 

    Cerebras Systems

    Sunnyvale, CA
    3 days ago
  • $120.6k - $224k

     ...-PrimaryStandard Job DescriptionWe are seeking an experienced engineer who is an subject matter expert in Guidance and Control (GNC)...  ...tracking, multi-sensor fusion, multi-hypothesis tracking, Bayesian inference, angle-only tracking, feature-aided tracking, image tracking,... 
    Full time
    Temporary work
    Work experience placement
    Interim role
    Casual work
    Flexible hours

    Lockheed Martin

    Orlando, FL
    5 days ago
  • $206.4k - $379.1k

     ...’s Generative AI Services team is seeking a Principal Service Engineer to serve as the technical lead for our GenAI Services domain....  ...generative models into Adobe’s flagship products.Design and architect inference infrastructure for enterprise-scale model customization,... 
    Full time
    Temporary work
    Local area
    Worldwide

    Adobe Systems

    San Francisco, CA
    3 days ago
  •  ...architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud...  ...are seeking a skilled and motivated Manufacturing Automation Engineer to design, implement, maintain, and optimize automated... 
    Contract work
    Shift work
    Weekend work

    Cerebras Systems

    Sunnyvale, CA
    1 day ago
  • $120.3k - $237.8k

     ...accepting online applications for the position Process Control Engineer through September 18, 2026 at 11:59 p.m. (Pacific Time). A Process...  ...control applications (DMC3, GDOT) and supporting applications (inferred properties, etc.) to control and optimize process unit... 
    Full time
    Temporary work
    Local area
    Relocation
    Visa sponsorship
    Work visa

    Chevron Oil Company

    Richmond, CA
    1 day ago
  •  ...About the Role We are seeking a Tokens-as-a-Service (TaaS) Engineer to help build the systems that convert large-scale infrastructure...  ...or workload optimization. Familiarity with model porting, inference/training workloads, token economics, or compute efficiency analysis... 
    Full time

    OpenAI

    San Francisco, CA
    1 day ago
  • SummaryMachine Test Engineer 2 Salary Range: $37.00 - $56.00/ Hr. As the largest machine tool builder in the western world, we need world...  ...including but not limited to probability, statistical inference, fundamentals of plane and solid geometry, trigonometry, and/or... 
    Work at office

    Haas Automation

    Oxnard, CA
    3 days ago
  •  ...planning, organization, control, integration, and completion of engineering project within area of assigned responsibility by performing...  ...with mathematical concepts such as probability and statistical inference, and fundamentals of plane and solid geometry and trigonometry... 
    Contract work
    Interim role
    Work at office
    All shifts

    TPI Composites

    Newton, IA
    5 days ago
  • $72.48k - $141.36k

     ...In this position... In this role, you will be supporting the engineering services team within Parts Supply & Logistics (PS&L) within Ford...  ...data sets, performing data visualization to draw key inferences for key PS&L stakeholders.• Assist in benchmarking best in world... 
    Immediate start
    3 days per week

    Ford

    Livonia, MI
    1 day ago
  • $195k - $285k

     ...the architecture of AI compute. As a Principal Hardware Design Engineer, you will be a cornerstone of our hardware organization,...  ...Neural Processing Unit) to solve the industry's most massive LLM inference challenges.Required Qualifications• Education: BS/MS in Electrical... 
    3 days per week

    d-Matrix

    Santa Clara, CA
    2 days ago
  •  ...input from customers, technical and product teams, manufacturing engineers, supplier partners, and other stakeholders to deliver...  ...probability distributions, graphical analysis, and statistical inference (population and sample, confidence intervals, and hypothesis testing... 
    Temporary work
    Internship

    Cummins

    Columbus, IN
    1 day ago