Get new jobs by email
- ...part of an inclusive, adaptable, and forward-thinking organization, apply now.We are currently seeking a On-Premise LLM Inference & GPU Systems Engineer to join our team in Charlotte, North Carolina (US-NC), United States (US).Role Overview We are seeking an AI...SuggestedWork at officeRemote workFlexible hours
- ...comprehensive LLM serving platform with a focus on end-to-end inference research. You will work with the research lead to pick high‑impact... .... You will collaborate with customers and Forward Deployed Engineers to deploy and tune models, and you will push frontier...Suggested
$184k - $287.5k
We are seeking highly skilled and motivated software engineers to join us and build AI inference systems that serve large-scale models with extreme efficiency. You’ll architect and implement high-performance inference stacks, optimize GPU kernels and compilers, drive industry...SuggestedFull time$200k - $420k
...rewriting the entire stack from scratch: personal hardware for local inference, bespoke training infrastructure, next-generation UIs, and... ...deep learning research. Who we are We are scientists, engineers, and builders from the industry's top tech companies and AI...SuggestedFull timeLocal areaVisa sponsorshipRelocation package- ...Title: Applied AI Engineer — Inference & Agent Systems Location: United States What We're Building Arcana is building AI agents that synthesize information across heterogeneous sources and deliver structured, reasoned answers in real time. The product only...SuggestedFull time
$190k - $260k
...they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft... ...), especially how they influence latency and throughput of inference.Strong understanding or working experience with distributed systems...SuggestedFull timeWork experience placementWork at officeLocal areaRemote workHome office- ...The AMD AI Group is looking for a Senior Software Development Engineer to own the end-to-end model execution stack on AMD Instinct GPUs... ...training infrastructure at scale and high-performance inference serving. THE PERSON: This role demands someone who has shipped...Suggested
- ...customers, better. And it means we prioritize a diverse F5 community where each individual can thrive.Job DescriptionThe AI Inference Engineer plays a critical role in the AI lifecycle by bridging the gap between high-performance model development and optimized deployment...SuggestedFull timeLocal areaImmediate start
$184k - $287.5k
We're now looking for a Sr. Inference Engineer, for GPU Kernel Optimization! What does it take to push every LLM inference operation to its performance ceiling? Our LLM Inference Performance Analysis and Optimization team builds the answer from the ground up. We develop...SuggestedFull time$193.3k - $261.5k
We are looking for a Senior Inference Engineer to own inference for real-time multimodalconversational AI. This is a full-stack inference role: you will work across the entire path a modeltakes from research to production — shaping model architecture so it is servable,...SuggestedInternshipLocal areaFlexible hours- Cloudflare is seeking a high-agency, systems engineer to design and build AI inference infrastructure across its global network. You will tackle sub-second model cold starts, multi-accelerator workload scheduling, and efficient cache management alongside AI/ML engineers...Suggested
$170k - $245k
...to be backed by Andreessen Horowitz, NEA, and Addition with $250+ million raised to date.About the roleAs a Distributed LLM Inference Engineer, you will help systems and optimizations that push the boundaries of performance for inference at large scale. This is an incredibly...SuggestedWork at office- ...us as we shape the future of AI and beyond. Together, we advance your career. THE ROLE:We are looking for a Senior GPU Inference Performance Engineer to own end-to-end performance analysis of GPU-accelerated AI inference workloads. You will profile, diagnose, and explain...Suggested
$167k - $209k
...profound difference for the dreamers and builders in the world. We are seeking a Senior Engineer II to implement and contribute to the design and optimization of our Serverless Inference infrastructure and APIs. In this role, you will tackle the challenges of large-scale...SuggestedFull timeLocal areaWorldwideFlexible hours$184k - $287.5k
We are now looking for a Senior DL Algorithms Engineer! NVIDIA is seeking senior engineers who are mindful of performance analysis and... ...you will be doing:Implement language and multimodal model inference as part of NVIDIA Inference Microservices (NIMs).Contribute new...SuggestedFull time- ...LLM Inference Engineer San Francisco or Remote Locations: San Francisco or Remote About The Role The NEAR AI team is building decentralized and confidential machine learning infrastructure to enable user-owned AI. Our mission is to build highly scalable and efficient...Remote work
- ...Role Mission As HAI's LLM Inference Engineer, you will own the serving infrastructure that determines whether our breakthrough healthcare AI reaches patients efficiently and reliably. You'll optimize the systems that translate raw model capability into sub-100ms responses...Work at office
- ...ultimate goal of enabling human life on Mars.AUTOMATION AND CONTROLS ENGINEER, AI SATELLITES (STARMIND) SpaceX is leveraging its experience... ..., power distribution, and processing needed for orbital AI inference. The Starmind team designs, builds, and operates the orbital...Permanent employmentWeekend work
- ...the ultimate goal of enabling human life on Mars.OPTICAL TEST ENGINEER, AI SATELLITES (STARMIND) SpaceX is leveraging its experience... ...control, power distribution, and processing needed for orbital AI inference. As we continue to upgrade and expand the constellation, we’re...Permanent employmentWork experience placementInternshipWeekend work
- ...architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud... ...high-speed inference.About The RoleWe're hiring a Principal Engineer for our Inference Cloud Platform. This team owns the cloud...
$120.6k - $224k
...-PrimaryStandard Job DescriptionWe are seeking an experienced engineer who is an subject matter expert in Guidance and Control (GNC)... ...tracking, multi-sensor fusion, multi-hypothesis tracking, Bayesian inference, angle-only tracking, feature-aided tracking, image tracking,...Full timeTemporary workWork experience placementInterim roleCasual workFlexible hours$206.4k - $379.1k
...’s Generative AI Services team is seeking a Principal Service Engineer to serve as the technical lead for our GenAI Services domain.... ...generative models into Adobe’s flagship products.Design and architect inference infrastructure for enterprise-scale model customization,...Full timeTemporary workLocal areaWorldwide- ...architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud... ...are seeking a skilled and motivated Manufacturing Automation Engineer to design, implement, maintain, and optimize automated...Contract workShift workWeekend work
$120.3k - $237.8k
...accepting online applications for the position Process Control Engineer through September 18, 2026 at 11:59 p.m. (Pacific Time). A Process... ...control applications (DMC3, GDOT) and supporting applications (inferred properties, etc.) to control and optimize process unit...Full timeTemporary workLocal areaRelocationVisa sponsorshipWork visa- ...About the Role We are seeking a Tokens-as-a-Service (TaaS) Engineer to help build the systems that convert large-scale infrastructure... ...or workload optimization. Familiarity with model porting, inference/training workloads, token economics, or compute efficiency analysis...Full time
- SummaryMachine Test Engineer 2 Salary Range: $37.00 - $56.00/ Hr. As the largest machine tool builder in the western world, we need world... ...including but not limited to probability, statistical inference, fundamentals of plane and solid geometry, trigonometry, and/or...Work at office
- ...planning, organization, control, integration, and completion of engineering project within area of assigned responsibility by performing... ...with mathematical concepts such as probability and statistical inference, and fundamentals of plane and solid geometry and trigonometry...Contract workInterim roleWork at officeAll shifts
$72.48k - $141.36k
...In this position... In this role, you will be supporting the engineering services team within Parts Supply & Logistics (PS&L) within Ford... ...data sets, performing data visualization to draw key inferences for key PS&L stakeholders.• Assist in benchmarking best in world...Immediate start3 days per week$195k - $285k
...the architecture of AI compute. As a Principal Hardware Design Engineer, you will be a cornerstone of our hardware organization,... ...Neural Processing Unit) to solve the industry's most massive LLM inference challenges.Required Qualifications• Education: BS/MS in Electrical...3 days per week- ...input from customers, technical and product teams, manufacturing engineers, supplier partners, and other stakeholders to deliver... ...probability distributions, graphical analysis, and statistical inference (population and sample, confidence intervals, and hypothesis testing...Temporary workInternship

