Senior Staff Applied AI Inference Engineer
$250k - $300kCrusoe
Crusoe is on a mission to accelerate the abundance of energy and intelligence . As the only vertically integrated AI infrastructure company built from the ground up, we own and operate each layer of the stack — from electrons to tokens — to power the world's most ambitious AI workloads. When you join Crusoe, you join a team that is building the future, faster.
We're in the midst of the greatest industrial revolution of our time. The demand for AI compute is boundless, and power is a bottleneck. We're solving that — with an energy-first approach that makes AI infrastructure better for the world and faster for the people innovating with AI.
We're looking for problem-solving, opportunity-finding teammates with a sense of urgency, who believe in the scale of our ambition and thrive on a path not fully paved — people who want to grow their careers alongside a team of experts across energy, manufacturing, data center construction, and cloud services.
If you want to do the most meaningful work of your career, help our customers and partners advance their AI strategies, and be part of a high-performing team that believes in each other, come build with us at Crusoe.
About the RoleYou will spend your time making large language models run faster, cheaper, and more reliably in production. That means owning the inference stack end to end: profiling where time and cost go, bringing modern optimization techniques into real deployments, and getting deep into the serving code when the defaults are not good enough. This is core systems and performance work on some of the most demanding models in use today.
The work is applied, not academic. The optimizations you build land in real customer deployments, each with its own models, traffic patterns, latency targets, and cost constraints. So while performance is the heart of the role, you will also work directly with customer engineering teams to tailor deployments to their needs, take a workload from an early proof of concept to a fully monitored production service, and make sure the gains you engineer actually show up for the people running the workload.
To set expectations clearly, this is a hands-on engineering role built around coding, profiling, and low-level optimization. It also carries a customer-facing side, along with elements of product and technical solutions work, because that is where the performance work gets proven.
What You'll Be Working On:- Bring current inference techniques into production and refine them.
- Design and optimize serving architectures, including prefill and decode disaggregation, request routing, and related approaches.
- Work down into the serving stack, from frameworks like vLLM and SGLang to the CUDA kernels underneath, profiling and running in-depth analysis to find and fix performance problems.
- Adapt and scale optimization methods across many kinds of ML models, with an emphasis on large language models.
- Profile and tune deployments against clear targets for latency, throughput, and cost, and keep them dependable under real traffic.
- Tailor deployments to each customer's models and constraints, partnering with their engineering teams to move a workload from an early proof of concept through to a live, well-monitored production service.
- Build and support the software and product features around the inference stack in a production setting, using one or more general-purpose languages, with Python preferred given how central it is to ML work.
- Experiment quickly: take fuzzy goals, shape them into clear specs and focused proofs of concept, run fast experiments to find what works, and ship well-tested results without delay.
- Own delivery end to end, from the first experiment through to the optimization running in production, keeping the underlying performance goals, clear specs, and follow-through front of mind, and drafting features and product requirement documents together with other engineering and product teams.
- Work through ambiguity and make sound calls on tradeoffs and tooling, steering away from complexity that is not needed.
- Take real pride and ownership in your work, hold yourself accountable, and look for the same from the people around you.
- A Bachelor's, Master's, or Ph.D. in Computer Science, Engineering, Mathematics, or a related field.
- Hands-on experience shipping code in production with one or more general-purpose languages, such as Python or C++, with a strong preference for Python.
- Familiarity with methods for optimizing LLMs for high throughput / low latency inference.
- Comfort with modern LLM serving frameworks such as vLLM or SGLang, and with profiling and analyzing performance down to the kernel level.
- A firm grasp of how GPUs are built and how they behave.
- Clear interest and hands-on experience with large language models.
- A working knowledge of AI/ML pipelines and the full path of developing and deploying ML models.
- Strong communication skills, particularly when explaining hard technical topics to customers and teammates.
- A track record of making software systems run faster, especially for large language models.
- Experience with CUDA or comparable technologies.
- A strong command of software engineering fundamentals, with a record of building and shipping AI/ML inference systems.
- Experience with Docker and Kubernetes.
- Prior work building or tuning AI/ML projects, particularly in a customer-facing setting.
- Competitive compensation and equity packages
- Restricted Stock Units
- Paid time off, paid holidays & leave of absence programs
- Comprehensive health, dental & vision insurance
- Employer contributions to HSA account
- Paid parental leave
- Paid life insurance, short-term and long-term disability
- Professional development & tuition reimbursement
- Mental health & wellness support
- Commuter benefits (parking & transit)
- Cell phone stipend
- 401(k) Retirement plan with company match up to 4% of salary
- Volunteer time off
- Global travel insurance & emergency assistance
- Daily meals allowance
- Additional perks & programs specific to location
Compensation will be paid in the range of up to $250,000 - $300,000 + Bonus. Restricted Stock Units are included in all offers. Compensation to be determined by the applicant's knowledge, education, and abilities, as well as internal equity and alignment with market data.
Crusoe is an Equal Opportunity Employer. Employment decisions are made without regard to race, color, religion, disability, genetic information, pregnancy, citizenship, marital status, sex/gender, sexual preference/ orientation, gender identity, age, veteran status, national origin, or any other status protected by law or regulation.
$188k - $275k
...CoreWeave is The Essential Cloud for AI™. Built for pioneers by pioneers, CoreWeave... ...Description of the team: The Inference team is responsible for delivering high... ...About the role: We are looking for an Applied AI Engineer to help us understand, measure, and improve...SuggestedPermanent employmentFull timeTemporary workCasual workWork at officeFlexible hours$200k - $250k
...Description About the Role Roger is an AI platform that frees home health... ...accuracy. We are now looking for a Senior Applied AI Engineer to build the intelligence layer at the... ...improvements. Build scalable, cost-efficient inference infrastructure with great monitoring...SeniorRemote workWork from home$175k - $200k
...everyone. Patients interact with advanced AI systems at every step of their care... ...latest publications . About the Senior Applied AI Engineer role We are hiring Senior Applied AI... ...evaluating models, productionizing inference, and measuring real-world clinical and...SeniorFull timeRemote workWork from homeFlexible hours$170k - $233k
...Responsible for serving as a senior technical leader... ...deployment of advanced AI solutions that address... ...and internal expert in applied AI, shaping the organization... ...(5) + years technical engineering experience with coding... ...learning, or real-time inference systems. Track...SeniorFull timeLocal area$139.2k - $208.8k
...aim to leave a positive mark on culture. Job Title : Senior Applied AI Engineer Team : Global Quality Engineering Location : New... ...ML workflows o Vertex AI Online Endpoints for real-time inference ● Integrate Vertex AI with GCP services (BigQuery, Cloud...SeniorLocal area- ...by Fast Company.For more information visit cultureamp.com.How you can help make a better world of workWe're looking for a Senior Applied AI Engineer to join the Frontier team, helping build Agentic AI Solutions at Culture Amp. You'll work at the intersection of applied...SeniorWork at officeLocal areaShift work2 days per week
$197.4k - $266.6k
...Forbes Best Startup Employers 2022 List.AI is transforming customer support and, with... ...product teams to identify opportunities for applying generative AI to solve user needs and... ...data structures, and principles of software engineering.Nice to Have:Knowledge in specialized...SeniorWork at officeImmediate startRemote workWork from homeMonday to Friday- ...better way, and we create possibilities. Interested in joining us on our journey? We are looking for a highly motivated Senior Applied AI Engineer to join our digital innovation team and help revolutionize how we serve our primarily B2C (Business-to-Consumer) customers...SeniorFull timeWork at officeRemote workFlexible hours
$177k - $226k
...landscape of critical care through our rapid seizure detection technology, come join the movement!Position Overview:The Senior Manager, Applied AI Engineering is a senior individual contributor role with broad ownership across Ceribell's internal AI engineering portfolio....SeniorFor contractorsWork at officeLocal areaImmediate startRemote work$228.6k - $342.8k
...do this by building and running the world's best data and AI infrastructure platform so our customers can use deep data... ...and how well it understands what it finds.We are hiring a Senior Staff Applied AI Engineer to own context retrieval for Databricks agents across SaaS...SeniorWork at officeLocal areaImmediate startWorldwide$104.9k - $199.07k
...BackgroundThe rapid evolution of artificial intelligence (AI) presents a transformative opportunity for Milliman to... ...objectives.Role PurposeMilliman AI Solutions is seeking a Senior Insurance Applied AI Engineer to establish and refine the business engagement and engineering...SeniorFull timeWork experience placementRemote workWorldwide- ...Senior Applied AI Engineer Our client is a fast-growing European fintech company in the business spend management space. We're looking for a Senior Applied AI Engineer to design, build and ship customer-facing AI features end-to-end. Responsibilities Design...SeniorFreelanceWork at officeImmediate startRemote workFlexible hours
$140k - $170k
...Senior Applied AI Engineer At Veracity, we aim to be a different kind of insurance partner – one that is free from outside investors, venture capital, or the pressures of a corporate parent. Ours is a culture of empowerment – one that believes in effort, results...SeniorRemote work$160k - $195k
...remote-first SaaS company building practical AI-powered solutions that improve real... ...continuous learning. We are expanding our engineering team to build the next generation of AI-... ...The Role We are looking for a Senior Applied AI Engineer to help design, build, and...SeniorFull timeLocal areaRemote work- ...is a hands-on, high-ownership role for an engineer who has built agents at scale and wants... ...rather than just view it. ~Make health AI that people trust with their lives: every... ...APIs and streaming services for real-time inference, speech, and vision. ~Own production...SeniorFull timeFlexible hours
- Role Description The Senior Applied AI Engineer is a hands-on software engineer specializing in the practical application of AI within production systems. This role sits at the intersection of software engineering, system design, and applied machine learning, with a focus...SeniorFull time
- ...to help shape how Pleo builds AI-powered product features, working alongside software engineers, data engineers and data scientists... ...You'll be reporting to the Senior Manager for Data & AI Products... ...distinctive skills while you bring applied AI engineering and data context...SeniorFull time
- ...of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by... ...world's most innovative companies unlock the power of AI. As a Senior Applied AI Engineer on our Cortex AI team, you will be a hands-on technical...SeniorFull time
- Role Description LTS is seeking a highly skilled Senior Applied AI Engineer to focus on continuously improving the intelligence behind the platform. You'll experiment with models, optimize retrieval strategies, refine agent reasoning, evaluate AI performance, and transform...SeniorFull time
- ...their health goals through sustainable behavioral change. As a Senior AI Engineer at Omada, you will lead the deployment of sophisticated AI... ...deep technical expertise with a practical understanding of applying AI in impactful ways. Qualifications ~5+ years of experience...SeniorFull timeRemote workWork from homeFlexible hours
- Role Description Building production systems powered by software engineering, data science and modern AI. We are seeking a hands-on Senior Applied AI Software Engineer to design, build and deploy production systems that combine software engineering, data science, large...SeniorRemote work
$162k - $190k
...boundaries of what’s possible. We invest heavily in advanced AI capabilities—specifically our Process Intelligence Graph—to turn... ...for people, companies, and the planet.The Role:As a Lead Applied Value Engineer, you are spearheading our mission of solving critical operational...SeniorFull timeWork experience placementWork at officeLocal areaImmediate startRemote workWorldwideFlexible hours- ...Applied AI Engineer Location: Onsite 5 days a week in Austin, TX, relocation is offered. Level is a learning technology company dedicated... ...experience (AWS or GCP) and containerized deployment. Senior or staff-level software engineering foundation (formal or...SeniorFull timeRelocation
$160k - $190k
...a difference while enjoying the journey, come join us and let's Tango! Role Summary: We are looking for a Senior Applied AI Engineer to help build and ship Tango's first AI-powered product, marking an important next chapter after 18 years as an industry...SeniorFull timeWork at officeRemote workVisa sponsorshipWork visaFlexible hours- Role Description Foresite is looking for a Senior Applied AI Engineer to shape how state-of-the-art AI and autonomous agents power our Managed Security Services Provider (MSSP) platform. We are building the next generation of AI-driven security operations, transforming...SeniorFull timeTemporary work
$103k - $143k
...require 5+ years of professional software engineering experience, including at least 2+ years... ...proactive interest in staying current with AI advances. We need the ability to... ...Group acts as our internal incubator for applied AI, using the latest large language models...SeniorFull timeRelocationHome officeFlexible hours$220k - $300k
...software company in NYC. They move fast and ship daily. AI-assisted engineering isn't a talking point here; agents like Devin, Claude, and... ...ship multiples of what they could alone. About the Team Applied AI owns the intelligence layer, including aspects of...SeniorAfternoon shift- ...About Qualitate Qualitate is building the AI-native primary intelligence platform for enterprises, investment firms, and... ...venture capital firms. The Role We’re looking for a Senior Applied AI Engineer to help build the LLM and retrieval systems at the core of our...SeniorFull timeFlexible hours
$8k
...Top Secret (TS/SCI) clearance with polygraph is required. Visionist has an exciting new, fully FUNDED opportunity for a Senior Applied AI Engineer - Software Engineering on our largest PRIME contract. Our team of Analysts and Engineers is motivated by the direct...SeniorPermanent employmentFull timeContract workTemporary workImmediate startFlexible hours- ...builds the world's largest AI chip, 56 times larger... ...-leading training and inference speeds; over 10 times... ...working directly with Engineering, Product, Infrastructure... ...functional ritual involving senior engineers and LTDirect... ...work at Cerebras here! Apply today and become part...SeniorRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Staff Applied AI Inference Engineer. Be the first to apply!
- project engineer assistant project manager Remote
- senior staff systems engineer Remote
- staff data engineer Remote
- assistant chief engineer Remote
- assistant engineer Remote
- engineering aide Remote
- software engineer staff Remote
- staff design engineer Remote
- senior staff engineer Remote
- staff security engineer Remote


