Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior Staff Applied AI Inference Engineer

$250k - $300k
Full-time

Crusoe

Crusoe is on a mission to accelerate the abundance of energy and intelligence . As the only vertically integrated AI infrastructure company built from the ground up, we own and operate each layer of the stack — from electrons to tokens — to power the world's most ambitious AI workloads. When you join Crusoe, you join a team that is building the future, faster.

We're in the midst of the greatest industrial revolution of our time. The demand for AI compute is boundless, and power is a bottleneck. We're solving that — with an energy-first approach that makes AI infrastructure better for the world and faster for the people innovating with AI.

We're looking for problem-solving, opportunity-finding teammates with a sense of urgency, who believe in the scale of our ambition and thrive on a path not fully paved — people who want to grow their careers alongside a team of experts across energy, manufacturing, data center construction, and cloud services.

If you want to do the most meaningful work of your career, help our customers and partners advance their AI strategies, and be part of a high-performing team that believes in each other, come build with us at Crusoe.

About the Role

You will spend your time making large language models run faster, cheaper, and more reliably in production. That means owning the inference stack end to end: profiling where time and cost go, bringing modern optimization techniques into real deployments, and getting deep into the serving code when the defaults are not good enough. This is core systems and performance work on some of the most demanding models in use today.

The work is applied, not academic. The optimizations you build land in real customer deployments, each with its own models, traffic patterns, latency targets, and cost constraints. So while performance is the heart of the role, you will also work directly with customer engineering teams to tailor deployments to their needs, take a workload from an early proof of concept to a fully monitored production service, and make sure the gains you engineer actually show up for the people running the workload.

To set expectations clearly, this is a hands-on engineering role built around coding, profiling, and low-level optimization. It also carries a customer-facing side, along with elements of product and technical solutions work, because that is where the performance work gets proven.

What You'll Be Working On:
  • Bring current inference techniques into production and refine them.
  • Design and optimize serving architectures, including prefill and decode disaggregation, request routing, and related approaches.
  • Work down into the serving stack, from frameworks like vLLM and SGLang to the CUDA kernels underneath, profiling and running in-depth analysis to find and fix performance problems.
  • Adapt and scale optimization methods across many kinds of ML models, with an emphasis on large language models.
  • Profile and tune deployments against clear targets for latency, throughput, and cost, and keep them dependable under real traffic.
  • Tailor deployments to each customer's models and constraints, partnering with their engineering teams to move a workload from an early proof of concept through to a live, well-monitored production service.
  • Build and support the software and product features around the inference stack in a production setting, using one or more general-purpose languages, with Python preferred given how central it is to ML work.
  • Experiment quickly: take fuzzy goals, shape them into clear specs and focused proofs of concept, run fast experiments to find what works, and ship well-tested results without delay.
  • Own delivery end to end, from the first experiment through to the optimization running in production, keeping the underlying performance goals, clear specs, and follow-through front of mind, and drafting features and product requirement documents together with other engineering and product teams.
  • Work through ambiguity and make sound calls on tradeoffs and tooling, steering away from complexity that is not needed.
  • Take real pride and ownership in your work, hold yourself accountable, and look for the same from the people around you.
What You'll Bring to the Team:
  • A Bachelor's, Master's, or Ph.D. in Computer Science, Engineering, Mathematics, or a related field.
  • Hands-on experience shipping code in production with one or more general-purpose languages, such as Python or C++, with a strong preference for Python.
  • Familiarity with methods for optimizing LLMs for high throughput / low latency inference.
  • Comfort with modern LLM serving frameworks such as vLLM or SGLang, and with profiling and analyzing performance down to the kernel level.
  • A firm grasp of how GPUs are built and how they behave.
  • Clear interest and hands-on experience with large language models.
  • A working knowledge of AI/ML pipelines and the full path of developing and deploying ML models.
  • Strong communication skills, particularly when explaining hard technical topics to customers and teammates.
Bonus points:
  • A track record of making software systems run faster, especially for large language models.
  • Experience with CUDA or comparable technologies.
  • A strong command of software engineering fundamentals, with a record of building and shipping AI/ML inference systems.
  • Experience with Docker and Kubernetes.
  • Prior work building or tuning AI/ML projects, particularly in a customer-facing setting.
Benefits:
  • Competitive compensation and equity packages
  • Restricted Stock Units
  • Paid time off, paid holidays & leave of absence programs
  • Comprehensive health, dental & vision insurance
  • Employer contributions to HSA account
  • Paid parental leave
  • Paid life insurance, short-term and long-term disability
  • Professional development & tuition reimbursement
  • Mental health & wellness support
  • Commuter benefits (parking & transit)
  • Cell phone stipend
  • 401(k) Retirement plan with company match up to 4% of salary
  • Volunteer time off
  • Global travel insurance & emergency assistance
  • Daily meals allowance
  • Additional perks & programs specific to location
Compensation Range

Compensation will be paid in the range of up to $250,000 - $300,000 + Bonus. Restricted Stock Units are included in all offers. Compensation to be determined by the applicant's knowledge, education, and abilities, as well as internal equity and alignment with market data.

Crusoe is an Equal Opportunity Employer. Employment decisions are made without regard to race, color, religion, disability, genetic information, pregnancy, citizenship, marital status, sex/gender, sexual preference/ orientation, gender identity, age, veteran status, national origin, or any other status protected by law or regulation.

Vacancy posted 14 days ago
Similar jobs that could be interesting for youBased on the Senior Staff Applied AI Inference Engineer in Remote vacancy
  • $188k - $275k

     ...CoreWeave is The Essential Cloud for AI™. Built for pioneers by pioneers, CoreWeave...  ...Description of the team: The Inference team is responsible for delivering high...  ...About the role: We are looking for an Applied AI Engineer to help us understand, measure, and improve... 
    Suggested
    Permanent employment
    Full time
    Temporary work
    Casual work
    Work at office
    Flexible hours

    Core Weave

    Remote
    2 days ago
  • $200k - $250k

     ...Description About the Role Roger is an AI platform that frees home health...  ...accuracy. We are now looking for a Senior Applied AI Engineer to build the intelligence layer at the...  ...improvements. Build scalable, cost-efficient inference infrastructure with great monitoring... 
    Senior
    Remote work
    Work from home

    Roger Healthcare

    San Francisco, CA
    7 days ago
  • $175k - $200k

     ...everyone. Patients interact with advanced AI systems at every step of their care...  ...latest  publications . About the Senior Applied AI Engineer role We are hiring Senior Applied AI...  ...evaluating models, productionizing inference, and measuring real-world clinical and... 
    Senior
    Full time
    Remote work
    Work from home
    Flexible hours

    Curai

    Remote
    21 days ago
  • $170k - $233k

     ...Responsible for serving as a senior technical leader...  ...deployment of advanced AI solutions that address...  ...and internal expert in applied AI, shaping the organization...  ...(5) + years technical engineering experience with coding...  ...learning, or real-time inference systems. Track... 
    Senior
    Full time
    Local area

    Loan Depot

    Remote
    2 days ago
  • $139.2k - $208.8k

     ...aim to leave a positive mark on culture. Job Title : Senior Applied AI Engineer Team : Global Quality Engineering Location : New...  ...ML workflows o Vertex AI Online Endpoints for real-time inference ● Integrate Vertex AI with GCP services (BigQuery, Cloud... 
    Senior
    Local area
    Burbank, CA
    11 days ago
  •  ...by Fast Company.For more information visit cultureamp.com.How you can help make a better world of workWe're looking for a Senior Applied AI Engineer to join the Frontier team, helping build Agentic AI Solutions at Culture Amp. You'll work at the intersection of applied... 
    Senior
    Work at office
    Local area
    Shift work
    2 days per week

    Culture Amp

    Melbourne, FL
    3 days ago
  • $197.4k - $266.6k

     ...Forbes Best Startup Employers 2022 List.AI is transforming customer support and, with...  ...product teams to identify opportunities for applying generative AI to solve user needs and...  ...data structures, and principles of software engineering.Nice to Have:Knowledge in specialized... 
    Senior
    Work at office
    Immediate start
    Remote work
    Work from home
    Monday to Friday

    FrontApp

    San Francisco, CA
    3 days ago
  •  ...better way, and we create possibilities. Interested in joining us on our journey? We are looking for a highly motivated Senior Applied AI Engineer to join our digital innovation team and help revolutionize how we serve our primarily B2C (Business-to-Consumer) customers... 
    Senior
    Full time
    Work at office
    Remote work
    Flexible hours

    GE Appliances

    Louisville, KY
    17 hours ago
  • $177k - $226k

     ...landscape of critical care through our rapid seizure detection technology, come join the movement!Position Overview:The Senior Manager, Applied AI Engineering is a senior individual contributor role with broad ownership across Ceribell's internal AI engineering portfolio.... 
    Senior
    For contractors
    Work at office
    Local area
    Immediate start
    Remote work

    Ceribell

    Sunnyvale, CA
    4 days ago
  • $228.6k - $342.8k

     ...do this by building and running the world's best data and AI infrastructure platform so our customers can use deep data...  ...and how well it understands what it finds.We are hiring a Senior Staff Applied AI Engineer to own context retrieval for Databricks agents across SaaS... 
    Senior
    Work at office
    Local area
    Immediate start
    Worldwide

    DataBricks

    San Francisco, CA
    3 days ago
  • $104.9k - $199.07k

     ...BackgroundThe rapid evolution of artificial intelligence (AI) presents a transformative opportunity for Milliman to...  ...objectives.Role PurposeMilliman AI Solutions is seeking a Senior Insurance Applied AI Engineer to establish and refine the business engagement and engineering... 
    Senior
    Full time
    Work experience placement
    Remote work
    Worldwide

    Milliman

    Dallas, TX
    3 days ago
  •  ...Senior Applied AI Engineer Our client is a fast-growing European fintech company in the business spend management space. We're looking for a Senior Applied AI Engineer to design, build and ship customer-facing AI features end-to-end. Responsibilities Design... 
    Senior
    Freelance
    Work at office
    Immediate start
    Remote work
    Flexible hours

    N-iX

    United States
    1 day ago
  • $140k - $170k

     ...Senior Applied AI Engineer At Veracity, we aim to be a different kind of insurance partner – one that is free from outside investors, venture capital, or the pressures of a corporate parent. Ours is a culture of empowerment – one that believes in effort, results... 
    Senior
    Remote work

    Veracity Insurance

    United States
    4 days ago
  • $160k - $195k

     ...remote-first SaaS company building practical AI-powered solutions that improve real...  ...continuous learning. We are expanding our engineering team to build the next generation of AI-...  ...The Role We are looking for a Senior Applied AI Engineer to help design, build, and... 
    Senior
    Full time
    Local area
    Remote work

    Brightplan

    United States
    2 days ago
  •  ...is a hands-on, high-ownership role for an engineer who has built agents at scale and wants...  ...rather than just view it. ~Make health AI that people trust with their lives: every...  ...APIs and streaming services for real-time inference, speech, and vision. ~Own production... 
    Senior
    Full time
    Flexible hours

    Function Health

    Remote
    a month ago
  • Role Description The Senior Applied AI Engineer is a hands-on software engineer specializing in the practical application of AI within production systems. This role sits at the intersection of software engineering, system design, and applied machine learning, with a focus... 
    Senior
    Full time

    CINC Systems

    Remote
    a month ago
  •  ...to help shape how Pleo builds AI-powered product features, working alongside software engineers, data engineers and data scientists...  ...You'll be reporting to the Senior Manager for Data & AI Products...  ...distinctive skills while you bring applied AI engineering and data context... 
    Senior
    Full time

    Pleo

    Remote
    20 days ago
  •  ...of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by...  ...world's most innovative companies unlock the power of AI. As a Senior Applied AI Engineer on our Cortex AI team, you will be a hands-on technical... 
    Senior
    Full time

    Snowflake

    Remote
    2 days ago
  • Role Description LTS is seeking a highly skilled Senior Applied AI Engineer to focus on continuously improving the intelligence behind the platform. You'll experiment with models, optimize retrieval strategies, refine agent reasoning, evaluate AI performance, and transform... 
    Senior
    Full time

    LTS

    Remote
    a month ago
  •  ...their health goals through sustainable behavioral change. As a Senior AI Engineer at Omada, you will lead the deployment of sophisticated AI...  ...deep technical expertise with a practical understanding of applying AI in impactful ways. Qualifications ~5+ years of experience... 
    Senior
    Full time
    Remote work
    Work from home
    Flexible hours

    Omada Health

    Remote
    a month ago
  • Role Description Building production systems powered by software engineering, data science and modern AI. We are seeking a hands-on Senior Applied AI Software Engineer to design, build and deploy production systems that combine software engineering, data science, large... 
    Senior
    Remote work

    M3USA

    Remote
    a month ago
  • $162k - $190k

     ...boundaries of what’s possible. We invest heavily in advanced AI capabilities—specifically our Process Intelligence Graph—to turn...  ...for people, companies, and the planet.The Role:As a Lead Applied Value Engineer, you are spearheading our mission of solving critical operational... 
    Senior
    Full time
    Work experience placement
    Work at office
    Local area
    Immediate start
    Remote work
    Worldwide
    Flexible hours

    Celonis

    New York, NY
    3 days ago
  •  ...Applied AI Engineer Location: Onsite 5 days a week in Austin, TX, relocation is offered. Level is a learning technology company dedicated...  ...experience (AWS or GCP) and containerized deployment. Senior or staff-level software engineering foundation (formal or... 
    Senior
    Full time
    Relocation

    Level.

    Austin, TX
    2 days ago
  • $160k - $190k

     ...a difference while enjoying the journey, come join us and let's Tango! Role Summary: We are looking for a Senior Applied AI Engineer to help build and ship Tango's first AI-powered product, marking an important next chapter after 18 years as an industry... 
    Senior
    Full time
    Work at office
    Remote work
    Visa sponsorship
    Work visa
    Flexible hours

    Tango

    Remote
    2 days ago
  • Role Description Foresite is looking for a Senior Applied AI Engineer to shape how state-of-the-art AI and autonomous agents power our Managed Security Services Provider (MSSP) platform. We are building the next generation of AI-driven security operations, transforming... 
    Senior
    Full time
    Temporary work

    Foresite

    Remote
    11 days ago
  • $103k - $143k

     ...require 5+ years of professional software engineering experience, including at least 2+ years...  ...proactive interest in staying current with AI advances. We need the ability to...  ...Group acts as our internal incubator for applied AI, using the latest large language models... 
    Senior
    Full time
    Relocation
    Home office
    Flexible hours

    BOLD

    Miami, FL
    2 days ago
  • $220k - $300k

     ...software company in NYC.  They move fast and ship daily. AI-assisted engineering isn't a talking point here; agents like Devin, Claude, and...  ...ship multiples of what they could alone. About the Team Applied AI owns the intelligence layer, including aspects of... 
    Senior
    Afternoon shift

    Simplex

    New York, NY
    17 days ago
  •  ...About Qualitate Qualitate is building the AI-native primary intelligence platform for enterprises, investment firms, and...  ...venture capital firms. The Role We’re looking for a Senior Applied AI Engineer to help build the LLM and retrieval systems at the core of our... 
    Senior
    Full time
    Flexible hours

    Qualitate

    New York, NY
    2 days ago
  • $8k

     ...Top Secret (TS/SCI) clearance with polygraph is required.  Visionist has an exciting new, fully FUNDED opportunity for a Senior Applied AI Engineer - Software Engineering on our largest PRIME contract. Our team of Analysts and Engineers is motivated by the direct... 
    Senior
    Permanent employment
    Full time
    Contract work
    Temporary work
    Immediate start
    Flexible hours

    Visionist, Inc.

    Remote
    2 days ago
  •  ...builds the world's largest AI chip, 56 times larger...  ...-leading training and inference speeds; over 10 times...  ...working directly with Engineering, Product, Infrastructure...  ...functional ritual involving senior engineers and LTDirect...  ...work at Cerebras here! Apply today and become part... 
    Senior
    Remote work

    Cerebras Systems

    Sunnyvale, CA
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior Staff Applied AI Inference Engineer. Be the first to apply!