Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Software Engineer, Inference AI/ML

$92k - $135k
Full-time

Coreweave


CoreWeave is the AI Hyperscaler™, delivering a cloud platform of cutting edge services powering the next wave of AI. Our technology provides enterprises and leading AI labs with the most performant, efficient and resilient solutions for accelerated computing. Since 2017, CoreWeave has operated a growing footprint of data centers covering every region of the US and across Europe. CoreWeave was ranked as one of the TIME100 most influential companies of 2024.

As the leader in the industry, we thrive in an environment where adaptability and resilience are key. Our culture offers career-defining opportunities for those who excel amid change and challenge. If you’re someone who thrives in a dynamic environment, enjoys solving complex problems, and is eager to make a significant impact, CoreWeave is the place for you. Join us, and be part of a team solving some of the most exciting challenges in the industry.

CoreWeave powers the creation and delivery of the intelligence that drives innovation. 

What You’ll Do:


Join the Inference team to ship production features that improve latency, reliability, and cost for model serving on our GPU platform. As an IC1, you’ll implement well-scoped changes, learn our operational practices, and grow quickly with mentorship from experienced engineers.

About the role:



  • Implement well-scoped features and fixes in Python/Go/C++ for model-serving services (e.g., Triton, vLLM, TensorRT-LLM, Ray Serve).

  • Write tests, code comments, and short design docs; participate in code reviews.

  • Add basic metrics and dashboards; assist with alarms and runbooks.

  • Follow on-call runbooks and learn incident response in a guided rotation.

  • Contribute to performance experiments (e.g., request batching, concurrency, caching) with guidance.

Who You Are:



  • BS/MS in CS, EE, or related field, or equivalent practical experience.

  • Foundations in data structures, algorithms, and networked services.
    Experience with Python or Go (C++ a plus) and Linux fundamentals; Git/CI basics.
    Exposure to containers and Kubernetes (coursework or projects welcome).
    Curiosity about GPU inference concepts (micro-batching, KV cache, streaming).


Preferred:


  • Internship or project that deployed a microservice or ML inference demo.

  • Coursework/research with PyTorch or TensorFlow; simple CUDA projects a plus.

  • Familiarity with Grafana/Prometheus/OpenTelemetry or similar tooling.

Why CoreWeave?


At CoreWeave, we work hard, have fun, and move fast! We’re in an exciting stage of hyper-growth that you will not want to miss out on. We’re not afraid of a little chaos, and we’re constantly learning. Our team cares deeply about how we build our product and how we work together, which is represented through our core values: 


  • Be Curious at Your Core

  • Act Like an Owner

  • Empower Employees

  • Deliver Best-in-Class Client Experiences

  • Achieve More Together

We support and encourage an entrepreneurial outlook and independent thinking. We foster an environment that encourages collaboration and provides the opportunity to develop innovative solutions to complex problems. As we get set for take off, the growth opportunities within the organization are constantly expanding. You will be surrounded by some of the best talent in the industry, who will want to learn from you, too. Come join us!

 

The base salary range for this role is $92,000 to $135,000. The starting salary will be determined based on job-related knowledge, skills, experience, and market location. We strive for both market alignment and internal equity when determining compensation. In addition to base salary, our total rewards package includes a discretionary bonus, equity awards, and a comprehensive benefits program (all based on eligibility). 

What We Offer

The range we’ve posted represents the typical compensation range for this role. To determine actual compensation, we review the market rate for each candidate which can include a variety of factors. These include qualifications, experience, interview performance, and location.

In addition to a competitive salary, we offer a variety of benefits to support your needs, including:


  • Medical, dental, and vision insurance - 100% paid for by CoreWeave

  • Company-paid Life Insurance 

  • Voluntary supplemental life insurance 

  • Short and long-term disability insurance 

  • Flexible Spending Account

  • Health Savings Account

  • Tuition Reimbursement 

  • Ability to Participate in Employee Stock Purchase Program (ESPP)

  • Mental Wellness Benefits through Spring Health 

  • Family-Forming support provided by Carrot

  • Paid Parental Leave 

  • Flexible, full-service childcare support with Kinside

  • 401(k) with a generous employer match

  • Flexible PTO

  • Catered lunch each day in our office and data center locations

  • A casual work environment

  • A work culture focused on innovative disruption

Our Workplace

While we prioritize a hybrid work environment, remote work may be considered for candidates located more than 30 miles from an office, based on role requirements for specialized skill sets. New hires will be invited to attend onboarding at one of our hubs within their first month. Teams also gather quarterly to support collaboration

California Consumer Privacy Act - California applicants only

CoreWeave is an equal opportunity employer, committed to fostering an inclusive and supportive workplace. All qualified applicants and candidates will receive consideration for employment without regard to race, color, religion, sex, disability, age, sexual orientation, gender identity, national origin, veteran status, or genetic information.

As part of this commitment and consistent with the Americans with Disabilities Act (ADA) , CoreWeave will ensure that qualified applicants and candidates with disabilities are provided reasonable accommodations for the hiring process, unless such accommodation would cause an undue hardship. If reasonable accommodation is needed, please contact: View email address on jobs.jobcopilot.com .

 

Export Control Compliance

This position requires access to export controlled information. To conform to U.S. Government export regulations applicable to that information, applicant must either be (A) a U.S. person, defined as a (i) U.S. citizen or national, (ii) U.S. lawful permanent resident (green card holder), (iii) refugee under 8 U.S.C. § 1157, or (iv) asylee under 8 U.S.C. § 1158, (B) eligible to access the export controlled information without a required export authorization, or (C) eligible and reasonably likely to obtain the required export authorization from the applicable U.S. government agency. CoreWeave may, for legitimate business reasons, decline to pursue any export licensing process.

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Software Engineer, Inference AI/ML in Washington DC vacancy
  •  ...Model Optimization & Deployment Engineer, you will focus on bringing...  ...vehicle SOCs. You will optimize the ML models, write custom CUDA...  ..., and build highly concurrent inference code to ensure real-time, deterministic...  ...maximize memory bandwidth on AI accelerators. Write... 
    Suggested
    Temporary work
    Relocation package

    Zoox

    Washington DC
    a month ago
  • $141.8k - $173.3k

     ...donation matchIMPACT YOU’LL MAKE:As a Sr. AI Software Engineer, you’ll play a key role in bringing...  ...high‑impact components that integrate AI/ML models or AI service APIs.Design...  ...such as model performance degradation, inference bottlenecks, prompt optimization challenges... 
    Suggested
    Full time
    Remote work

    Boeing Employees Credit Union

    Washington DC
    1 day ago
  • $188k - $275k

     ...Description CoreWeave is The Essential Cloud for AI™. Built for pioneers by pioneers,...  ...'ll Do Description of the team: The Inference team is responsible for delivering high-performance...  ...: We are looking for an Applied AI Engineer to help us understand, measure, and... 
    Suggested
    Permanent employment
    Full time
    Temporary work
    Casual work
    Work at office
    Flexible hours

    CoreWeave

    Washington DC
    12 days ago
  • $185.4k - $232.05k

     ...security solutions to deliver AI-driven, actionable...  ...in 2024. Product & Software Excellence: We were...  ...seeking a Senior Software Engineer, Applied AI to own end-...  ...engineers, with the ML/LLM Ops platform you deploy...  ...vector databases, and inference/serving platforms.... 
    Suggested
    Full time
    Work at office
    Flexible hours

    LVT

    Washington DC
    a month ago
  • $165k - $250k

     ...Software Engineer, Applied AI Washington, DC State Affairs is the nation's leading news and policy...  ...development Knowledge of various AI/ML concepts such as computer vision, image processing, statistical modeling/inference, data mining, natural language processing... 
    Suggested
    Work experience placement
    Work at office
    Local area

    State Affairs

    Washington DC
    2 days ago
  • $229.9k - $262.4k

    Senior Lead AI Engineer (FM Hosting, LLM Inference) Overview: At Capital One, we are creating...  ...real time, our applications of AI & ML are bringing humanity and simplicity...  ..., test, deploy, and support AI software components including foundation... 
    Full time
    Part time
    Local area

    Capital One Financial Corporation

    McLean, VA
    1 day ago
  •  ...Software Engineer II - Backend/Platform Agentic AI Mastercard is a global technology company in the payments industry...  ...Hands-on experience in applied AI/ML (LLM integration, RAG pipelines,...  ...agentic workflows, model serving, or inference services) Familiar with... 
    Worldwide

    Dynamic Yield

    Arlington, VA
    2 days ago
  •  ...Expression is seeking an experienced Senior AI Software Engineer and Technical Lead to lead the...  ...intelligent data processing, distributed inference, and resilient edge computing in communication...  .... Experience integrating AI/ML inference into operational software systems... 
    For contractors
    Work at office
    Immediate start
    Remote work

    Expression

    Washington DC
    5 days ago
  •  ...AI/ML Software Engineer At Gallatin, we are rebuilding logistics infrastructure for the national security missions of the United States and...  ...evaluating and deploying large scale ML pipelines and real-time inference systems—while collaborating with cross-functional teams to... 
    Local area

    Gallatin AI, Inc.

    Washington DC
    4 days ago
  • $229.9k - $262.4k

    Sr. Lead AI Engineer (Inference Optimization, FM hosting, AI Platform) Overview: At Capital One, we...  ...in real time, our applications of AI & ML are bringing humanity and simplicity to...  ...develop, test, deploy, and support AI software components including foundation model training... 
    Full time
    Part time
    Local area

    Capital One

    McLean, VA
    4 days ago
  •  ...government applications. Our engineers and space scientists...  ...a Senior Full‑Stack AI Application Engineer with...  ...machine learning inference for real‑time analysis...  ...5 years of full‑stack software development experienceAt...  ...RealityKit, ARKitCore ML, on‑device inference frameworksReact... 
    Work at office
    Local area

    AST SpaceMobile

    Lanham, MD
    3 days ago
  • $167.85k - $209.75k

     ...Senior Software Engineer, AI/ML Platforms and Infrastructure – Brain Health Accelerator The Allen Institut e accelerates science for a healthier...  ..., design, delivery, and operation of AI training and inference services and infrastructure to support modelling efforts... 
    Work at office
    Local area
    Remote work
    Visa sponsorship
    Work visa
    Relocation package

    Allen Institute

    Washington DC
    23 days ago
  •  ...thriving multi-discipline, specialty engineering services and consulting firm,...  ...a capable and motivated Software Engineer to join our team. If...  ...engineering team building cutting-edge AI applications for the nuclear...  ...full-stack technologies.AI/ ML Integration: Implement... 

    MPR Associates

    Alexandria, VA
    7 hours ago
  •  ...to deploying reliable inference and retrieval systems in...  ...closely with product and engineering to translate real...  ...needs into high-performing AI features, operating across...  ...with engineers and non-ML partners, and you write...  ...years of professional software development experience... 
    Full time
    Contract work
    Flexible hours

    Twenty Inc.

    Arlington, VA
    3 days ago
  •  ...Title: Senior Embodied AI Engineer About Us: UnitX builds the world's leading physical AI systems to automate repetitive...  ...Familiarity with deploying, profiling, and optimizing ML models for real-time inference on robotic hardware (e.g., NVIDIA Jetson, TensorRT, CUDA... 
    Full time

    Unitx

    Washington DC
    1 day ago
  •  ...technology to make systems simpler, faster, and more human. As a Software Engineer - AI, you'll design and deliver end-to-end generative AI and large...  ...: Contribute to the vision, goals, and roadmap for AI/ML development at Promise as an individual contributor with meaningful... 
    Permanent employment
    Full time
    Work at office
    Local area
    Flexible hours

    Promise Co.

    Washington DC
    1 day ago
  • $166k - $203k

     ...Software Engineer, Applied AI HackerOne is revolutionizing offensive security by combining human intelligence with artificial intelligence to help...  ...AI models into applications Hands-on experience with ML frameworks such as PyTorch, TensorFlow, or HuggingFace Transformers... 
    Apprenticeship
    Work at office
    Local area
    Remote work
    Flexible hours
    1 day per week

    HackerOne

    Washington DC
    4 days ago
  • $190k - $230k

     ...Senior Software Engineer, Applied AI At HackerOne, we're revolutionizing offensive security by combining human intelligence with artificial intelligence...  ...implementing features within production-grade AI or ML systems, including integrating LLMs or generative AI models... 
    Apprenticeship
    Work at office
    Local area
    Remote work
    Flexible hours
    1 day per week

    HackerOne

    Washington DC
    4 days ago
  • $132.5k - $338.3k

     ...forefront of a new era in enterprise AI — one defined not by model...  ...AI research and production engineering — investigating the foundational...  ..., model selection and inference routing strategies, autonomy and...  ...services, research prototypes, or AI/ML systems. Minimum of 5 years... 
    Full time
    Work experience placement
    Live in
    Work at office
    Local area
    Relocation

    Accenture

    Arlington, VA
    1 day ago
  •  ...JOB DESCRIPTION Iron EagleX is seeking a Senior Agentic AI Full Stack Software Engineer to support our AI team in Crystal City, VA. This role will...  .../versioning, reproducible environments, and CI/CD for ML-enabled systems.  WHAT YOU’LL NEED TO SUCCEED ~ Clearance... 
    Contract work

    General Dynamics Information Technology

    Arlington, VA
    a month ago
  • $180k - $260k

     ...We own the data centres, software, and applications that power today’s AI stack using sustainable...  ...for a Senior AI Product Engineer to join our product...  ...integrated product features (LLM inference APIs, fine-tuning UX,...  ...Working knowledge of AI/ML infrastructure concepts:... 
    Full time
    Contract work
    Flexible hours

    Nscale

    Washington DC
    1 day ago
  • $109k - $203k

    Software Engineer - AI - CoCounsel Forward Deployed EngineeringAre you excited about building AI solutions that help legal professionals work faster...  ...-on experience with Python and familiarity with modern AI/ML tooling and APIs.Understanding of core machine learning and LLM... 
    Full time
    Contract work
    Work at office
    Local area
    Flexible hours

    Thomson Reuters

    Washington DC
    2 days ago
  • $131.3k - $237.35k

     ...of deploying enterprise-scale AI, data, and mission platform capabilities...  ...the next level.As a Senior AI Engineer, you will:Support the...  ..., developing analytics and AI/ML solutions (4+ years with a...  ...compatible APIsPython developmentGPU inference optimizationPreferred... 
    Full time
    Remote work
    Flexible hours

    Leidos

    Alexandria, VA
    4 days ago
  • $103.2k - $203.4k

     ...the government forward! Build AI that matters . We ship...  ...for building and integrating AI/ML applications. Owned AI solutions...  ...cloud AI services or on prem inference stacks Background in LLM...  ...or opensource; mentorship of engineers. Clear communication with engineers... 
    Live in
    Work at office
    Local area

    Accenture

    Washington DC
    3 days ago
  •  .... Your Role ~ The Software Architect is responsible...  ...on Java-based systems, AI-enabled capabilities,...  ...partners closely with engineering, product, security, and...  ...Architect and integrate AI/ML capabilities (e.g.,...  ...integration, LLM/RAG patterns, inference services) into... 
    Flexible hours

    IntraFi

    Arlington, VA
    a month ago
  •  ...Associate Software Engineer – Full Stack (AI First) We are seeking an enthusiastic and motivated Associate Software Engineer to join our Master Data...  ...clauses, failure atomicity) Basic knowledge of Python or AI/ML concepts Familiarity with CI/CD pipelines and DevOps... 

    PlaceIQ

    Washington DC
    1 day ago
  • The CERT Division of the Software Engineering Institute (SEI) is seeking applicants for the role of AI Security Software Engineer. Established in response to the Morris worm...  ..., demonstrating strong expertise in ML development and deploymentCollaborate with researchers... 
    Full time
    Work experience placement
    Relocation package

    Carnegie Mellon University

    Arlington, VA
    1 day ago
  •  ...Senior Software Engineer - Backend/Platform Agentic AI Mastercard is a global technology company in the payments industry. Our mission is to connect and...  ...About You: Proven experience productionizing AI/ML systems, delivering reliable, scalable services used in... 
    Worldwide

    Dynamic Yield

    Arlington, VA
    2 days ago
  •  ...businesses manage networks. There AI Core group pioneers’...  ...The Role As one of our AI ML Engineer’s, you'll be a key technical...  ...platforms Own the end-to-end software development lifecycle — from...  ...services Build real-time inference pipelines for complex models... 
    Full time
    Shift work

    C-serv

    Washington DC
    1 day ago
  • $190k - $230k

     ...). The HackerOne Platform unites agentic AI solutions with the ingenuity of the world...  ...inclusion, respect, and accountability. Senior Software Engineer, Applied AI Location: Seattle, WA; Austin...  ...features within production‑grade AI or ML systems, including integrating LLMs or... 
    Apprenticeship
    Work at office
    Local area
    Remote work
    Flexible hours
    Shift work
    1 day per week

    HackerOne

    Washington DC
    5 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Software Engineer, Inference AI/ML. Be the first to apply!