Software Engineer, Inference AI/ML
$92k - $135kCoreweave
CoreWeave is the AI Hyperscaler™, delivering a cloud platform of cutting edge services powering the next wave of AI. Our technology provides enterprises and leading AI labs with the most performant, efficient and resilient solutions for accelerated computing. Since 2017, CoreWeave has operated a growing footprint of data centers covering every region of the US and across Europe. CoreWeave was ranked as one of the TIME100 most influential companies of 2024.
As the leader in the industry, we thrive in an environment where adaptability and resilience are key. Our culture offers career-defining opportunities for those who excel amid change and challenge. If you’re someone who thrives in a dynamic environment, enjoys solving complex problems, and is eager to make a significant impact, CoreWeave is the place for you. Join us, and be part of a team solving some of the most exciting challenges in the industry.
CoreWeave powers the creation and delivery of the intelligence that drives innovation.
What You’ll Do:
Join the Inference team to ship production features that improve latency, reliability, and cost for model serving on our GPU platform. As an IC1, you’ll implement well-scoped changes, learn our operational practices, and grow quickly with mentorship from experienced engineers.
About the role:
- Implement well-scoped features and fixes in Python/Go/C++ for model-serving services (e.g., Triton, vLLM, TensorRT-LLM, Ray Serve).
- Write tests, code comments, and short design docs; participate in code reviews.
- Add basic metrics and dashboards; assist with alarms and runbooks.
- Follow on-call runbooks and learn incident response in a guided rotation.
- Contribute to performance experiments (e.g., request batching, concurrency, caching) with guidance.
Who You Are:
- BS/MS in CS, EE, or related field, or equivalent practical experience.
- Foundations in data structures, algorithms, and networked services.
Experience with Python or Go (C++ a plus) and Linux fundamentals; Git/CI basics.
Exposure to containers and Kubernetes (coursework or projects welcome).
Curiosity about GPU inference concepts (micro-batching, KV cache, streaming).
Preferred:
- Internship or project that deployed a microservice or ML inference demo.
- Coursework/research with PyTorch or TensorFlow; simple CUDA projects a plus.
- Familiarity with Grafana/Prometheus/OpenTelemetry or similar tooling.
Why CoreWeave?
At CoreWeave, we work hard, have fun, and move fast! We’re in an exciting stage of hyper-growth that you will not want to miss out on. We’re not afraid of a little chaos, and we’re constantly learning. Our team cares deeply about how we build our product and how we work together, which is represented through our core values:
- Be Curious at Your Core
- Act Like an Owner
- Empower Employees
- Deliver Best-in-Class Client Experiences
- Achieve More Together
We support and encourage an entrepreneurial outlook and independent thinking. We foster an environment that encourages collaboration and provides the opportunity to develop innovative solutions to complex problems. As we get set for take off, the growth opportunities within the organization are constantly expanding. You will be surrounded by some of the best talent in the industry, who will want to learn from you, too. Come join us!
The base salary range for this role is $92,000 to $135,000. The starting salary will be determined based on job-related knowledge, skills, experience, and market location. We strive for both market alignment and internal equity when determining compensation. In addition to base salary, our total rewards package includes a discretionary bonus, equity awards, and a comprehensive benefits program (all based on eligibility).
What We Offer
The range we’ve posted represents the typical compensation range for this role. To determine actual compensation, we review the market rate for each candidate which can include a variety of factors. These include qualifications, experience, interview performance, and location.
In addition to a competitive salary, we offer a variety of benefits to support your needs, including:
- Medical, dental, and vision insurance - 100% paid for by CoreWeave
- Company-paid Life Insurance
- Voluntary supplemental life insurance
- Short and long-term disability insurance
- Flexible Spending Account
- Health Savings Account
- Tuition Reimbursement
- Ability to Participate in Employee Stock Purchase Program (ESPP)
- Mental Wellness Benefits through Spring Health
- Family-Forming support provided by Carrot
- Paid Parental Leave
- Flexible, full-service childcare support with Kinside
- 401(k) with a generous employer match
- Flexible PTO
- Catered lunch each day in our office and data center locations
- A casual work environment
- A work culture focused on innovative disruption
Our Workplace
While we prioritize a hybrid work environment, remote work may be considered for candidates located more than 30 miles from an office, based on role requirements for specialized skill sets. New hires will be invited to attend onboarding at one of our hubs within their first month. Teams also gather quarterly to support collaboration
California Consumer Privacy Act - California applicants only
CoreWeave is an equal opportunity employer, committed to fostering an inclusive and supportive workplace. All qualified applicants and candidates will receive consideration for employment without regard to race, color, religion, sex, disability, age, sexual orientation, gender identity, national origin, veteran status, or genetic information.
As part of this commitment and consistent with the Americans with Disabilities Act (ADA) , CoreWeave will ensure that qualified applicants and candidates with disabilities are provided reasonable accommodations for the hiring process, unless such accommodation would cause an undue hardship. If reasonable accommodation is needed, please contact: View email address on jobs.jobcopilot.com .
Export Control Compliance
This position requires access to export controlled information. To conform to U.S. Government export regulations applicable to that information, applicant must either be (A) a U.S. person, defined as a (i) U.S. citizen or national, (ii) U.S. lawful permanent resident (green card holder), (iii) refugee under 8 U.S.C. § 1157, or (iv) asylee under 8 U.S.C. § 1158, (B) eligible to access the export controlled information without a required export authorization, or (C) eligible and reasonably likely to obtain the required export authorization from the applicable U.S. government agency. CoreWeave may, for legitimate business reasons, decline to pursue any export licensing process.
- ...Model Optimization & Deployment Engineer, you will focus on bringing... ...vehicle SOCs. You will optimize the ML models, write custom CUDA... ..., and build highly concurrent inference code to ensure real-time, deterministic... ...maximize memory bandwidth on AI accelerators. Write...SuggestedTemporary workRelocation package
$141.8k - $173.3k
...donation matchIMPACT YOU’LL MAKE:As a Sr. AI Software Engineer, you’ll play a key role in bringing... ...high‑impact components that integrate AI/ML models or AI service APIs.Design... ...such as model performance degradation, inference bottlenecks, prompt optimization challenges...SuggestedFull timeRemote work$188k - $275k
...Description CoreWeave is The Essential Cloud for AI™. Built for pioneers by pioneers,... ...'ll Do Description of the team: The Inference team is responsible for delivering high-performance... ...: We are looking for an Applied AI Engineer to help us understand, measure, and...SuggestedPermanent employmentFull timeTemporary workCasual workWork at officeFlexible hours$185.4k - $232.05k
...security solutions to deliver AI-driven, actionable... ...in 2024. Product & Software Excellence: We were... ...seeking a Senior Software Engineer, Applied AI to own end-... ...engineers, with the ML/LLM Ops platform you deploy... ...vector databases, and inference/serving platforms....SuggestedFull timeWork at officeFlexible hours$165k - $250k
...Software Engineer, Applied AI Washington, DC State Affairs is the nation's leading news and policy... ...development Knowledge of various AI/ML concepts such as computer vision, image processing, statistical modeling/inference, data mining, natural language processing...SuggestedWork experience placementWork at officeLocal area$229.9k - $262.4k
Senior Lead AI Engineer (FM Hosting, LLM Inference) Overview: At Capital One, we are creating... ...real time, our applications of AI & ML are bringing humanity and simplicity... ..., test, deploy, and support AI software components including foundation...Full timePart timeLocal area- ...Software Engineer II - Backend/Platform Agentic AI Mastercard is a global technology company in the payments industry... ...Hands-on experience in applied AI/ML (LLM integration, RAG pipelines,... ...agentic workflows, model serving, or inference services) Familiar with...Worldwide
- ...Expression is seeking an experienced Senior AI Software Engineer and Technical Lead to lead the... ...intelligent data processing, distributed inference, and resilient edge computing in communication... .... Experience integrating AI/ML inference into operational software systems...For contractorsWork at officeImmediate startRemote work
- ...AI/ML Software Engineer At Gallatin, we are rebuilding logistics infrastructure for the national security missions of the United States and... ...evaluating and deploying large scale ML pipelines and real-time inference systems—while collaborating with cross-functional teams to...Local area
$229.9k - $262.4k
Sr. Lead AI Engineer (Inference Optimization, FM hosting, AI Platform) Overview: At Capital One, we... ...in real time, our applications of AI & ML are bringing humanity and simplicity to... ...develop, test, deploy, and support AI software components including foundation model training...Full timePart timeLocal area- ...government applications. Our engineers and space scientists... ...a Senior Full‑Stack AI Application Engineer with... ...machine learning inference for real‑time analysis... ...5 years of full‑stack software development experienceAt... ...RealityKit, ARKitCore ML, on‑device inference frameworksReact...Work at officeLocal area
$167.85k - $209.75k
...Senior Software Engineer, AI/ML Platforms and Infrastructure – Brain Health Accelerator The Allen Institut e accelerates science for a healthier... ..., design, delivery, and operation of AI training and inference services and infrastructure to support modelling efforts...Work at officeLocal areaRemote workVisa sponsorshipWork visaRelocation package- ...thriving multi-discipline, specialty engineering services and consulting firm,... ...a capable and motivated Software Engineer to join our team. If... ...engineering team building cutting-edge AI applications for the nuclear... ...full-stack technologies.AI/ ML Integration: Implement...
- ...to deploying reliable inference and retrieval systems in... ...closely with product and engineering to translate real... ...needs into high-performing AI features, operating across... ...with engineers and non-ML partners, and you write... ...years of professional software development experience...Full timeWork at officeFlexible hours
- ...Title: Senior Embodied AI Engineer About Us: UnitX builds the world's leading physical AI systems to automate repetitive... ...Familiarity with deploying, profiling, and optimizing ML models for real-time inference on robotic hardware (e.g., NVIDIA Jetson, TensorRT, CUDA...Full time
- ...technology to make systems simpler, faster, and more human. As a Software Engineer - AI, you'll design and deliver end-to-end generative AI and large... ...: Contribute to the vision, goals, and roadmap for AI/ML development at Promise as an individual contributor with meaningful...Permanent employmentFull timeWork at officeLocal areaFlexible hours
$166k - $203k
...Software Engineer, Applied AI HackerOne is revolutionizing offensive security by combining human intelligence with artificial intelligence to help... ...AI models into applications Hands-on experience with ML frameworks such as PyTorch, TensorFlow, or HuggingFace Transformers...ApprenticeshipWork at officeLocal areaRemote workFlexible hours1 day per week$157.25k - $212.75k
...None Job Family: Data Science and Data Engineering Job Qualifications: Skills: AI Ops, CI/CD, Docker (Software), User Interfaces (UI), Web Applications Certifications... ..., reproducible environments, and CI/CD for ML-enabled systems. WHAT YOU'LL NEED TO...Full timeContract workTemporary workPart timeImmediate startRemote workWorldwideFlexible hours$190k - $230k
...Senior Software Engineer, Applied AI At HackerOne, we're revolutionizing offensive security by combining human intelligence with artificial intelligence... ...implementing features within production-grade AI or ML systems, including integrating LLMs or generative AI models...ApprenticeshipWork at officeLocal areaRemote workFlexible hours1 day per week$180k - $260k
...We own the data centres, software, and applications that power today’s AI stack using sustainable... ...for a Senior AI Product Engineer to join our product... ...integrated product features (LLM inference APIs, fine-tuning UX,... ...Working knowledge of AI/ML infrastructure concepts:...Full timeContract workFlexible hours$132.5k - $338.3k
...forefront of a new era in enterprise AI — one defined not by model... ...AI research and production engineering — investigating the foundational... ..., model selection and inference routing strategies, autonomy and... ...services, research prototypes, or AI/ML systems. Minimum of 5 years...Full timeWork experience placementLive inWork at officeLocal areaRelocation$109k - $203k
Software Engineer - AI - CoCounsel Forward Deployed EngineeringAre you excited about building AI solutions that help legal professionals work faster... ...-on experience with Python and familiarity with modern AI/ML tooling and APIs.Understanding of core machine learning and LLM...Full timeContract workWork at officeLocal areaFlexible hours$131.3k - $237.35k
...of deploying enterprise-scale AI, data, and mission platform capabilities... ...the next level.As a Senior AI Engineer, you will:Support the... ..., developing analytics and AI/ML solutions (4+ years with a... ...compatible APIsPython developmentGPU inference optimizationPreferred...Full timeRemote workFlexible hours$103.2k - $203.4k
...the government forward! Build AI that matters . We ship... ...for building and integrating AI/ML applications. Owned AI solutions... ...cloud AI services or on prem inference stacks Background in LLM... ...or opensource; mentorship of engineers. Clear communication with engineers...Live inWork at officeLocal area- ...Associate Software Engineer – Full Stack (AI First) We are seeking an enthusiastic and motivated Associate Software Engineer to join our Master Data... ...clauses, failure atomicity) Basic knowledge of Python or AI/ML concepts Familiarity with CI/CD pipelines and DevOps...
- The CERT Division of the Software Engineering Institute (SEI) is seeking applicants for the role of AI Security Software Engineer. Established in response to the Morris worm... ..., demonstrating strong expertise in ML development and deploymentCollaborate with researchers...Full timeWork experience placementRelocation package
- .... Your Role ~ The Software Architect is responsible... ...on Java-based systems, AI-enabled capabilities,... ...partners closely with engineering, product, security, and... ...Architect and integrate AI/ML capabilities (e.g.,... ...integration, LLM/RAG patterns, inference services) into...Flexible hours
- ...Senior Software Engineer - Backend/Platform Agentic AI Mastercard is a global technology company in the payments industry. Our mission is to connect and... ...About You: Proven experience productionizing AI/ML systems, delivering reliable, scalable services used in...Worldwide
$190k - $230k
...). The HackerOne Platform unites agentic AI solutions with the ingenuity of the world... ...inclusion, respect, and accountability. Senior Software Engineer, Applied AI Location: Seattle, WA; Austin... ...features within production‑grade AI or ML systems, including integrating LLMs or...ApprenticeshipWork at officeLocal areaRemote workFlexible hoursShift work1 day per week- Xcelerate Solutions in Bethesda, MD is seeking a TS/SCI-cleared Senior Software Engineer to design and implement a data‑centric, cloud‑based architecture leveraging AI/ML and cross‑domain transfer systems. You will join a mission‑driven team delivering secure, scalable...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Software Engineer, Inference AI/ML. Be the first to apply!
- agile software developer Washington DC
- software developer internship no experience Washington DC
- intermediate software engineer Washington DC
- software engineer staff Washington DC
- experienced software developer Washington DC
- work from home software developer Washington DC
- software developer fintech Washington DC
- software data engineer Washington DC
- financial software developer Washington DC
- software developer internship Washington DC



