Staff + Sr. Software Engineer, Cloud Inference Launch Engineering
$320kUnited States Digital Space LLC
About the company
the company’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems.
About the role
The Cloud Inference team scales and optimizes Claude to serve the massive audiences of developers and enterprise companies across AWS, GCP, Azure, and future cloud service providers (CSPs). We own the end-to-end product of Claude on each cloud platform, from API integration and intelligent request routing to inference execution, capacity management, and day-to-day operations.
Within Cloud Inference, the model & inference launch team owns the validation pipeline for our inference server and load balancer on these platforms. We're responsible for every inference change — model launches, performance improvements, safeguard integrations — landing on cloud platforms with correctness, performance, and reliability intact.
This is high-leverage infrastructure work: validation has to be fast and cheap enough to run on the same accelerators that serve customers, trustworthy enough to replace manual checks, and consistent enough that a change working on the company first-party means it works everywhere. This directly determines how fast frontier models and features ship to every cloud platform, and how quickly performance wins reach production — reclaiming capacity at a time when compute is our scarcest resource.
Key responsibilities
- Be on the critical path for frontier model launches, bringing up inference for new model architectures and shipping them to cloud platforms in lockstep with our first-party platform
- Work with the core inference team to bring new inference features (e.g. structured sampling, prompt caching, and more) to cloud platforms, owning the platform-specific integration that gets them to production
- Identify and dive deep on the gaps that make inference behave differently across first-party and CSPs — config drift, observability, deployment patterns, hard cross-platform bugs — and fix them at the source rather than building platform-specific workarounds
- Design, build, and own the CI/CD infrastructure for the inference server and load balancer across cloud platforms, with shadow traffic, performance baselines (throughput and latency), and correctness checks that catch regressions before production
- Drive down merge-to-production cycle time by making validation faster, more parallel, and cost-effective enough to run on the same constrained accelerator pool that serves customers, without trading away reliability
- Analyze observability data across providers to identify performance bottlenecks, cost anomalies, and regressions, and drive remediation based on real-world production workloads
Minimum qualifications
- Have a strong interest in LLM serving; prior inference or ML experience is not required
- Have significant software engineering experience, with a strong background in high-performance, large-scale distributed systems serving millions of users
- Have a track record of building automation or test infrastructure that measurably improved release velocity or reliability
- Have experience building or operating services on at least one major cloud platform (AWS, GCP, or Azure), with exposure to Kubernetes, Infrastructure as Code, or container orchestration
- Thrive in cross-functional collaboration with both internal teams and external partners
- Are a fast learner who can quickly ramp up on new technologies, hardware platforms and provider ecosystems
- Are highly autonomous and take ownership of problems end-to-end, including work that falls outside your job description
Preferred qualifications
- LLM inference optimization, batching, and caching strategies
- Capacity-constrained scheduling or shared-resource test infrastructure
- Solid understanding of multi-region deployments, request routing, load balancing, global traffic management
- Working with CSP partner teams to scale infrastructure across multiple platforms, navigating differences in networking, security, privacy, and managed service
- Proficiency in Python or Rust
The annual compensation range for this role is listed below.
For sales roles, the range provided is the role’s On Target Earnings ("OTE") range, meaning that the range includes both the sales commissions/sales bonuses target and annual base salary for the role.
Annual Salary:
$320,000—$485,000 USD
Logistics
Minimum education: Bachelor’s degree or an equivalent combination of education, training, and/or experience
Required field of study: A field relevant to the role as demonstrated through coursework, training, or professional experience
Minimum years of experience: Years of experience required will correlate with the internal job level requirements for the position
Location-based hybrid policy: Currently, we expect all staff to be in one of our offices at least 25% of the time. However, some roles may require more time in our offices.
Visa sponsorship: We do sponsor visas! However, we’re not able to successfully sponsor visas for every role and every candidate. But if we make you an offer, we will make every reasonable effort to get you a visa, and we retain an immigration lawyer to help with this.
How we're different
We believe that the highest-impact AI research will be big science. At the company we work as a single cohesive team on just a few large-scale research efforts. And we value impact — advancing our long-term goals of steerable, trustworthy A
#J-18808-Ljbffr$168.1k - $227.4k
...Sr. Software Development Engineer, Inference Team - AWS Neuron AWS Neuron is the complete software stack for AWS... ...AWS purpose-built accelerators for cloud-scale machine learning. This senior... ...other employees, supervisors, and staff; adhere to standards of excellence...SeniorWork experience placementInternshipLocal areaFlexible hours- ...Google LLC is hiring a Senior Staff Software Engineer for Google Cloud Business Platform. You will define multi-year architectural strategy for pre-production environments and automated testing ecosystems, and drive a brand-new testing platform for GCP Lead to Cash systems...Senior
- ...United States Digital Space LLC is seeking an engineer to join the Cloud Inference team and scale Claude across AWS, GCP, Azure, and future CSPs. You will own end-to-end inference on each cloud platform, from API integration to deployment and daily operations. You...Suggested
$174k - $299k
...fuels us to continue our growth and launch new services at the speed we have been... ...-connected world.Job Overview:As our Sr. Staff Backend Engineer for the Inventory and Promise Platform... ...10+ years of experience in backend software developmentDemonstrated experience in...SeniorTemporary workFlexible hours$262k - $364k
...Senior Staff Software Engineer, Google Cloud Business Platform Share Senior Staff Software Engineer, Google Cloud Business Platform In most instances... ...-business projects. ~ Experience with release and launch readiness, test frameworks, and SDLC management....SeniorTemporary work$229k - $343k
...addition to Bitmoji , Saturn, and other digital services.Snap Engineering teams build fun and technically sophisticated products that... ...execute with privacy at the forefront.We’re looking for a Staff Software Engineer to join the Delivery Platform team at Snap.What you...Live inWork at officeLocal area$188k - $275k
CoreWeave is The Essential Cloud for AI™. Built for pioneers by pioneers, CoreWeave... ...025. Learn more at What You'll Do: Inference Platform Team The Inference team... ...inference systems. About the role: As a Staff Software Engineer (IC5) on the Inference team, you will...Permanent employmentFull timeTemporary workCasual workWork at officeFlexible hours$405k
Staff+ Software Engineer, Privacy San Francisco, CA | New York City, NY | Seattle, WA About Anthropic... ...preserving architectures for AI training and inference systems operating at very large scale,... ...with distributed systems and cloud infrastructure at scale Experience serving...Work at officeVisa sponsorshipFlexible hours- ...industry. Our mission is to enable every engineering organization to build its own self-... ...-accelerated computing, and production software systems. We work closely with leading... ...AI Engineer to own the model training, inference, and infrastructure that power its agentic...Full time
$168.1k - $227.4k
...available, scalable, distributed engineering systems for one of the... ...efficient analytics platform’. A Sr SDE on this team has a unique... ...expertise to solve complex software problems in a fast-paced environment... ...- Strategize, design and launch services for medium-to-large...SeniorInternshipFlexible hours- Staff Storage Software Engineer Lambda, the superintelligence cloud, is a leader in AI cloud infrastructure serving tens of thousands of customers. Our customers range... ...systems, making groundbreaking AI training and inference possible. The Lambda Infrastructure Engineering...Local areaFlexible hours
$260.1k
...“surges.”learn more about working at Coinbase. As a Senior Staff Software Engineer on thePlatform Payments team, you'll define the engineering... ...expertise in backend programming (e.g., Go, Java, Python) and cloud-native architecture. Deep money-movement architecture...SeniorLocal area- ...Shield AI, a defense-tech innovator, seeks a senior power systems engineer in the Seattle area to define and execute the Launch and Recovery Vehicle power system from concept through production. You will lead design, integration, and verification of high-power systems...Senior
$190k - $225k
...voice applications. Our models serve 600M+ inference calls monthly, process 1M+ hours of... ...apply! About the role: We're hiring a Software Engineer to help turn cutting-edge AI research into... ..., owning them from prototype through launch and ongoing iteration. Scale inference...Senior$119.8k - $234.7k
...EngineeringCompany: MicrosoftOverviewMicrosoft Silicon, Cloud Hardware, and Infrastructure Engineering (SCHIE) is the team behind Microsoft’s expanding... ...deployments using customer workloads, AI training and inference applications, and synthetic benchmarks to guide data-...SeniorOngoing contractWork at officeLocal areaWorldwide3 days per week$174k - $299k
...us to continue our growth and launch new services at the speed we... ...a part of the Observability Engineering team at Coupang to build and... ...solutions, as well as building software components from scratch. You... ...Professional certifications in cloud platforms, monitoring tools,...SeniorFull timeTemporary workWork experience placementFlexible hours$61k - $101k
...training or certification in software engineering concepts, plus 5+ years of... ...in one or more areas such as cloud platforms, AI, machine learning... ...such as model hosting, inference services, model gateways, managed... ...design through build, launch, and early operational support...Full time- Cloud Engineer Either strong Google Cloud or Azure Engineering skills (bonus if they have exposure to both) Strong CI/CD skills with preferences... ...technologies like Powershell Knowledge of configuration technologies like Ansible, Terraform or equivalent Software Technology IncSenior
$170k - $250k
Staff Software Engineer, Cloud Infrastructure & Systems Engineering, Seattle TaxBit is looking for a Staff Software Engineer to join our Systems, Security & Compliance organization and take ownership of the cloud infrastructure that powers our platform. In this role, you...Work at officeWork from homeFlexible hours- Blue Origin LLC is seeking a Mechanical Engineer III to join the team in Seattle, WA. This position is crucial for the production support... ...separation mechanisms for projects such as the New Glenn launch vehicle. The ideal candidate will have a B.S. in engineering with...Senior
$148.5k - $237.6k
Sr Software Engineer II Seattle, Washington, United States Join Axon and be a Force for Good. At Axon... ...with our ecosystem of devices and cloud software. Like our products, we work better... ...technical projects from concept to launch, ensuring solution meets business and technical...SeniorWork experience placementWork at officeRemote work$188k - $303k
...CoreWeave is The Essential Cloud for AI™. Built for pioneers by... ...CoreWeave is looking for a Sr. Engineering Manager to lead the core team for our next-generation Inference Platform. What You'll Do:... ...will lead a team of senior and staff engineers building and operating...SeniorPermanent employmentFull timeTemporary workCasual workWork at officeFlexible hours$174k - $299k
...us to continue our growth and launch new services at the speed we... ...Collaboration:• Work closely with other engineering teams to understand and meet... ...with hardware and software vendors to evaluate and integrate... ...working in cloud environments, particularly AWSDemonstrated...SeniorTemporary workFlexible hours- ...team owns the core of Remitly's software development lifecycle... ...AI-assisted tooling that both engineers and non-engineer Builders now... ...and confidence. As a Senior or Staff Software Development Engineer... ...tools, build/release systems, or cloud infrastructure platforms...Full time
$193.8k - $285k
...systems and abstractions that DoorDash Engineering depends on: reliable, efficient, secure... ...under exponential load. We're hiring a Staff Software Engineer to be the domain specialist for... ...architectural decisions.Kubernetes and cloud-native infrastructure; strong...Hourly payWork at officeLocal areaImmediate startRemote workFlexible hours- ...Amazon Kuiper Manufacturing Enterprises LLC is seeking a Software Development Engineer for the Leo Regulus AI-native team. You will design, build... ...operate autonomous AI agents that automate end-to-end ecommerce launch workflows across 100+ markets. You’ll work on scalable...
- ...intelligence. As a Machine Learning and System Optimization Engineer, you will orchestrate and allocate overall system capacity to... ...well as drive large initiatives that allow for more efficient inference by sharing various parts of the perception stack with one another...SeniorTemporary workRelocation package
$61k - $101k
...require formal training or certification in software engineering concepts, along with 5+ years of... ...oversight. We require strong knowledge of cloud delivery models such as IaaS, PaaS,... ...architecture, training, and inference. We require experience with Infrastructure...SeniorFull timeFor contractors$175k - $185k
...You'll Be Doing: Sr. Director, AI Platform Engineering Job Summary The... ..., delivering the software, and operating it in... ...organizations on the underlying cloud, messaging and... ...and technical staff, foster leadership and... ...least one initiative launched from inception...SeniorFull timeImmediate startWork from home$143.7k - $194.4k
...very rapid.The right candidate will possess proven software engineering skills, with experience creating and launching large distributed systems with the help of a... ...fundamentals, including architecture, training/inference lifecycles, and optimization of model executionAmazon...InternshipFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Staff + Sr. Software Engineer, Cloud Inference Launch Engineering. Be the first to apply!
- cloud engineer Washington State
- cloud developer Washington State
- informatica cloud developer Washington State
- senior cloud network engineer Washington State
- cloud architect Washington State
- senior aws cloud engineer Washington State
- senior software engineer ruby on rails Washington State
- senior compensation manager Washington State
- senior manager Washington State
- senior living Washington State




