Cloud Inference Launch Engineer: Scale & Optimize LLM Infra
Jobleads-US
United States Digital Space LLC is seeking an engineer to join the Cloud Inference team and scale Claude across AWS, GCP, Azure, and future CSPs. You will own end-to-end inference on each cloud platform, from API integration to deployment and daily operations.
You will focus on fast, cost-efficient validation, performance improvements, and reliability to ensure consistent behavior across providers and accelerate model delivery.
#J-18808-Ljbffr Jobleads-US$320k
...of committed researchers, engineers, policy experts, and... ...About the role The Cloud Inference team scales and optimizes Claude to serve the massive... ...Inference, the model & inference launch team owns the validation... ...a strong interest in LLM serving; prior inference...CloudVisa sponsorship$152.2k - $205.9k
...building the future of cloud financial... ...in production at scale. Spend is driven... ...tokens, model choice, inference patterns, and... ...cost, usage, and optimization insights that power... ...most consequential launches, run the customer... ...directly with engineering, and dig into consumption...CloudFlexible hours$150k - $190k
...reliable open-model inference platform. Backed... ...of researchers and engineers on a mission to... ...accessibility. Our cloud provides instant access... ..., and model optimization. About The Role We... ...technical assets that scale the team: demo... ...knowledge of modern LLM inference — serving...CloudImmediate start$152k
...continue our growth and launch new services at... ...Detection Engineering (DE) team within Coupang... ...threats at scale. The team builds and... ...develops AI Agent and LLM-based analysis capabilities... .... Analyze and optimize automation... ...operating systems in cloud environments, preferably...CloudTemporary workFlexible hours$121.6k - $243.2k
...on building large-scale and highly available cloud infrastructure, which... ...technologies to optimize our AI Infra stack, including training infra, inference infra, and AI agents... ...Science, Computer Engineering, or a related discipline... ...following areas: - LLM training infra,...CloudTemporary workLocal area$191.2k - $239k
...to build the simplest scalable cloud. If you have a growth mindset,... ...DigitalOcean is seeking a Senior Engineer 2 to play a key technical role in our AI Inference Optimization team. DigitalOcean aims to be... ...familiarity with the Gen AI (LLM, VLM, LMM) landscape, including...CloudFull timeLocal areaWorldwideFlexible hours- ...analysis and AI training and inference. Designed from the ground... ...data center, edge, and cloud.As a Forward Deployed Engineer (FDE), you are a core member... ...our most strategic, large-scale customer environments. You... ...'s technical team to optimize infrastructure throughput,...Cloud
$145.19k - $203.26k
...physical constraints into tractable optimization problems solved at constellation scale — and feeds directly into... ...integrated modelling framework alongside engineers modelling power, thermal,... ...whether from satellite, telecom, cloud infrastructure, or logistics domains...CloudPermanent employmentFull timeTemporary workLocal areaWorldwideRelocationShift work- ...DocuSign is seeking a Director of Engineering for Search to architect and lead the discovery engine powering... ...of Engineering, AI, you will mentor leaders, optimize indexing latency, and ensure scalable operations across cloud providers. This role requires strong...Cloud
$234.4k - $296.6k
...for hybrid, multi-cloud environments. Join... ...expertise with the scale and operational... ...and Cisco’s global engineering capabilities. Our... ...distributed training and inference pipelines to... ...model deployment, optimization, and evaluation.... ...in the code-gen / LLM community. Why...CloudFull timeTemporary workLocal areaFlexible hours$146k - $194k
...Platform is the internal engineering force multiplier behind... ...is how Anduril scales its business systems with... ...production stability.Support launch readiness, production... ...distributed systems, CI/CD, and cloud or platform... ...AI coding assistants, LLM-enabled applications, or...CloudFull timeWork experience placementImmediate start$155k - $205k
...are seeking a Systems Engineer to join our team in Redmond... ..., WA. We build and optimize the infrastructure... ...spanning edge devices to cloud GPU clusters — with a... ...; and ML training and inference optimization. By applying... ...designing and scaling distributed training infrastructure...Cloud$171k - $231.4k
...a broad set of global cloud-based services including... ...faster, lower IT costs, and scale. You will be a part of... ...AWS Product Compliance Engineering team within AWS Global... ...region planning and launches. You will develop compliance... ...Engineering, AWS Infra Service Supply Chain, Compliance...CloudLocal areaFlexible hours- Principal Quality Assurance Engineer - Generative AI & LLM Platforms This role has been designed as 'Hybrid... ...Enterprise is the global edge-to-cloud company advancing the way people live... ...distributed systems, traffic behavior, scaling, failover, disaster recovery, and adverse...CloudWork at office2 days per week
- ...hiring for a founding engineer at at the Staff and Senior... ...can move faster and scale AI with confidence. Why... ...dollar segment at a leading cloud provider, and served as... ...work from idea through launch. Excellent... ...infrastructure, including training, inference, or the platforms and...Cloud
$152.2k - $243.7k
...opportunity to create impact at scale — tackling meaningful... ...entrepreneurs and engineers in 2016, Pismo is a... ...in the market. Pismo’s cloud-based platform empowers firms to build and launch financial products rapidly... ...Implement and optimize high-performance data services...CloudFull timeWork experience placementWork at officeLocal area- ...Hardware / Machine Design Engineer Location: Hybrid... ...a next-generation cloud platform designed to power... ...—including large-scale compute, model training, fine-tuning, inference, and emerging agentic AI... ...architectures into highly optimized, production-ready systems...CloudWork at officeRelocation3 days per week
$182k - $242k
CoreWeave is The Essential Cloud for AI™. Built for... ...to build and scale AI with confidence. Trusted... ...internal and customer engineering teams, offering valuable... ...concept, onboarding, and optimizing workloads, you will... ...(AI/ML) training and inference workloads on technologies...CloudPermanent employmentFull timeTemporary workCasual workWork at officeFlexible hours$61k - $101k
...certification in software engineering concepts, along... ...knowledge of cloud delivery models such... ...architecture, training, and inference. We require... ...building and scaling a high-performance LLM inference platform... ...and GPU/CPU serving optimization for low-latency, high...CloudFull timeFor contractors- ...Google LLC in the Seattle region seeks software engineers to help build next‑generation technologies at massive scale. You will work on a project crucial to Google’s needs with opportunities to switch teams as the business grows and evolves. We value versatility, demonstrated...Cloud
$182k - $242k
...CoreWeave is The Essential Cloud for AI™. Built for... ...innovators to build and scale AI with confidence. Trusted... ...for a Senior Storage Engineer, File & Block to help... ...AI training and inference workloads and internal... ...monitor, analyze, and optimize using telemetry, metrics...CloudPermanent employmentFull timeTemporary workCasual workWork at officeFlexible hours$200.8k - $251k
...Francisco seeks a team member to build and optimize a machine learning framework for large... ...experience and solid software engineering skills, particularly in tools like CUDA... ...range of $200,800 - $251,000, along with comprehensive benefits. #J-18808-Ljbffr Scale AIFull time$209.1k - $282.9k
...seek a Staff Robotics Engineer to help create the next... ..., independent of cloud services. You will... ...reliability, and safety Optimize end-to-end latency, determinism... ...AI and perception inference pipelines deployed on... ...are built and scaled at Arm. In addition...CloudWork at officeLocal areaVisa sponsorshipRelocation package$83.2k - $104k
...the simplest scalable cloud. If you have a growth... ...projects around, on a scale unmatched by most. We... ...are passionate about optimizing and troubleshooting data... ...Reporting to the Manager of Infra::Machines::Design, you... ...support sustaining engineering efforts for the...CloudFull timeLocal areaWorldwideFlexible hours$182k - $242k
...CoreWeave is The Essential Cloud for AI™. Built for... ...to build and scale AI with confidence.... ...rendering, and real-time inference. Our stack is engineered for speed, scale,... ...kernel authoring and optimization. You will write, profile... ...critical path of LLM inference....CloudPermanent employmentFull timeTemporary workCasual workWork at officeFlexible hours$188k - $275k
...is The Essential Cloud for AI™. Built for... ...innovators to build and scale AI with confidence... ...What You'll Do: Inference Platform Team The... ..., and system-wide optimizations that drive... ...a Staff Software Engineer (IC5) on the Inference... ...Triton, TensorRT-LLM, Ray Serve, or TorchServe...CloudPermanent employmentFull timeTemporary workCasual workWork at officeFlexible hours$142.8k - $274.8k
...MicrosoftOverviewMicrosoft Silicon, Cloud Hardware, and Infrastructure Engineering (SCHIE) is the team... ..., quality, delivery, scale and sustainability... ...solutions that will manage and optimize the Cloud infrastructure... ..., AI training and inference applications, and synthetic...CloudOngoing contractWork at officeLocal areaWorldwide3 days per week$198k - $264k
...CoreWeave is The Essential Cloud for AI™. Built for... ...to build and scale AI with confidence.... ...with Product, Engineering, Research, Infrastructure... ...Manager focused on inference, you will lead... ...customer onboarding, launch readiness, and runtime optimization. The Inference team...CloudPermanent employmentFull timeTemporary workCasual workWork at officeFlexible hours$168.1k - $227.4k
...Sr. Software Development Engineer, Inference Team - AWS Neuron AWS Neuron... ...-built accelerators for cloud-scale machine learning. This senior... ...in inference framework optimization and engineering best practices... ...engineering to deliver best-in-class LLM serving performance. The...CloudWork experience placementInternshipLocal areaFlexible hours- ...comfortable talking with analytics engineers, data architects, and... ...partnering with our product team on launches. Not marketing content.... ...and the engagement model we scale behind you. Requirements... ...line yourself.Real fluency with cloud data warehouses (Snowflake, Databricks...Cloud
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Cloud Inference Launch Engineer: Scale & Optimize LLM Infra. Be the first to apply!
- cloud engineer Washington State
- cloud developer Washington State
- informatica cloud developer Washington State
- senior cloud network engineer Washington State
- cloud architect Washington State
- senior aws cloud engineer Washington State
- aws cloud Washington State
- senior cloud service delivery manager Washington State
- cloud security Washington State
- cloud Washington State





