Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Cloud Inference Launch Engineer: Scale & Optimize LLM Infra

Jobleads-US

United States Digital Space LLC is seeking an engineer to join the Cloud Inference team and scale Claude across AWS, GCP, Azure, and future CSPs. You will own end-to-end inference on each cloud platform, from API integration to deployment and daily operations.

You will focus on fast, cost-efficient validation, performance improvements, and reliability to ensure consistent behavior across providers and accelerate model delivery.

#J-18808-Ljbffr Jobleads-US
Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Cloud Inference Launch Engineer: Scale & Optimize LLM Infra in Washington State vacancy
  • $320k

     ...of committed researchers, engineers, policy experts, and...  ...About the role The Cloud Inference team scales and optimizes Claude to serve the massive...  ...Inference, the model & inference launch team owns the validation...  ...a strong interest in LLM serving; prior inference... 
    Cloud
    Visa sponsorship

    United States Digital Space LLC

    Washington State
    1 day ago
  • $152.2k - $205.9k

     ...building the future of cloud financial...  ...in production at scale. Spend is driven...  ...tokens, model choice, inference patterns, and...  ...cost, usage, and optimization insights that power...  ...most consequential launches, run the customer...  ...directly with engineering, and dig into consumption... 
    Cloud
    Flexible hours

    Amazon

    Seattle, WA
    3 days ago
  • $150k - $190k

     ...reliable open-model inference platform. Backed...  ...of researchers and engineers on a mission to...  ...accessibility. Our cloud provides instant access...  ..., and model optimization. About The Role We...  ...technical assets that scale the team: demo...  ...knowledge of modern LLM inference — serving... 
    Cloud
    Immediate start

    Featherless AI

    Seattle, WA
    2 days ago
  • $152k

     ...continue our growth and launch new services at...  ...Detection Engineering (DE) team within Coupang...  ...threats at scale. The team builds and...  ...develops AI Agent and LLM-based analysis capabilities...  .... Analyze and optimize automation...  ...operating systems in cloud environments, preferably... 
    Cloud
    Temporary work
    Flexible hours

    Coupang

    Seattle, WA
    12 hours ago
  • $121.6k - $243.2k

     ...on building large-scale and highly available cloud infrastructure, which...  ...technologies to optimize our AI Infra stack, including training infra, inference infra, and AI agents...  ...Science, Computer Engineering, or a related discipline...  ...following areas: - LLM training infra,... 
    Cloud
    Temporary work
    Local area

    ByteDance

    Seattle, WA
    1 day ago
  • $191.2k - $239k

     ...to build the simplest scalable cloud. If you have a growth mindset,...  ...DigitalOcean is seeking a Senior Engineer 2 to play a key technical role in our AI Inference Optimization team. DigitalOcean aims to be...  ...familiarity with the Gen AI (LLM, VLM, LMM) landscape, including... 
    Cloud
    Full time
    Local area
    Worldwide
    Flexible hours

    DigitalOcean

    Seattle, WA
    3 days ago
  •  ...analysis and AI training and inference. Designed from the ground...  ...data center, edge, and cloud.As a Forward Deployed Engineer (FDE), you are a core member...  ...our most strategic, large-scale customer environments. You...  ...'s technical team to optimize infrastructure throughput,... 
    Cloud

    VAST Data

    Seattle, WA
    2 days ago
  • $145.19k - $203.26k

     ...physical constraints into tractable optimization problems solved at constellation scale — and feeds directly into...  ...integrated modelling framework alongside engineers modelling power, thermal,...  ...whether from satellite, telecom, cloud infrastructure, or logistics domains... 
    Cloud
    Permanent employment
    Full time
    Temporary work
    Local area
    Worldwide
    Relocation
    Shift work

    BLUE ORIGIN

    Seattle, WA
    2 days ago
  •  ...DocuSign is seeking a Director of Engineering for Search to architect and lead the discovery engine powering...  ...of Engineering, AI, you will mentor leaders, optimize indexing latency, and ensure scalable operations across cloud providers. This role requires strong... 
    Cloud

    Jobleads-US

    Seattle, WA
    19 hours ago
  • $234.4k - $296.6k

     ...for hybrid, multi-cloud environments. Join...  ...expertise with the scale and operational...  ...and Cisco’s global engineering capabilities. Our...  ...distributed training and inference pipelines to...  ...model deployment, optimization, and evaluation....  ...in the code-gen / LLM community. Why... 
    Cloud
    Full time
    Temporary work
    Local area
    Flexible hours

    Cisco Systems, Inc

    Seattle, WA
    1 day ago
  • $146k - $194k

     ...Platform is the internal engineering force multiplier behind...  ...is how Anduril scales its business systems with...  ...production stability.Support launch readiness, production...  ...distributed systems, CI/CD, and cloud or platform...  ...AI coding assistants, LLM-enabled applications, or... 
    Cloud
    Full time
    Work experience placement
    Immediate start

    Anduril Industries

    Seattle, WA
    2 days ago
  • $155k - $205k

     ...are seeking a Systems Engineer to join our team in Redmond...  ..., WA. We build and optimize the infrastructure...  ...spanning edge devices to cloud GPU clusters — with a...  ...; and ML training and inference optimization. By applying...  ...designing and scaling distributed training infrastructure... 
    Cloud

    General Robotics

    Redmond, WA
    3 days ago
  • $171k - $231.4k

     ...a broad set of global cloud-based services including...  ...faster, lower IT costs, and scale. You will be a part of...  ...AWS Product Compliance Engineering team within AWS Global...  ...region planning and launches. You will develop compliance...  ...Engineering, AWS Infra Service Supply Chain, Compliance... 
    Cloud
    Local area
    Flexible hours

    AmazonWebServices

    Seattle, WA
    12 hours ago
  • Principal Quality Assurance Engineer - Generative AI & LLM Platforms This role has been designed as 'Hybrid...  ...Enterprise is the global edge-to-cloud company advancing the way people live...  ...distributed systems, traffic behavior, scaling, failover, disaster recovery, and adverse... 
    Cloud
    Work at office
    2 days per week

    Hewlett Packard Enterprise Company

    Friday Harbor, WA
    1 day ago
  •  ...hiring for a founding engineer at at the Staff and Senior...  ...can move faster and scale AI with confidence. Why...  ...dollar segment at a leading cloud provider, and served as...  ...work from idea through launch. Excellent...  ...infrastructure, including training, inference, or the platforms and... 
    Cloud

    Neuromorphic Labs

    Bellevue, WA
    2 days ago
  • $152.2k - $243.7k

     ...opportunity to create impact at scale — tackling meaningful...  ...entrepreneurs and engineers in 2016, Pismo is a...  ...in the market. Pismo’s cloud-based platform empowers firms to build and launch financial products rapidly...  ...Implement and optimize high-performance data services... 
    Cloud
    Full time
    Work experience placement
    Work at office
    Local area

    VISA

    Bellevue, WA
    a month ago
  •  ...Hardware / Machine Design Engineer Location: Hybrid...  ...a next-generation cloud platform designed to power...  ...—including large-scale compute, model training, fine-tuning, inference, and emerging agentic AI...  ...architectures into highly optimized, production-ready systems... 
    Cloud
    Work at office
    Relocation
    3 days per week

    Designworks Talent LLC

    Bellevue, WA
    3 days ago
  • $182k - $242k

    CoreWeave is The Essential Cloud for AI™. Built for...  ...to build and scale AI with confidence. Trusted...  ...internal and customer engineering teams, offering valuable...  ...concept, onboarding, and optimizing workloads, you will...  ...(AI/ML) training and inference workloads on technologies... 
    Cloud
    Permanent employment
    Full time
    Temporary work
    Casual work
    Work at office
    Flexible hours

    CoreWeave

    Bellevue, WA
    2 days ago
  • $61k - $101k

     ...certification in software engineering concepts, along...  ...knowledge of cloud delivery models such...  ...architecture, training, and inference. We require...  ...building and scaling a high-performance LLM inference platform...  ...and GPU/CPU serving optimization for low-latency, high... 
    Cloud
    Full time
    For contractors

    J.P. Morgan

    Seattle, WA
    10 days ago
  •  ...Google LLC in the Seattle region seeks software engineers to help build next‑generation technologies at massive scale. You will work on a project crucial to Google’s needs with opportunities to switch teams as the business grows and evolves. We value versatility, demonstrated... 
    Cloud

    Jobleads-US

    Seattle, WA
    1 day ago
  • $182k - $242k

     ...CoreWeave is The Essential Cloud for AI™. Built for...  ...innovators to build and scale AI with confidence. Trusted...  ...for a Senior Storage Engineer, File & Block to help...  ...AI training and inference workloads and internal...  ...monitor, analyze, and optimize using telemetry, metrics... 
    Cloud
    Permanent employment
    Full time
    Temporary work
    Casual work
    Work at office
    Flexible hours

    CoreWeave

    Bellevue, WA
    12 days ago
  • $200.8k - $251k

     ...Francisco seeks a team member to build and optimize a machine learning framework for large...  ...experience and solid software engineering skills, particularly in tools like CUDA...  ...range of $200,800 - $251,000, along with comprehensive benefits. #J-18808-Ljbffr Scale AI
    Full time

    Scale AI

    Seattle, WA
    1 day ago
  • $209.1k - $282.9k

     ...seek a Staff Robotics Engineer to help create the next...  ..., independent of cloud services. You will...  ...reliability, and safety Optimize end-to-end latency, determinism...  ...AI and perception inference pipelines deployed on...  ...are built and scaled at Arm. In addition... 
    Cloud
    Work at office
    Local area
    Visa sponsorship
    Relocation package

    ARM

    Seattle, WA
    19 hours ago
  • $83.2k - $104k

     ...the simplest scalable cloud. If you have a growth...  ...projects around, on a scale unmatched by most. We...  ...are passionate about optimizing and troubleshooting data...  ...Reporting to the Manager of Infra::Machines::Design, you...  ...support sustaining engineering efforts for the... 
    Cloud
    Full time
    Local area
    Worldwide
    Flexible hours

    DigitalOcean

    Seattle, WA
    1 day ago
  • $182k - $242k

     ...CoreWeave is The Essential Cloud for AI™. Built for...  ...to build and scale AI with confidence....  ...rendering, and real-time inference. Our stack is engineered for speed, scale,...  ...kernel authoring and optimization. You will write, profile...  ...critical path of LLM inference.... 
    Cloud
    Permanent employment
    Full time
    Temporary work
    Casual work
    Work at office
    Flexible hours

    CoreWeave

    Bellevue, WA
    a month ago
  • $188k - $275k

     ...is The Essential Cloud for AI™. Built for...  ...innovators to build and scale AI with confidence...  ...What You'll Do: Inference Platform Team The...  ..., and system-wide optimizations that drive...  ...a Staff Software Engineer (IC5) on the Inference...  ...Triton, TensorRT-LLM, Ray Serve, or TorchServe... 
    Cloud
    Permanent employment
    Full time
    Temporary work
    Casual work
    Work at office
    Flexible hours

    CoreWeave

    Bellevue, WA
    1 day ago
  • $142.8k - $274.8k

     ...MicrosoftOverviewMicrosoft Silicon, Cloud Hardware, and Infrastructure Engineering (SCHIE) is the team...  ..., quality, delivery, scale and sustainability...  ...solutions that will manage and optimize the Cloud infrastructure...  ..., AI training and inference applications, and synthetic... 
    Cloud
    Ongoing contract
    Work at office
    Local area
    Worldwide
    3 days per week

    Microsoft

    Redmond, WA
    4 days ago
  • $198k - $264k

     ...CoreWeave is The Essential Cloud for AI™. Built for...  ...to build and scale AI with confidence....  ...with Product, Engineering, Research, Infrastructure...  ...Manager focused on inference, you will lead...  ...customer onboarding, launch readiness, and runtime optimization. The Inference team... 
    Cloud
    Permanent employment
    Full time
    Temporary work
    Casual work
    Work at office
    Flexible hours

    CoreWeave

    Bellevue, WA
    a month ago
  • $168.1k - $227.4k

     ...Sr. Software Development Engineer, Inference Team - AWS Neuron AWS Neuron...  ...-built accelerators for cloud-scale machine learning. This senior...  ...in inference framework optimization and engineering best practices...  ...engineering to deliver best-in-class LLM serving performance. The... 
    Cloud
    Work experience placement
    Internship
    Local area
    Flexible hours

    Jobleads-US

    Seattle, WA
    8 hours ago
  •  ...comfortable talking with analytics engineers, data architects, and...  ...partnering with our product team on launches. Not marketing content....  ...and the engagement model we scale behind you. Requirements...  ...line yourself.Real fluency with cloud data warehouses (Snowflake, Databricks... 
    Cloud

    Golden Analytics

    Bellevue, WA
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Cloud Inference Launch Engineer: Scale & Optimize LLM Infra. Be the first to apply!