Staff Software Engineer, AI Inference Cloud
$209.1k - $282.9kARM
As an engineer on Arm’s AI Inference Cloud team, you will shape the technical direction and develop highly available, scalable services for running AI inference workloads. You will guide architecture and actively contribute to development across Kubernetes orchestration, workload management, service delivery, and observability. Partnering with AI compute, Inference Runtime, and product teams to enhance the performance and usability of Arm’s AI platform.Responsibilities:Define and build the architecture for cloud-based AI inference services.Develop Kubernetes controllers and platform capabilities supporting workload deployment, scheduling, recovery, scaling, upgrades, and lifecycle management.Establish production practices for health validation, progressive rollout, rollback, observability, and service objectives. Improve platform reliability, scalability, performance, and resource efficiency.Lead production readiness reviews and resolve complex issues across services, Kubernetes, networking, and compute infrastructure. Turn incidents and operational bottlenecks into durable platform improvements.Lead design and build reviews, mentor engineers, and drive technical alignment across teams.Necessary Skills and Experience:5+ years of experience, or equivalent proven impact, building distributed systems, cloud platforms, or production infrastructure.Deep production experience with Kubernetes, including controllers, operators, scheduling, resource management, networking, and workload lifecycle management.Strong software and production engineering skills, including hands-on programming in Go, C++, Rust, Python, or a similar language, and experience with reliable services, APIs, concurrency, observability, deployment safety, capacity planning, and incident response.A track record of leading complex technical initiatives while remaining hands-on, including architecture, implementation, debugging, mentoring, and influencing technical direction across teams.Ability to troubleshoot complex systems and communicate clearly with engineers from different technical backgrounds.Preferred Skills and Experience:Experience with AI infrastructure, model serving, or accelerator-backed workloads.Familiarity with frameworks such as PyTorch, Ray, vLLM, SGLang, or TensorRT-LLM, or experience qualifying accelerators and tuning distributed workloads.Knowledge of inference performance and resource-efficiency considerations.In Return:You will be part of our AI Platforms team - A driven and diverse group passionate about developing foundational production capabilities for AI inference at Arm. We provide a collaborative setting where your ideas can come to life quickly. Your work will directly impact the success of our AI projects, shape Arm’s AI inference capabilities, defining and operating production inference workloads. This is an outstanding opportunity to work with world-class teams and contribute to groundbreaking advances in AI technology. Join us in building the next generation of AI inference infrastructure!(Recruiter to Complete)Additional Information:Please note that a relocation package (including visa sponsorship support) is available for this role, for candidates who require it.Salary Range:$209,100-$282,900 per yearWe value people as individuals and our dedication is to reward people competitively and equitably for the work they do and the skills and experience they bring to Arm. Salary is only one component of Arm's offering. The total reward package will be shared with candidates during the recruitment and selection process.Accommodations at ArmAt Arm, we want to build extraordinary teams. If you need an adjustment or an accommodation during the recruitment process, please email View email address on click.appcast.io. To note, by sending us the requested information, you consent to its use by Arm to arrange for appropriate accommodations. All accommodation or adjustment requests will be treated with confidentiality, and information concerning these requests will only be disclosed as necessary to provide the accommodation. Although this is not an exhaustive list, examples of support include breaks between interviews, having documents read aloud, or office accessibility. Please email us about anything we can do to accommodate you during the recruitment process.Hybrid Working at ArmArm’s approach to hybrid working is designed to create a working environment that supports both high performance and personal wellbeing. We believe in bringing people together face to face to enable us to work at pace, whilst recognizing the value of flexibility. Within that framework, we empower groups/teams to determine their own hybrid working patterns, depending on the work and the team’s needs. Details of what this means for each role will be shared upon application. In some cases, the flexibility we can offer is limited by local legal, regulatory, tax, or other considerations, and where this is the case, we will collaborate with you to find the best solution. Please talk to us to find out more about what this could look like for you.Equal Opportunities at ArmArm is an equal opportunity employer, committed to providing an environment of mutual respect where equal opportunities are available to all applicants and colleagues. We are a diverse organization of dedicated and innovative individuals, and don’t discriminate on the basis of race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or status as a protected veteran.
$209.1k - $282.9k
As a Software Engineer on our AI Inference Runtime team, you will set technical direction for critical components of distributed Inference runtime for... ...throughput, reliability, and resource efficiency.Partner with cloud, framework, compiler, hardware, and research teams; lead...CloudWork at officeLocal area$236k - $330k
...usher in this new era, we seek AI-native thinkers across every... ...the state of the art in LLM inference systems and optimization.Our... ...tuning. We embrace AI-native engineering, using AI not only as the workload... ...approaches to accelerate software development, experimentation,...SuggestedShift work$209.1k - $282.9k
As an engineer on the AI Compute Infra team, you will design, build, and operate... ...-tuning, evaluation, and inference. You will work across... ...issues across applications, cloud infrastructure, clusters, and... ...developing reliable infrastructure software.Practical knowledge of...CloudWork at officeLocal areaVisa sponsorshipRelocation package$201k - $315k
...Seattle, WASoftware - Software Systems /Full-time /... ...looking for an experienced Staff Software Engineer to build, scale, and... ...to training our AI models in Perception,... ...workloadsExperience with cloud infrastructure on AWS... ...workloads (training, inference, data generation)...CloudFull timeTemporary workRelocation package$209.1k - $282.9k
We seek a Software Engineer to contribute to the next generation of Physical AI platforms on Arm. Your work will focus on building... ...on the device, independent of cloud services.You will focus on building... ...systems.Exposure to AI/ML inference systems, even from a platform...CloudWork at officeLocal area- ...Description:DataRobot delivers AI that maximizes impact... ...’s Fleet team is the engine behind how our... ...where you come in.As a Staff Software Engineer, you’ll be responsible... ...for multi-cloud and hybrid environments... ...infrastructure for training and inference.Why Join the Fleet...CloudFull timeLocal areaRemote workWorldwideFlexible hours
$229k - $343k
...services.We’re looking for a Staff Software Engineer to join Snap Inc on our... ...reliably for large-scale batch inference and low-latency online inference... ...responsible use of emerging AI technologies to improve... ...large-scale microservices, cloud infrastructure and/or platform...CloudFull timeLive inWork at officeLocal area- ...services that enable ML engineers and data... ...monitoring, and agentic AI capabilities. We... ...driven products. As a Staff Engineer, you'll... ...low-latency model inference, large-scale... ...and operating great software systems.Who you areWe... ....Familiarity with cloud services (e.g.,...CloudFlexible hours
- ...technology products.As a Senior Lead Software Engineer at JPMorgan Chase within the... ...Architect and deploy secure, scalable cloud platforms optimized for AI/ML workloads.Partner with AI teams... ...architecture, ML training, and inference.Experience with Infrastructure as...CloudFor contractors
$168.1k - $227.4k
AWS Neuron is the complete software stack for the AWS Inferentia and Trainium cloud-scale machinelearning accelerators. This role is for a senior software engineer in the Machine Learning Inference Applications team. This role is responsible for development and performance...CloudInternshipFlexible hours$168.1k - $227.4k
...Neuron is the complete software stack for AWS... ...accelerators for cloud-scale machine learning... ...senior software engineering role is part of... ...Machine Learning Inference Applications team... ...Neuron, TPUs, or other AI accelerator... ...supervisors, and staff; adhere to standards...CloudWork experience placementInternshipLocal areaFlexible hours$188k - $275k
...Staff Software Engineer, InferenceCoreWeave is The Essential Cloud for AI™. Built for pioneers by pioneers, CoreWeave delivers a platform of technology, tools, and teams... ...traded company (Nasdaq: CRWV) in March 2025.Inference Platform Team The Inference team builds and...CloudPermanent employmentFull timeCasual workWork at office- ...Lambda, The Superintelligence Cloud, is a leader in AI cloud infrastructure... ...future. We are seeking a Staff Engineer to help our development of... ...generation of AI training and inference at scale. As a Staff... ...0+ years of experience in software engineering, platform engineering...CloudWork at officeLocal areaImmediate startWork from homeFlexible hours
- ...accelerators and multiple cloud providers. Build... ...high-performance inference infrastructure for machine... ...model architectures and AI accelerator platforms.... ...Significant software engineering experience, particularly... ...work policy requiring staff to work from an Anthropic...CloudFull timeWork at officeWorldwideVisa sponsorshipFlexible hours
$209.1k - $282.9k
As a software Engineer on the AI Compute Platform team, you will design and build a secure, reliable,... ...developer tools for distributed training and inference, working closely with compute... ...experience building distributed systems, cloud platforms, or production backend...CloudWork at officeLocal areaVisa sponsorshipRelocation package- ...Rippling.com addresses.About the roleData Cloud is Rippling’s unified data platform... ...product workflows.We’re hiring a Staff Software Engineer to build the AI platforms within Data Cloud. At... ...primitives for model tuning, inference, evaluation, and development of task...CloudWork at office3 days per week
$92k - $135k
...Description CoreWeave is The Essential Cloud for AI™. Built for pioneers by pioneers,... ...more at What You'll Do: Join the Inference team to ship production features that improve... ...quickly with mentorship from experienced engineers. About the role: Implement well-...CloudPermanent employmentFull timeTemporary workCasual workInternshipWork at officeFlexible hours- ...usher in this new era, we seek AI-native thinkers across every... ...core part of Snowflake’s Data Engineering strategy. The Dynamic Tables... ...increasingly complex query shapes. As a Staff Engineer on this team, you... ...operating systems at cloud scale (multi-tenant SaaS, petabyte...CloudFull time
$236k - $339.2k
...usher in this new era, we seek AI-native thinkers across every... ...and passionate StaffSoftware Engineer for our Snowpark Container... ...high-scale, high-performance, cloud native compute platform to enable... ...’s Data Cloud Mission.AS A STAFF SOFTWARE ENGINEER, YOU WILL:Design and...CloudWork at office- ...individual can thrive.Job DescriptionThe AI Inference Engineer plays a critical role in the AI... ...inference routing and orchestration. Ensure software solutions are optimized for peak... ...technologies, including Docker, Kubernetes, and cloud platforms such as AWS, GCP, and Azure....CloudFull timeLocal areaImmediate start
- ...usher in this new era, we seek AI-native thinkers across every... .... There is only one Data Cloud. Snowflake’s founders started... ...But it didn’t stop there. They engineered Snowflake to power the Data Cloud... ...engineers. AS A STAFF SOFTWARE ENGINEER - IDENTITY & ACCESS...CloudFull time
$405k
...interpretable, and steerable AI systems. We want AI to be safe... ...of committed researchers, engineers, policy experts, and business... ...seeking talented and experienced Staff+ Software Engineers to join our... ...Kubernetes clusters across multiple cloud providers, handling...CloudFull timeWork at officeVisa sponsorshipFlexible hours$189k - $303k
...We are searching for an exceptional Staff-level Backend Software Engineer to join the Aurora Services Engineering... ...teams within Aurora. Embrace AI tools to add new features which delight... ...backend services running in Aurora’s AWS cloud used to monitor and manage the...CloudFull timeRemote work$220k - $250k
...you are Metropolis is seeking a Staff Software Engineer to lead our Visit Experience team. At this... ...Experience with leveraging AI technology to streamline engineering activities... ...Datastores: MySQL, PostgreSQL, Snowflake Cloud: AWS Version control: Git & GitHub...CloudFull timeTemporary workWork experience placementWork at officeLocal area$210k - $250k
...developer-first workflows so engineering and data teams can ship with... ...is growing and hiring a new Staff Software Engineer - Front End, who will... ...Help evaluate and integrate AI-assisted development tools and... ...full stack Familiarity with cloud platforms, CI/CD workflows,...CloudFull timeLocal area3 days per week$129.5k - $214.02k
...your work matters—and so do you. Staff Software Engineer-ENG: We are seeking a highly experienced... ...projects. • Proficiency with cloud technologies like Azure, AWS, GCP, and... ...of workforce insights, and people-first AI, our ability to reveal unseen ways to build...CloudFull timeWorldwide- ...Databricks, we are obsessed with Data + AI to solve the world's toughest... ...strategies and cutting-edge engineering. Our team ensures timely,... ...across different pricing plans and cloud providers (AWS, GCP & Azure).As a Sr. Staff Software Engineer on the Money team, you...CloudWork at officeWorldwide
$140.6k - $173.1k
...opportunities.We are looking for a passionate Senior Software Development Engineer who enjoys solving complex problems, passionate... ..., and enhancing agentic solutions in a cloud-native environment using industry-leading AI tools/platforms, Java, AWS, Kafka event streaming...CloudFull timeFlexible hours$143k - $286k
...WalmartBusiness Segment: Home OfficeRole summary:As a Staff Software Engineer at Walmart, you will lead the design and delivery of... ...full software development lifecycle with a focus on cloud-native architecture and AI/ML integration. You will provide technical leadership...CloudFull timeTemporary workPart time$200.4k - $260.5k
...will enable game developers to bring their ideas to life in an entirely new way.As a Staff Software Engineer on our gaming AI team, you will be responsible for leading the delivery of cloud services and features that power this platform. This is new and challenging work...CloudFull timeWork at officeWorldwide
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Staff Software Engineer, AI Inference Cloud. Be the first to apply!
- senior aws cloud engineer Seattle, WA
- aws cloud architect Seattle, WA
- remote cloud architect Seattle, WA
- informatica cloud developer Seattle, WA
- senior principal cloud computing engineer Seattle, WA
- cloud network engineer Seattle, WA
- google cloud architect Seattle, WA
- cloud engineer Seattle, WA
- senior cloud solutions architect Seattle, WA
- entry level cloud engineer Seattle, WA



