Principal Software Engineer, AI Inference Cloud
$262.7k - $355.4kARM
As a Principal Engineer on Arm’s AI Inference Cloud team, you will shape the technical direction and develop highly available, scalable services for running AI inference workloads. You will guide architecture and actively contribute to development across Kubernetes orchestration, workload management, service delivery, and observability. Partnering with AI compute, Inference Runtime, and product teams to enhance the performance and usability of Arm’s AI platform.Responsibilities:Define and build the architecture for cloud-based AI inference services.Develop Kubernetes controllers and platform capabilities supporting workload deployment, scheduling, recovery, scaling, upgrades, and lifecycle management.Establish production practices for health validation, progressive rollout, rollback, observability, and service objectives. Improve platform reliability, scalability, performance, and resource efficiency.Lead production readiness reviews and resolve complex issues across services, Kubernetes, networking, and compute infrastructure. Turn incidents and operational bottlenecks into durable platform improvements.Lead design and build reviews, mentor engineers, and drive technical alignment across teams.Necessary Skills and Experience:8+ years of experience, or equivalent proven impact, building distributed systems, cloud platforms, or production infrastructure.Deep production experience with Kubernetes, including controllers, operators, scheduling, resource management, networking, and workload lifecycle management.Strong software and production engineering skills, including hands-on programming in Go, C++, Rust, Python, or a similar language, and experience with reliable services, APIs, concurrency, observability, deployment safety, capacity planning, and incident response.A track record of leading complex technical initiatives while remaining hands-on, including architecture, implementation, debugging, mentoring, and influencing technical direction across teams.Ability to troubleshoot complex systems and communicate clearly with engineers from different technical backgrounds.Preferred Skills and Experience:Experience with AI infrastructure, model serving, or accelerator-backed workloads.Familiarity with frameworks such as PyTorch, Ray, vLLM, SGLang, or TensorRT-LLM, or experience qualifying accelerators and tuning distributed workloads.Knowledge of inference performance and resource-efficiency considerations.In Return:You will be part of our AI Platforms team - A driven and diverse group passionate about developing foundational production capabilities for AI inference at Arm. We provide a collaborative setting where your ideas can come to life quickly. Your work will directly impact the success of our AI projects, shape Arm’s AI inference capabilities, defining and operating production inference workloads. This is an outstanding opportunity to work with world-class teams and contribute to groundbreaking advances in AI technology. Join us in building the next generation of AI inference infrastructure!(Recruiter to Complete)Additional Information:Please note that a relocation package (including visa sponsorship support) is available for this role, for candidates who require it.Salary Range:$262,700-$355,400 per yearWe value people as individuals and our dedication is to reward people competitively and equitably for the work they do and the skills and experience they bring to Arm. Salary is only one component of Arm's offering. The total reward package will be shared with candidates during the recruitment and selection process.Accommodations at ArmAt Arm, we want to build extraordinary teams. If you need an adjustment or an accommodation during the recruitment process, please email View email address on click.appcast.io. To note, by sending us the requested information, you consent to its use by Arm to arrange for appropriate accommodations. All accommodation or adjustment requests will be treated with confidentiality, and information concerning these requests will only be disclosed as necessary to provide the accommodation. Although this is not an exhaustive list, examples of support include breaks between interviews, having documents read aloud, or office accessibility. Please email us about anything we can do to accommodate you during the recruitment process.Hybrid Working at ArmArm’s approach to hybrid working is designed to create a working environment that supports both high performance and personal wellbeing. We believe in bringing people together face to face to enable us to work at pace, whilst recognizing the value of flexibility. Within that framework, we empower groups/teams to determine their own hybrid working patterns, depending on the work and the team’s needs. Details of what this means for each role will be shared upon application. In some cases, the flexibility we can offer is limited by local legal, regulatory, tax, or other considerations, and where this is the case, we will collaborate with you to find the best solution. Please talk to us to find out more about what this could look like for you.Equal Opportunities at ArmArm is an equal opportunity employer, committed to providing an environment of mutual respect where equal opportunities are available to all applicants and colleagues. We are a diverse organization of dedicated and innovative individuals, and don’t discriminate on the basis of race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or status as a protected veteran.
$262.7k - $355.4k
As a Principal Software Engineer on our AI Inference Runtime team, you will set technical direction for critical components of distributed Inference runtime... ...throughput, reliability, and resource efficiency.Partner with cloud, framework, compiler, hardware, and research teams;...CloudWork at officeLocal area$262.7k - $355.4k
As a Principal Engineer on the AI Compute Infra team, you will design, build, and operate... ...-tuning, evaluation, and inference. You will guide work across... ...across applications, cloud infrastructure, clusters, and... ...developing reliable infrastructure software.Practical knowledge of...CloudWork at officeLocal areaVisa sponsorshipRelocation package- ...come to the right place.As a Principal Software Engineer at JPMorganChase within the... ...that intersect with AI/ML. Thus, you are collaborative... ...patterns to optimize training and inference of ML models on various... ...discipline).Practical cloud native experienceExperience...Cloud
- ...technology products.As a Senior Lead Software Engineer at JPMorgan Chase within the... ...Architect and deploy secure, scalable cloud platforms optimized for AI/ML workloads.Partner with AI teams... ...architecture, ML training, and inference.Experience with Infrastructure as...CloudFor contractors
- ...Lambda, The Superintelligence Cloud, is a leader in AI cloud infrastructure serving tens of thousands... ...be joining a team of highly capable software, hardware and network engineers building one of the largest AI training and inference networks in the world. You'll...CloudWork at officeLocal areaWork from homeFlexible hours
- ...through innovation and AI-powered automation... ...solutions with engineering excellence, and... ...our Spectrum-NET software solution performs... ...Spectrum-NET is a cloud-native, horizontally... ...We are seeking a Principal Software Engineer... ...systems, including ML inference Experience with...CloudFull timeRelocationFlexible hoursShift work
$168.1k - $227.4k
...Neuron is the complete software stack for AWS Inferentia... ...-built accelerators for cloud-scale machine learning. This senior software engineering role is part of the Machine Learning Inference Applications team and focuses... ...Neuron, TPUs, or other AI accelerator hardware-...CloudWork experience placementInternshipLocal areaFlexible hours- Senior Lead Software Engineer Be an integral part of an agile team that's constantly pushing... ...and deploy secure, scalable cloud platforms optimized for AI/ML workloads. Partner with AI teams... ...architecture, ML training, and inference. Experience with Infrastructure as...CloudFor contractors
$135.2k - $306.4k
The Oracle Cloud Infrastructure (OCI) team offers the... ....Oracle Kubernetes Engine (OKE) is OCI's managed... ...demanding cloud native, AI, and GPU workloads.We are... ...looking for a senior IC5 software engineer with deep... ...compute, model training or inference platforms, GPU scheduling...CloudTemporary workRemote workFlexible hours$262.7k - $355.4k
As a Principal Software Engineer on the AI Compute Platform team, you will design and build a secure, reliable... ...tools for distributed training and inference, working closely with compute infrastructure... ...building distributed systems, cloud platforms, or production backend...CloudWork at officeLocal areaVisa sponsorshipRelocation package- ...in critical industries through AI transformation. We specialize... ...two systems. The first is our inference control plane — open-weight models... ...operated across managed GPU clouds and customer-managed... ...you go home — building the fork engine, the guest agent, and the multi...CloudFull timeRemote workWork visaFlexible hoursDay shift
- ...come to the right place. As a Principal Engineer at JPMorgan Chase on the Core AI Infrastructure Platform team... ...unifies our training and inference pipelines across hybrid-cloud and Neo-cloud environments.... ...training or certification on software engineering concepts and 10+...Cloud
$188k - $275k
CoreWeave is The Essential Cloud for AI™. Built for pioneers by pioneers, CoreWeave delivers... ...025. Learn more at What You'll Do: Inference Platform Team The Inference team... ...systems. About the role: As a Staff Software Engineer (IC5) on the Inference team, you will...CloudPermanent employmentFull timeTemporary workCasual workWork at officeFlexible hours- ...individual can thrive.Job DescriptionThe AI Inference Engineer plays a critical role in the AI... ...inference routing and orchestration. Ensure software solutions are optimized for peak... ...technologies, including Docker, Kubernetes, and cloud platforms such as AWS, GCP, and Azure....CloudFull timeLocal areaImmediate start
- ...enterprise. To usher in this new era, we seek AI-native thinkers across every function... ..., and data marketplace. AS A PRINCIPAL SOFTWARE ENGINEER AT SNOWFLAKE YOU WILL: Solve real... ...? Build an industry-leading Cloud Data and AI Platform. Solve challenging...CloudFull time
$188.7k - $258.39k
...committed to building breakthrough software with a spark of magic. We believe... ...the Role We are looking for a Principal Software Engineer to join our Applied AI team. This is a rare opportunity... ...Experience developing and operating cloud services at enterprise scale (AWS,...CloudFull timeImmediate startFlexible hours$61k - $101k
...training or certification in software engineering concepts, along with 5+... ...require strong knowledge of cloud delivery models such as IaaS... ...architecture, training, and inference. We require experience with... ...use of enterprise-approved AI-assisted software development...CloudFull timeFor contractors$304k
...enterprise. To usher in this new era, we seek AI-native thinkers across every function who... ...Snowflake’s AI, Analytics and Data Engineering capabilities. We lead innovations across... ...helping customers build peta-byte scale multi-cloud data lakes on Snowflake. We deliver core...Cloud$135.2k - $306.4k
Oracle Cloud Infrastructure (OCI) delivers mission-critical applications... .... We are hoping to enhance engineering efficiency by concentrating... ...life-saving care. And with AI embedded across our products... ...Level - IC5As a Senior Principal Engineer, you will lead the design...CloudTemporary workWorldwideFlexible hours$114.6k - $234.6k
Oracle Cloud Infrastructure (OCI) is seeking a highly motivated Software Developer 4 to join the Infrastructure Planning and Capacity... ...support critical business and engineering processes that influence... ...to life-saving care. And with AI embedded across our products and...CloudTemporary workWorldwideFlexible hours$190.5k - $300k
...opportunity may be the right fit for you.Principal Software Engineer, C08 (IC)Our OpportunityChewy is... ...will advance the platform using modern cloud-first technologies while enabling more... ..., optimization, machine learning, and AI-assisted decision-making.You will be both...CloudLocal areaFlexible hours$143k - $286k
...WalmartBusiness Segment: Home OfficeRole summary:The Principal Software Engineer leads platform engineering efforts by delivering scalable, cloud-native solutions and reusable... ...software development lifecycle, integrating AI/ML technologies to build intelligent systems...CloudFull timeTemporary workPart timeWorldwide- ...dependable experiences.The RoleAs Principal Engineer for High Value Senders, you... ...New Products and Lead AI AdoptionCreate new HVS propositions... ...in customer experience and software delivery. Turn effective approaches... ...design expertise, including cloud infrastructure, safe...CloudFull timeWork at officeWorldwideFlexible hours
- ...one of the world’s most influential companies.As a Principal Software Engineer at JPMorganChase within the CDAO AI/ML Data Platforms Team, you provide deep... ...various technical disciplinesExtensive practical cloud native experienceExpertise in Computer Science, Computer...Cloud
- ...building the tools that define how software gets built and delivered. As AI agents redefine software... ...environments. We’re looking for a Principal Backend Engineer who thrives at the intersection... ...and evolving large-scale, cloud-native systems, with deep knowledge...CloudFull timeTemporary workRemote workHome officeShift work
$180k - $210k
...and incubated at the Allen Institute for AI. We're valued at $300M and grew revenue... ...-year. Our customers include Google Cloud, Snowflake, Databricks, RingCentral,... ...yoodli.ai . About this team As a Principal Backend Engineer on the Admin Workflows team, you'll own...CloudFull timeWork at office3 days per week$163.9k - $235.55k
...for an exceptionally skilled and visionary Principal Software Engineer with deep expertise in building, deploying, and scaling generative AI applications. As a key individual... ...TensorFlow, PyTorch). Strong understanding of cloud platforms (AWS, Google Cloud, Azure) and...CloudFull timeLocal area$114.6k - $234.6k
As a Principal Software Development Engineer in the Oracle Cloud Infrastructure (OCI) Security Platform division, you will play a critical leadership role in the... ...industry innovations to life-saving care. And with AI embedded across our products and services, we help...CloudTemporary workFlexible hours$220k - $250k
...California, United StatesProducts - Engineering /Fulltime /HybridOver 50,000... ...trust our end-to-end, cloud-driven networking solutions.... ...detailsTitle of position: Principal Software EngineerPosition type: Full... ...intelligent networking, generative AI, and autonomous agentic...CloudFull timeH1bLocal areaWork from homeWork visaShift work$197.3k - $313.7k
...SalesforceSalesforce is the #1 AI CRM, where humans with agents... ...person and virtually. As the Principal Engineer focused on architecture... ...requires a deep understanding of software development, architecture principles... ...), GCP, or other public cloud substrates. Examples of AWS...CloudFull time
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Principal Software Engineer, AI Inference Cloud. Be the first to apply!
- principal software engineer Seattle, WA
- senior principal software engineer Seattle, WA
- big data cloud engineer Seattle, WA
- salesforce marketing cloud developer Seattle, WA
- cloud architect Seattle, WA
- google cloud architect Seattle, WA
- principal cloud engineer Seattle, WA
- software engineer - cloud services Seattle, WA
- informatica cloud developer Seattle, WA
- senior aws cloud engineer Seattle, WA



