AI Inference Core - Senior SW Engineer for Platform & DevOps
Cerebras Systems
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services. This order of magnitude increase in speed is transforming the user experience of AI applications, unlocking real-time iteration and increasing intelligence via additional agentic computation.Cerebras works with the leading model labs, global enterprises, and cutting-edge AI-native startups. OpenAI recently announced a multi-year partnership with Cerebras, to deploy 750 megawatts of scale, transforming key workloads with ultra high-speed inference.About the TeamThe Core Infrastructure team builds the software systems that power engineering workflows across Cerebras.Our infrastructure coordinates complex work across machines, clusters, development environments, and hardware systems. We build orchestration frameworks, execution engines, scheduling systems, test infrastructure, developer tools, and reusable software platforms that allow engineers to build, test, qualify, and deliver software reliably at scale.These systems are primarily built in Python, but the work goes far beyond scripting or automation. Our frameworks act as the control plane for distributed workflows, managing resources, execution state, concurrency, failures, retries, dependencies, and observability across large and complex environments.About the RoleWe are hiring a Software Engineer to build and operate the platform layer behind Cerebras engineering infrastructure.You will work on CI/CD systems, Kubernetes, deployment automation, cloud and on-premises infrastructure, developer environments, artifact management, and observability. You will help make the systems engineers depend on reliable, scalable, and easy to operate.This is an engineering-focused infrastructure role rather than a primarily ticket-driven operations position. You will automate repeated work, debug failures across system boundaries, and turn operational problems into durable software and platform improvements.We value strong systems fundamentals, independent problem solving, and sound engineering judgment more than familiarity with any particular infrastructure product.ResponsibilitiesDesign, build, and maintain CI/CD systems supporting build, test, integration, qualification, and release workflows.Build and operate Kubernetes-based platforms and services used by engineering teams across Cerebras.Develop deployment systems, internal tools, and self-service workflows that make infrastructure changes repeatable, reviewable, and safe.Improve infrastructure reliability, capacity, performance, cost efficiency, monitoring, and operational readiness.Debug issues spanning CI pipelines, Kubernetes workloads, networking, storage, authentication, operating systems, and distributed applications.Perform root-cause analysis and implement lasting fixes rather than relying on repeated manual intervention.Partner with software, IT, security, networking, release, and developer-productivity teams to deliver scalable infrastructure solutions.Skills & Qualifications5+ years of professional experience in platform engineering, DevOps, infrastructure engineering, site reliability engineering, or software engineering.Hands-on experience building or maintaining CI/CD pipelines and automated software-delivery workflows.Experience deploying and operating services using Kubernetes and containerized environments.Experience with a major cloud platform, preferably AWS, and programmatic infrastructure provisioning.Strong understanding of Linux or Unix operating-system fundamentals.Understanding of networking concepts such as DNS, routing, load balancing, proxies, ports, TLS, and service connectivity.Proficiency in Python, Shell, or another language used to build infrastructure automation and operational tooling.Experience with monitoring, logging, alerting, dashboards, and incident investigation.Strong debugging and problem-solving skills, including the ability to investigate issues spanning applications, infrastructure, networking, and operating systems.Preferred Skills & QualificationsExperience with infrastructure-as-code tools, specifically Terraform.Experience with Kubernetes controllers, operators, custom resources, Helm, Argo CD, or similar platform technologies.Experience managing artifact repositories, package registries, build caches, or software-distribution infrastructure.Familiarity with build systems, dependency management, and reproducible-build practices.Experience supporting hybrid environments spanning cloud infrastructure, on-premises systems, and specialized hardware.Experience with identity and access management, secrets, certificates, TLS, or mTLS.Experience building internal developer platforms or self-service infrastructure products.BS/MS in Computer Science or a related field, or equivalent practical experience.Why Join CerebrasPeople who are serious about software make their own hardware. At Cerebras, we have built a breakthrough architecture that is unlocking new opportunities for the AI industry. With dozens of model releases and rapid growth, we’ve reached an inflection point in our business. Members of our team tell us there are five main reasons they joined Cerebras:Build a breakthrough AI platform beyond the constraints of the GPU.Publish and open source their cutting-edge AI research.Work on one of the fastest AI supercomputers in the world.Enjoy job stability with startup vitality.Our simple, non-corporate work culture that respects individual beliefs.Find out more about what it's like to work at Cerebras here! Apply today and become part of the forefront of groundbreaking advancements in AI!Cerebras Systems is committed to creating an equal and diverse environment and is proud to be an equal opportunity employer. We celebrate different backgrounds, perspectives, and skills. We believe inclusive teams build better products and companies. We try every day to build a work environment that empowers people to do their best work through continuous learning, growth and support of those around them.This website or its third-party tools process personal data. For more details, click here to review our CCPA disclosure notice.LocationSunnyvale, CA; Toronto, CANEmployment TypeFull timeLocation TypeHybridDepartmentSoftware Engineering
$193.93k - $352.29k
...and profound opportunity for AI to drive positive change in the... ...building a universal autonomy platform: self-driving for all roads and... ...for building & improving the core infrastructure for autonomy teams... ....Maintain an in-house ML inference platform to serve large language...SeniorImmediate startFlexible hours$184k - $287.5k
...architect that wants to change how the AI inference industry works. With agent adoption taking... ...teach C-level executives, distinguished engineers and datacenter developers how to scale... ...with disaggregated inference.Evangelize devops best-practices for such as managing kubernetes...SeniorDevopsFull time$152k - $241.5k
...recently, GPU deep learning ignited modern AI — the next era of computing — with the... ...looking for an AI & Deep Learning Compiler Engineer. NVIDIA is hiring software engineers for... ...our DLC has been the backbone of NVIDIA’s inference engine, spanning across data centers,...SeniorFull timeRemote work- NVIDIA Corporation seeks a Senior Software Engineer to advance its JAX-based platform and performance optimizations in AI. You will design and implement JAX core components, collaborating with researchers to build scalable, fast tooling for data, training and analysis...Senior
- CrowdStrike, Inc. is seeking a Senior Engineer for the Core Libraries team. You will design, build, extend, and maintain shared Windows, Linux... ...design, contributing to a scalable endpoint security platform and an AI-enabled security stack. #J-18808-Ljbffr CrowdStrike...Senior
$152k - $241.5k
## Senior Software Engineer, Deep Learning Inference - TensorRTApplylocations: US, CA, Santa Claratime... ...be scaled to multiple platforms for functionality and performance... ..., GPU architects and DevOps engineers across diverse... ...vacancy.NVIDIA uses AI tools in its recruiting processes...SeniorDevops$165k - $200k
...robotics company developing AI-powered robots to... ...better. JOB SUMMARY At the core of every robot policy we ship is a simulation platform that has to be fast... ...driver. As Staff Simulation Engineer for Core Simulation... ...platform alongside senior team members, applying...SeniorLocal area$201.3k - $352.3k
...It all started when engineer Fred Luddy wrote... ...ServiceNow is the AI control tower for business... ...Our ServiceNow AI platform brings together any... ...About the Role — Senior Staff Data Platform... ...'ll work on top of core data architecture... ...Experience working in a DevOps environment...SeniorDevopsFull timeWork experience placementWork at officeImmediate startRemote workFlexible hours2 days per week$180k - $240k
...Senior AI Infrastructure EngineerSanta Clara,... ...AI Infrastructure Engineer to design, build,... ...high-performance AI platform powering our autonomous... ..., KubeFlow).Inference Performance Engineering... ...timeouts.Agentic DevOps & CI/CD: Develop... ...GitOps automation.Core Skills:...SeniorDevopsWork at office$165k - $242k
Join to apply for the Senior Platform Engineer II, Compute Services role at CoreWeave... ...is The Essential Cloud for AI™. Built for pioneers by... ..., AKS, GKS). Strong GitOps/DevOps with ArgoCD or similar helm... ...is represented through our core values: Be Curious at Your...SeniorDevopsPermanent employmentTemporary workCasual workWork at officeRemote workFlexible hours$186k - $282k
...this role exists FloQast's AI products have outgrown the infrastructure... ...patterns the rest of the platform runs on. Transform, AI... ...never catches. Today, DevOps engineers carry this work alongside... ...demand throughput, cross-region inference, quotas and throttles, and...SeniorDevops- ...Superintelligence Cloud, is a leader in AI cloud infrastructure serving tens... ...Tuesday.About the RoleLambda’s Core Cloud Platform powers compute provisioning and infrastructure... ...data centers. We are looking for a Senior Site Reliability Engineer to improve the reliability,...SeniorWork at officeLocal areaWork from homeFlexible hours
$262k - $364k
...coach a distributed engineering team, fostering... ...various NICs and platforms, tuning the communication... ..., storage, and AI/ML, staying ahead of training and inference advancements.... ...TheHPN team is at the core of Google's AI... ...Storage (GDS).As a Senior Staff Software Engineer...SeniorRemote workWorldwide$139k - $185k
...CoreWeave, the AI Hyperscaler™, acquired... ...end-to-end platform to develop,... ...building, and inference at scale, we’re... ...The Production Engineering team builds and... ...are seeking a Senior Production Engineer... ...at the core of CoreWeave's... ...is a hands-on DevOps/SRE-flavored role...SeniorDevopsPermanent employmentFull timeTemporary workCasual workWork at officeFlexible hours- NVIDIA seeks a Senior Product Manager for AI Platform Inference (Finance) in Santa Clara to lead tooling, SDKs, and libraries enabling GPU-based inference deployments. You will craft strategy, roadmaps, and market plans, collaborating with developers to optimize models...Senior
$119.8k - $234.7k
...builds the end-to-end AI stack and is core to Azure AI innovation... ...for a Principal Software Engineer - Responsible AI who is... ...build and drive the team’s DevOps culture. Drive and... ...Safety and governance platforms for AI models and agents Inference, routing, orchestration...SeniorDevopsOngoing contractWork at officeLocal area3 days per week$144k - $216k
# Senior Software Engineer, Core Platform## FloQastPublished 16 Jul 2026### Share this jobSan Jose, CA, USA144K - 216K USD AnnualFull Time## Role Highlights###... ...success in contributing to complex data pipelines and identity systems. #J-18808-Ljbffr Towards AI, Inc.Senior- HP, Inc. seeks a Senior Software Engineer to build software platforms across client apps, backend services, and cloud integrations... ...-grade software with edge AI as part of the PC Design team.... ..., hardware, security, validation, DevOps, and AI/ML teams to deliver robust...SeniorDevops
$153k - $204k
...CoreWeave, the AI Hyperscaler™, acquired... ...end-to-end platform to develop,... ...building, and inference at scale, we’re... ...keep our platform engineers building: we... ...guidance from senior teammates, you'... ...SRE, operations, DevOps, or technical-support... ...through our core values:...SeniorDevopsPermanent employmentFull timeTemporary workCasual workInternshipWork at officeFlexible hours- ...United StatesProducts - Engineering /Fulltime /HybridOver 5... ...is one of our core values and in our DNA.... ...Principal Engineer - Cloud Platform Reports To: Director of... ...Collaborate with product, DevOps, and security teams to... ...artificial intelligence (AI) tools to support parts...SeniorDevopsFull timeRemote workFlexible hours
$182.5k - $260.5k
...Taipei, and Tokyo. Our core values are openness... ...are available at Senior Staff and above.... ...Scientist, you own the inference and optimization layer that makes AI in agentic... ...larger models as the platform matures. Partner with... ...systems and backend engineers to ship capabilities...SeniorWork at office$168k - $258.75k
...Deep Learning Training and Inference Frameworks, with the goal... ...of supporting NVIDIA's top AI researchers and software engineers in driving the future of... ...The TPM will work alongside senior management and coordinate... ...development, release management, DevOps with a consistent track...SeniorDevopsFull timeRemote workShift work$184k - $287.5k
NVIDIA is the engine of modern Artificial Intelligence, Parallel Compute... ...Infrastructure, and Agentic AI - the biggest technology... ...storageExtensive experience with DevOps solutions, including Python, Ansible... ...& RAG-based workflows, inference at scale, large scale training...SeniorDevopsFull time$224k - $356.5k
...into the unlimited potential of AI to define the next era of... ...best experience. We own the platform — performance, CI/CD pipelines... ...innovations in leading open-source LLM inference frameworks — identify... ...Computer Science, Computer Engineering, Electrical Engineering, or equivalent...SeniorFull timeLocal area$152k - $241.5k
...workloads for Physical AI. These include... ...step model training, and inference, all on a large scale!... ...our product management, engineering, and business teams to... ...operating Kubernetes-based platforms for distributed GPU... ...Airflow, Argo, etc), modern DevOps practices (GitOps, IaC...SeniorDevopsFull time$152k - $241.5k
NVIDIA's high-performance computing platforms are powering the AI revolution across many applications and... ...linear algebra and Tensor Core primitives. Since 2017, it has provided... ...degree in Computer Science, Computer Engineering, or related field (or equivalent experience...SeniorFull time$184k - $287.5k
NVIDIA is looking for Senior Software Engineer to join the Cumulus Linux team! We present you with an... ...defined to meet the exploding growth in AI and high-performance computing. You'll... ...for defining and implementing core platform services, as well as Reliability, Availability...SeniorFull timeWork at office$193.93k - $291.15k
...driver, combining cutting-edge AI with automotive-grade hardware. Nuro licenses its core technology, the Nuro Driver, to... ...gives the automakers and mobility platforms a clear path to AVs at commercial... .... We are searching for an engineer with experience building reliable...Senior$196k - $310.5k
...Security organization is looking for a Senior Cybersecurity Engineer - Identity Platform & Access Management to lead the... ...safeguards developers, services, and AI agents throughout NVIDIA's... ...scale.At NVIDIA, we operate at the core of enterprise security, architecting...SeniorFull timeWorldwide$184k - $287.5k
We are looking for a Senior System Software Engineer, Software Defined... ...solutions for NVIDIA's AI Clouds hosting GPU-... ...multi-node training, inference, cloud gaming, and... ...tuningCollaborate with SRE, DevOps, and network... ...with observability platforms and tools (Prometheus...SeniorDevopsFull time
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to AI Inference Core - Senior SW Engineer for Platform & DevOps. Be the first to apply!
- senior platform engineer Sunnyvale, CA
- platform developer Sunnyvale, CA
- platform engineer Sunnyvale, CA
- platform engineering manager Sunnyvale, CA
- data platform engineer Sunnyvale, CA
- senior associate architect Sunnyvale, CA
- senior dynamics crm developer Sunnyvale, CA
- senior application security Sunnyvale, CA
- senior account director Sunnyvale, CA
- senior plumbing designer Sunnyvale, CA


