Senior SDET, Inference Platform
Cerebras Systems
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services. This order of magnitude increase in speed is transforming the user experience of AI applications, unlocking real-time iteration and increasing intelligence via additional agentic computation.Cerebras works with the leading model labs, global enterprises, and cutting-edge AI-native startups. OpenAI recently announced a multi-year partnership with Cerebras, to deploy 750 megawatts of scale, transforming key workloads with ultra high-speed inference.About the RoleWe are looking for an Inference Platform SDET to join the Inference Service Quality team at Cerebras and work on the inference platform. This team sits at the intersection of distributed systems, cloud and cluster infrastructure, and the software stack that serves the world's fastest AI inference.In this role, you will own the quality and reliability of the infrastructure that deploys and runs the Cerebras Inference Platform — from CI/CD pipelines and Kubernetes-based deployments to ingress, load balancing, and service discovery. You will validate the platform both in cloud environments and on real Cerebras clusters, working side by side with the Inference Platform development team to catch issues before our customers do.This is an excellent opportunity for engineers who enjoy infrastructure, automation, and debugging across the full deployment stack, and who want to ensure that a platform serving inference at massive scale stays fast, reliable, and production-ready.ResponsibilitiesDesign, build, and maintain test infrastructure and automation for deploying and validating the Cerebras Inference Platform.Validate the platform across environments — from cloud-managed Kubernetes to deployments running on Cerebras hardware.Test and verify deployment infrastructure including Kubernetes workloads, CI/CD pipelines, ingress and service discovery, NGINX, and load balancing.Collaborate closely with the Inference Platform development team to ensure new features and platform capabilities ship reliably.Investigate and debug complex issues spanning networking, orchestration, deployment, and distributed services.Develop and maintain testbeds used to validate platform performance, scalability, and reliability.Identify failure points, bottlenecks, and edge cases that impact platform stability and inference performance.Contribute to test plans and validation strategies for new platform features and releases.Improve observability, diagnostics, and debugging workflows across the inference platform stack.Partner with engineering teams to ensure high-quality, production-ready releases of the Cerebras Inference Platform.Minimum Skills & Qualifications3+ years of experience in software engineering, QA/quality engineering, systems engineering, or infrastructure development.Strong programming skills in Python and/or Go (experience with both is a plus).Experience building automation tools, testing frameworks, or internal developer tooling.Hands-on experience with CI/CD systems (e.g., Jenkins)Experience debugging complex systems, distributed services, or networked infrastructure.Familiarity with systems-level development, infrastructure tooling, or platform integration.Strong problem-solving skills and the ability to investigate issues across multiple system and infrastructure layers.Excellent communication and collaboration skills.Experience mentoring junior engineersPreferred SkillsHands-on experience with Kubernetes and container orchestration in a real production or staging environment.Experience with cloud-managed Kubernetes such as Amazon EKS.Experience with GitOps/deployment tooling (e.g., ArgoCD).Familiarity with ingress controllers, service discovery, NGINX, and load balancing.Experience with build systems such as Bazel.Experience with cluster tooling and operations (e.g., k9s).Exposure to performance debugging, profiling, or system observability tools.Experience with ML inference infrastructure, model serving systems, or GPU-accelerated workloadsLocation: Toronto / SunnyvaleTeam: Inference Service QualityWhy Join CerebrasPeople who are serious about software make their own hardware. At Cerebras, we have built a breakthrough architecture that is unlocking new opportunities for the AI industry. With dozens of model releases and rapid growth, we’ve reached an inflection point in our business. Members of our team tell us there are five main reasons they joined Cerebras:Build a breakthrough AI platform beyond the constraints of the GPU.Publish and open source their cutting-edge AI research.Work on one of the fastest AI supercomputers in the world.Enjoy job stability with startup vitality.Our simple, non-corporate work culture that respects individual beliefs.Find out more about what it's like to work at Cerebras here! Apply today and become part of the forefront of groundbreaking advancements in AI!Cerebras Systems is committed to creating an equal and diverse environment and is proud to be an equal opportunity employer. We celebrate different backgrounds, perspectives, and skills. We believe inclusive teams build better products and companies. We try every day to build a work environment that empowers people to do their best work through continuous learning, growth and support of those around them.This website or its third-party tools process personal data. For more details, click here to review our CCPA disclosure notice.LocationSunnyvale, CA; Sunnyvale, CA; Toronto, CANEmployment TypeFull timeLocation TypeHybridDepartmentSoftware Engineering
- ...to deliver industry-leading training and inference speeds; over 10 times faster than GPU-... ...resiliency are some of the key focus areas. As senior software development engineer in Test,... ...joined Cerebras:Build a breakthrough AI platform beyond the constraints of the GPU....PlatformSenior
- ...industry-leading training and inference speeds; over 10 times faster... ...Cloud. We work closely with platform, infrastructure, ML systems,... ...business.About The RoleAs a Senior Software Engineer in Test for... ...standards.Mentor and guide junior SDETs on testing methodology,...PlatformWork at officeRemote workShift work3 days per week
- ...to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based... ....About the RoleAs a Staff GPU Inference SDET, you will be the founding quality,... ...joined Cerebras:Build a breakthrough AI platform beyond the constraints of the GPU.Publish...Platform
$168k - $270.25k
...We are looking for Senior Software Development Engineer in Test to join our Confidential Computing team for NVIDIA's Enterprise SWQA... ...for Compute software releases on all new compute architecture platforms including Tesla GPUs, NVIDIA turnkey systems and OEM systems.Develop...PlatformSeniorFull timeWorldwide$168k - $270.25k
We are seeking a Senior Software Engineer in Test to join the Compute CUDA Quality Assurance team and drive validation readiness for... ...execution readiness, defect triage, and automation from early platform bring up through key release stages. We collaborate across CUDA...PlatformSeniorFull time$168k - $258.75k
Inference is the fastest growing and most competitive area in Generative AI today. It is where AI models impact our daily... ...algorithms, usecases, and deployment techniques. As a Senior Product Manager for AI Platform Inference you will be responsible for building the tools...PlatformSeniorFull time$124.5k - $272k
...the Fortune 100, trust the Netskope One platform, its Zero Trust Engine, and the powerful... ...and Instagram.Positions are available at Senior Staff and above. Candidates are assessed... ...Machine Learning Scientist, you own the inference and optimization layer that makes AI in...PlatformSenior$208k - $327.75k
...AI models and applications running on NVIDIA hardware. Every inference deployment — from a single-GPU workstation to a multi-thousand... ...decide what we build, what we adopt, and what we retire.Build platforms, not one-offs. Deliver capabilities that generalize across model...PlatformSeniorFull time$184k - $287.5k
...ISVs). These ISVs are developing a network stack for distributed inference which will be used to orchestrate wide area networks/cloud... ...above.Collaborate with developers and onboard them to AI-Grid platforms and SDKs by providing deep technical guidance.Anticipate Telecom...PlatformSeniorFull timeWork experience placement$148k - $235.75k
We are looking for a Senior Technical Product Marketing Manager. This role will be located... ...data center business and pivotal in our inference marketing. You will be focused on... ...Be Doing:Help drive NVIDIA’s inference platform technical go-to-market effortsWork closely...PlatformSeniorFull time$224k - $356.5k
We are now looking for a Senior System Software Engineer to work on Dynamo. NVIDIA is hiring software engineers for its GPU-accelerated... ...processing. We are a fast-paced team building Generative AI inference platform to make design and deployment of new AI models easier and...PlatformSeniorFull time- ...advance your career. THE ROLE:This role owns the framework-layer inference product strategy for ROCm, translating customer, ecosystem,... ...Provide software-informed input to future GPU architecture and platform decisions relevant to inference performance at scale.Drive...PlatformSeniorRemote workShift work
- ...architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud... ..., test infrastructure, developer tools, and reusable software platforms that allow engineers to build, test, qualify, and deliver...PlatformSenior
$184k - $287.5k
...skilled and motivated software engineers to join us and build AI inference systems that serve large-scale models with extreme efficiency.... ...such as containerd/CRI-O/CRIU.Experience with cloud platforms (AWS/GCP/Azure), infrastructure as code, CI/CD, and production...PlatformSeniorFull time$152k - $241.5k
...systems, image classification, speech recognition, etc. With the rapid advancement of AI, our DLC has been the backbone of NVIDIA’s inference engine, spanning across data centers, personal devices, automotive, and robotics. The compiler must deliver leading inference...PlatformSeniorFull timeRemote work$140k - $224.25k
...product documentation and feature requirements into test plans, automation coverage, and targeted hands-on validation for complex PC platform scenarios, including corner cases and obscure interactions across hardware, software, drivers, games, and system settings.Debug...PlatformSeniorFull time$108.8k - $163.2k
...provide an all-in-one point-of-sale (POS) and business management platform that handles everything from seamless payment processing to... ...confidence.What does a successful Software Development Engineer in Test (SDET) do at Clover?As Clover continues to grow globally, the EMEA...PlatformSeniorFull timeWorldwide$140k - $215k
...mission hasn’t changed — we’re here to stop breaches, and we’ve redefined modern security with the world’s most advanced AI-native platform. We work on large scale distributed systems, processing almost 3 trillion events per day and this traffic is growing daily. Our...PlatformSeniorFull timeContract workWork experience placementWork at officeLocal area- ...to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based... ...the RoleWe are looking for a hands-on SDET Technical Lead to establish and lead Release... ...joined Cerebras:Build a breakthrough AI platform beyond the constraints of the GPU.Publish...PlatformWork at officeRemote work3 days per week
- ...to deliver industry-leading training and inference speeds; over 10 times faster than GPU-... ...contribute to projects on our Inference Platform team. Our team primarily owns the orchestration... ....Raise the effectiveness of senior engineers through design feedback, pairing...PlatformSenior
- ...architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud... ...closely with hardware architects, runtime engineers, cloud platform teams, AI researchers, and strategic customers to rapidly...PlatformSenior
$193.93k - $352.29k
...world around us, that's why we're building a universal autonomy platform: self-driving for all roads and all rides. Founded in 2016,... ...generation to on-road validation. Maintain an in-house ML inference platform to serve large language models efficiently....PlatformSeniorImmediate startFlexible hours$193.93k - $352.29k
...world around us, that’s why we’re building a universal autonomy platform: self-driving for all roads and all rides.Founded in 2016,... ...data generation to on-road validation.Maintain an in-house ML inference platform to serve large language models efficiently.Maintain an...PlatformSeniorImmediate startFlexible hours$184k - $287.5k
NVIDIA seeks a Senior Software Engineer specializing in Deep Learning Inference for our growing team. As a key contributor, you will help design, build, and optimize... ...You will play a central role in improving these platforms, facilitating smooth deployment and serving of...PlatformSeniorFull time- ...to deliver industry-leading training and inference speeds; over 10 times faster than GPU-... ...inference.About The RoleWe’re hiring a Senior Frontend Engineer to own and scale critical... ...team and take ownership of major platform areas such as billing, request logs, and...PlatformSeniorShift work
$152k - $204k
...pioneers by pioneers, CoreWeave delivers a platform of technology, tools, and teams that... ...025. Learn more at What You'll Do: Senior engineers are area owners who lead designs... ...teams to evolve our Kubernetes-native inference platform and meet strict P99 SLAs at scale...PlatformSeniorPermanent employmentFull timeTemporary workCasual workWork at officeFlexible hoursShift work$193.3k - $261.5k
...PyTorch and JAX enabling unparalleled ML inference and training performance.The Inference... ...celebrates knowledge-sharing and mentorship. Our senior members enjoy one-on-one mentoring and... ...with vLLM, SGLang, TensorRT or similar platforms in production environments.- Deep...PlatformSeniorWork experience placementInternshipLocal areaFlexible hours- Senior Product Manager About the Role We're looking for a Senior Product Manager to lead... ...and execution for our skills and AI platform. In this role you'll own products at the... ...paths, content recommendations, skills inference, and generative/agentic features - balancing...PlatformSenior
$140k - $224.25k
...of technological innovation, pushing the boundaries of AI and accelerated computing. Our team in Santa Clara, CA is looking for a Senior Software Engineer in Test to join us in this exciting journey. This is a ground breaking opportunity to work with powerful technology...SeniorFull time- ...DDN is seeking a Senior Software Engineering Manager to lead the engineering organization responsible for our KV Cache Platform—a distributed memory and storage platform that accelerates large-scale LLM inference across GPU clusters.In this role, you will lead geographically...PlatformSenior
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior SDET, Inference Platform. Be the first to apply!
- sdet Sunnyvale, CA
- application tester Sunnyvale, CA
- software tester Sunnyvale, CA
- software development engineer in test sdet Sunnyvale, CA
- senior director product management Sunnyvale, CA
- senior grant accountant Sunnyvale, CA
- senior tax Sunnyvale, CA
- senior data management analyst Sunnyvale, CA
- senior consulting engineer Sunnyvale, CA
- sr electrical engineer Sunnyvale, CA

