System Engineer Datacenter GPU
Ampcus
Job-ID30284781Reference25-23064Ampcus Inc. is a certified global provider of a broad range of Technology and Business consulting services. We are in search of a highly motivated candidate to join our talented Team. Job Title: System Engineer Datacenter GPU Location(s): Austin, TX Client is looking for a System Engineer Datacenter GPU to work in IPP (Infrastructure, Planning and Process) Sanity Engineering. IPP is a core software infrastructure organization within Client. This group works with various other groups within Client Software such as Graphics Processors, Mobile Processors, Deep Learning, Artificial Intelligence and Driverless Cars to cater to their infrastructure needs. These cloud services provide almost half a million automated jobs per day on thousands of servers helping with the productivity of thousands of Client software engineers worldwide. The cloud hosts a heterogeneous mix of machines and devices with various operating systems (Windows/Linux/Android), a multitude of hardware platforms both Client GPUs and Tegra Processors. Are you passionate about distributed infrastructure and looking for a sophisticated workspace and are ready to build the next generation of cloud services for chip bringups, design creative solutions, mine through data to uncover real problems and fix them? We are looking forward to onboard a fun-loving person like you.What you'll be doingDevelop framework and scripts to automate workflows and deployments in the private cloud environment.Have a thorough understanding of Client GPU hardware and driver stack, SBIOS, VBIOS and should be able to enhance automation for farm wide updates.Solve complex problems involving multi site distributed infrastructure scaling, leading GPU product bringups (PCIe & Enterprise) in infrastructure, integrating GPU test suites to infrastructure harness etc.Deploy and maintain a large farm of machines using world-class Configuration Management & Infrastructure Automation (IaC) tools like Chef, Ansible, TerraformDevelop extensive monitoring dashboards systems to have fast, reliable and real-time pulse of the various infrastructure subsystems using state of the art telemetry.Will be leading a service charter completely, and will be essentially a responsible pilot in charge for the development, monitoring and automation of that infrastructure.Automation and performance tuning of regression test frameworks, creation of self healing/automated recovery solutions for multi-geo regression farms.Assist in roll-out and deployment of new development features aimed at supporting the latest Client hardware and technologies.What we need to seeBachelor's or Master's Degree in Computer Science or Software Engineering, or equivalent demonstrable experience.10+ years of relevant experience . Ability to analyze and debug source code to triage, root cause and resolve issues in the infrastructure. Collaborate with the development teams in improving the build and test infrastructure.Familiar with maintenance and setup of Linux, Windows hosts and popular open source applications such as Nginx, Apache Apache Tomcat and MySQL server.Hands-on programming experience with any including but not limited to Python (preferred), JAVA etc.Unix & TCL shell proficiency is expected.Experience in MySQL/No-SQL(plus), should be able to write complex queries.Experience with version control systems like Perforce, GIT.Demonstrable experience working in large scale enterprise production systems. 5+ years of operational experience required.Ways to stand out from the crowdBackground with automating bare metal and VM provisioning.Prior knowledge of VM isolation for GPUs and Client Confidential Computing is a plus.Experience with public clouds (AWS, GCP, Azure), VM and container virtualization technologies like VMware, KVM, Docker and Kubernetes Clusters.Experience with debugging GPU performance issues, embedded device software development and automation, software driver development and CUDA/TensorRT applications.All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, age, protected veterans or individuals with disabilities.
- ...data centers, to PCs, gaming and embedded systems. Grounded in a culture of innovation... ...ROLE:AMD is seeking a Lead Systems Debug Engineer to provide technical leadership for... ...validation, and debug of next-generation datacenter GPU platforms. This role will drive platform...Suggested
- ...Operations organization and key business units, engineering leadership, business operations, sales,... ...communication skills to debug systems-level issues from firmware level to applications... ...all aspects of data management for AMD datacenter GPU products. This is a high visibility...Suggested
- ...United States, Norway, Bhutan, and Ethiopia. To learn more, visit Position Overview We are seeking a Senior GPU Systems & Fabric Engineer to serve as the critical bridge between our physical GPU/network infrastructure and the Kubernetes abstraction layer....SuggestedRemote jobFull timeLocal area
$173.9k - $235.2k
...— this is the role.We are seeking a Systems Development Engineer to build automation software, diagnostic... ...You will work across PCIe topology, GPU diagnostics, Linux drivers, and... ...readiness of the platform3. Partner with datacenter operations to close the loop between...SuggestedPermanent employmentInternshipLocal areaWorldwideFlexible hoursNight shiftDay shift- ...NVIDIA Corporation in Austin, TX is seeking a Senior Software Engineer to build accelerated PyTorch-based solutions for large-scale ML models... ...TFMs, and ensemble models, advancing training and inference on GPU infrastructure. You will collaborate with researchers,...Suggested
- ...forward.THE ROLE: We are looking for a dynamic, energetic Sr. Systems Design Engineer with a passion for platform/system firmware. As part of... ...the development and implementation of firmware for AMD's Datacenter GPU Products.THE PERSON:As a firmware lead, you are...
$173.9k - $235.2k
We are seeking an experienced Senior Systems Development Engineer to lead the development of automation... ...level issues across storage, compute, GPU, networking in production... ...diagnosability, and reliability- Partner with datacenter operations teams to close the loop between...InternshipLocal areaWorldwideFlexible hours- ...supports advanced power management products for demanding compute, AI, datacenter, and infrastructure applications.Job DescriptionRenesas is seeking an experienced Staff Digital Mixed-Signal Systems Engineer to develop, model, and verify digitally implemented mixed-signal...
- ...that moves the world forward.THE ROLE: The Datacenter Platform Engineering Group (DPEG) organization is looking for an experienced system level debug engineer. Individual will be... ...issues is a must, understand the flow of a GPU through the different layers of a system and...
- ...centers, to PCs, gaming and embedded systems. Grounded in a culture of innovation and... ..., software, validation, and customer engineering teams.Influence future platform management... ...background in server, AI accelerator, GPU, or datacenter platforms.Deep expertise in BMC...
- ...advance your career.THE ROLE:Join AMD's Datacenter Infrastructure team in Rockdale, Texas,... ...that enable AMD's most advanced engineering initiatives.In this role, you will collaborate... ...network engineers to ensure critical systems remain reliable, performant, and ready...Visa sponsorshipAfternoon shift
- ...matters to the world. The Production Engineering Team Examples of key problems the team... ...technical level. You debug distributed systems methodically across layers you don't own.... ..., and follow up until they do. Bonus: GPU training workloads. InfiniBand or RoCE. Slurm...
- ...Renesas Electronics seeks a Staff Digital Mixed-Signal Systems Engineer to develop, model, and verify digitally implemented mixed-signal... ...next‑generation digital power management products used in AI, datacenter, and HPC applications. You will translate product...
- ...advanced power management products for demanding compute, AI, datacenter, and infrastructure applications. Job Description... ...Renesas is seeking an experienced Staff Digital Mixed-Signal Systems Engineer to develop, model, and verify digitally implemented mixed-signal...Full time
$152k - $241.5k
Sr Software Engineer - Distributed Systems Engineer, EDA InfrastructureNVIDIA is hiring engineers to build and scale the infrastructure that supports... ...and platform services that manage large fleets of GPU-based and CPU-based compute systems used by engineering teams...Full timeRemote work$173k - $245k
Meta is seeking a Hardware Systems Engineer to support the new product introduction (NPI) of next-generation AI and high-performance computing... ...for AI and HPC hardware platforms, including AI accelerators, GPU clusters, and high-bandwidth memory subsystems in data center...Hourly payLocal areaRemote work- ...data centers, to PCs, gaming and embedded systems. Grounded in a culture of innovation and... ...opportunity for a memory systems design engineer within the client and graphics platform engineering... ...of computer hardware architecture (GPU, CPU/APU, memory and bus logic...
- ...days per week (T-W-Th) Who We AreWe are an engineering-focused IT organization responsible for... ...innovative.The RoleWe are seeking a Senior Systems Engineer to lead the implementation,... ...environments (e.g., job schedulers, containers, GPU resources, monitoring and logging tools)....Full timeH1bLocal areaWork from homeRelocation package3 days per week
- ...and technically curious Field Applications Engineer to join our Global Partner Support team.... ...intersection of data center hardware, system software, and AI workloads.In this role,... ...Participate in system-level debugging of GPU, driver, networking, and AI software issues...Internship
- ...data centers, to PCs, gaming and embedded systems. Grounded in a culture of innovation and... ...performant HPC and AI applications for use on AMD GPU products. THE PERSON:You are passionate... ...developers, customers, and product engineers to deliver programs on-time. You delight...
- ...applications.To support our extraordinary teams who build great products and contribute to our growth, we’re looking to add a Principal Engineer, Systems Architecture Engineering located in Austin, TX.Reporting to the Chief Technology Officer,the Principal Engineer, Systems...Full timeVisa sponsorshipFlexible hours
- ...domains in the interest of national security. L3Harris Engineering Hiring Event – Rockwall, Texas (11/10) Must have a... ...in Plano, Rockwall, Greenville, and Waco, Texas: Systems Engineers: FPGA/GPU RF Signal Processing SIGINT COMINT ELINT...Local area
- ...Senior Systems EngineerSaronic Technologies is a leader in revolutionizing autonomy at sea, dedicated to developing state-of-the-art... ...OverviewWe are seeking a highly skilled and motivated Senior Systems Engineer to lead the design, deployment, and ongoing maintenance of our...Permanent employmentTemporary workWork at office
- ...we’ll advance your career.THE ROLE: The Data Center Platform Engineering Group (DPEG) organization is looking for skilled individuals that... ...can contribute to the bring-up, support and debug of complex system / SOC problems. Individuals will be part of a growing lab team...Local areaRemote work
- ...centralized Search, Q&A, and Conversational AI system that integrates seamlessly with all... ...About This RoleAs a Senior ML System Engineer on the AI & ML Platform’s Inference team,... ...scaling) to deep low-level optimizations (GPU kernels, quantization, speculative decoding...Work at officeLocal area
- ...advance your career.THE ROLE: As the Power Feature Enablement Systems Engineer you will be expected to be a Cross Functional Technical Leader... ...at hardware, firmware, and software levels.Understands CPU, GPU, Memory and NPU power/performance, SW power management, and system...
- ...System Engineer LOCATION: Houston/Austin (Texas) CATEGORY: Exempt, Full-Time, 100% On-site. JOB FUNCTION: Manage cloud operations in all data centers. Travel may be required. RESPONSIBILITIES: Design, deploy, and maintain cluster servers across all data...Full time
- ...technology that delivers 100× the performance of traditional digital systems at the same power and cost. This unlocks bigger, more capable... ...models and CNNs to advanced signal processing, and is engineered to operate from –40 °ree;C to +125 °ree;C, making it ideal...
$99.1k - $160k
We're seeking a Systems Development Engineer to join the Unified Workcell Compute team. This is a hands-on role where you will help design and build... ...robotics hardware platforms including x86, ARM, NVIDIA GPU systems, and embedded devices. - Create tooling and frameworks...Full timeTemporary workSeasonal workWorldwideFlexible hoursDay shift$110.5k - $160k
...cloud-scale innovation with world-class expertise across silicon engineering, hardware design, verification, software, and operations to... ..., PCB, high-speed components (HBM, PCIe, chip-to-chip), inter-system connections, and system-to-system interfaces. You'll dive deep...Flexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to System Engineer Datacenter GPU. Be the first to apply!
- senior linux systems engineer Austin, TX
- system engineer contract Austin, TX
- data systems engineer Austin, TX
- senior staff systems engineer Austin, TX
- microsoft systems engineer Austin, TX
- operations support system engineer Austin, TX
- advanced systems engineer Austin, TX
- software system engineer Austin, TX
- space systems engineer Austin, TX
- system test engineer Austin, TX




