Infrastructure Engineer (Data Center Operations)
Cerebras Systems
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services. This order of magnitude increase in speed is transforming the user experience of AI applications, unlocking real-time iteration and increasing intelligence via additional agentic computation.Cerebras works with the leading model labs, global enterprises, and cutting-edge AI-native startups. OpenAI recently announced a multi-year partnership with Cerebras, to deploy 750 megawatts of scale, transforming key workloads with ultra high-speed inference.About the Role: We are looking for a hands-on Infrastructure Engineer to join our team and support our high-performance, on-premise server and networking infrastructure. You will be responsible for maintaining, provisioning, and troubleshooting hardware and Linux systems, working closely with network and system teams. This is an in-person role, ideal for someone who enjoys working across hardware, networking, and system layers.Key Responsibilities:Physically install, rack, cable, and maintain blade servers and hardware components (CPUs, DIMMs, NICs, storage devices, etc.)Connect servers to high-speed networks (100G/400G), verify optics/DACs, and check link statusConfigure BIOS, firmware, and out-of-band management (IPMI/iDRAC/iLO)Install and provision Linux OS; configure hostnames, IPs, routing, and NFS mount pointsDebug network issues at physical and OS level (VLAN, link issues, routing, etc.)Use Linux tools (e.g., ip, dmesg, netstat, ping) to isolate and fix issuesFollow provisioning playbooks and maintain accurate records of assets and changesUse scripting (Bash, Python) to automate routine tasks and improve efficiencyCollaborate with internal teams (network, systems, storage) and coordinate vendor RMAsDocument procedures and contribute to team knowledge baseTroubleshoot and replace failed server components with minimal downtimeQualifications:3–5+ years of experience in data center, lab, or infrastructure engineering rolesProficient in Linux system administration and network configurationStrong hands-on knowledge of x86 server hardware and enterprise networkingFamiliar with BIOS configuration, firmware updates, and remote management toolsSkilled in physical setup and troubleshooting of high-speed NICs and optical linksExperience with VLANs, static routing, and diagnosing layer 1–3 issuesAbility to write scripts for automation and diagnostics (Bash, Python preferred)Comfortable working on-site daily and lifting/moving server hardwarePreferred Skills:Experience with PXE, NFS, RAID controllers, and monitoring toolsFamiliarity with configuration management tools (e.g., Ansible)Prior experience in a lab or R&D hardware/software environmentThis is a unique opportunity to work with cutting-edge infrastructure and grow into more senior technical roles. If you enjoy bridging hardware and software with hands-on work, we’d love to hear from you.ADD DESCRIPTION HEREWhy Join CerebrasPeople who are serious about software make their own hardware. At Cerebras, we have built a breakthrough architecture that is unlocking new opportunities for the AI industry. With dozens of model releases and rapid growth, we’ve reached an inflection point in our business. Members of our team tell us there are five main reasons they joined Cerebras:Build a breakthrough AI platform beyond the constraints of the GPU.Publish and open source their cutting-edge AI research.Work on one of the fastest AI supercomputers in the world.Enjoy job stability with startup vitality.Our simple, non-corporate work culture that respects individual beliefs.Find out more about what it's like to work at Cerebras here! Apply today and become part of the forefront of groundbreaking advancements in AI!Cerebras Systems is committed to creating an equal and diverse environment and is proud to be an equal opportunity employer. We celebrate different backgrounds, perspectives, and skills. We believe inclusive teams build better products and companies. We try every day to build a work environment that empowers people to do their best work through continuous learning, growth and support of those around them.This website or its third-party tools process personal data. For more details, click here to review our CCPA disclosure notice.LocationHeadquarters/Sunnyvale OfficeEmployment TypeFull timeLocation TypeOn-siteDepartmentSoftware
- ...automation of tasks.Networking:• Setting up and maintaining network infrastructure.• Troubleshooting network connectivity problems.• Configuring... ...efficient inventory management processes and tools.Lab Operations:• Maintain a clean, organized, and safe lab environment.•...OperationsWork at officeRemote work
$165.2k - $223.6k
AWS Infrastructure Services owns the design, planning, delivery, and operation of all AWS global infrastructure. In other words, we... ...running. We support all AWS data centers and all of the servers, storage... ..., hardware, and network engineers, supply chain specialists, security...OperationsInternshipLocal areaFlexible hours$207k - $300k
...software systems, programming the data plane and control plane for... ...world's largest network infrastructure.Manage individual project... ...:Master’s degree or PhD in Engineering, Computer Science, or a... ...Google Global Networking, Data Center operations, systems research, and much...OperationsWorldwide$136.88k - $205k
...essential building blocks of the data infrastructure that connects our world.... ...of Ethernet and Data center products from New Production... ...part of a dynamic product engineering team working on most advanced... ...product characterization, test operation, qualification, yield ramp,...OperationsPermanent employmentFull timeInternshipWork from home- ...experienced Senior Customer Quality Engineer (CQE) to lead customer quality engagements for strategic Data Center, Networking and AI Infrastructure, customers. This role will support AMD... ...engineering, product line quality, operations, design, reliability, and...Operations
$109k - $160k
...Software Engineer - Data Infrastructure ServicesSunnyvale, CA / Bellevue, WA CoreWeave is The Essential... ...innovative solutions, automation and operations of our data platform infrastructure.... ...each day in our office and data center locationsA casual work environmentA...OperationsPermanent employmentFull timeCasual workWork at office- ...generation computing experiences—from AI and data centers, to PCs, gaming and embedded systems.... .... THE ROLEWe are seeking a Senior Data Engineer to design, build, and optimize our next... ...platform for high-tech semiconductor operations. You will lead the implementation of a...Operations
$115k - $130k
..., and networking solutions for Data Center, Cloud Computing, Enterprise IT... ...talented, passionate, and committed engineers, technologists, and business... ...Engineer to Support daily operations of high-performance networking infrastructure within our AI lab(s) and data center...OperationsWork at officeWorldwideNight shift- ...Supermicro in San Jose, CA is seeking a Data Center Engineer to support the design, deployment, operation, and maintenance of high-performance data center infrastructure, focusing on servers, storage, networking, power, and cooling. The role requires a Bachelor's degree...Operations
$147k - $210k
...devices products entering Google’s data centers focusing on verified boot and embedded... ...matter expertise to Platform Infrastructure Engineering (PIE) teams designing and developing... ...works to create and maintain the safest operating environment for Google's users and developers...$208.8k
...Responsibilities The Data Platform team at the TikTok USDS Joint Venture operates TikTok's US data processing ecosystem: a... ...compliance. About the Role As an Engineering Manager on the Data Platform... ...successes across multiple data centers. You will innovate with software...OperationsTemporary workLocal areaShift work$147k - $210k
...hardware, network, or service operations and quality.Minimum... ...with developing large-scale infrastructure, distributed systems or networks... ...services, cloud infrastructure, data processing systems or... ...infrastructure consumed by other engineering teams.Experience designing safety...Operations- Everforth Apex is seeking a Senior Network Engineer in Milpitas to design, implement, and support enterprise and data center network systems. You will serve as a subject... ...focuses on hands-on deployments, day-to-day operations, and strategic initiatives across data center...Operations
$184k - $287.5k
...forefront of technological advancement. NVIDIA data center systems have become core to NVIDIA's... ...are seeking an excellent Senior System Engineer to work on bring up, integration,... ...-on debugging skills in embedded Linux operating environments.Experience on using simulation...Full timeEarly shift$152k - $241.5k
The NVIDIA DGXC Data Services team builds cloud-native systems... ...hybrid and multi-cloud infrastructure. We are building the next-generation... ...build, train, deploy, and operate AI products at scale without... ...platform teams, and partner engineering teams to understand...OperationsFull time$174k - $252k
...hardware, network, or service operations and quality.Minimum... ...with developing large-scale infrastructure, distributed systems or networks... ....5 years of experience with data structures and algorithms.1... ...technologies.Google's software engineers develop the next-generation...Operations- Job Title: GPU Network Engineer Job Location: Sunnyvale, CA (hybrid... ...50k + BenefitsRequirements: Data center, HPC networking, InfiniBand,... ...we are an innovative energy infrastructure company that develops... ...Engineer to design, build, and operate the high-speed fabrics that...OperationsLocal areaRelocation
$120.75k - $161k
About the Role/TeamJoin Eightfold’s core Data Platform team to design, develop, and... ...scalability, and load balancing.Performance Engineering: Drive the scaling, optimization, and... ...evolving business and user needs.System Operations: Diagnose and resolve complex problems within...OperationsPermanent employmentWork at office3 days per week$147k - $210k
Develop and Maintain Testing Infrastructure, implement and support... ...testing systems focused on data storage and processing for large... ...performance at scale.Enhance engineering productivity, monitor engineering... ...(QE), Development and Operations (DevOps), engineering productivity...OperationsFlexible hours- MatX Inc. is seeking a senior optical interconnect engineer to own end-to-end development of next-generation interconnects for AI scale... ..., and collaborating with cross-functional teams to meet aggressive performance targets in data centers. #J-18808-Ljbffr MatX Inc.
$150k - $217k
...analysis of measurements or data and capacity forecast... ...with other engineering teams.Lead and improve... ...qualification, deployment, operation and optimization.Minimum... ...services.The AI and Infrastructure team is redefining what... ...Global Networking, Data Center operations, systems...OperationsWorldwide$105.3k - $175.21k
...seeking a Network Security Engineer. The candidate chosen... ...to support USG operations.Primary duties and responsibilities... ..., classification of data, etc.• Assist with... ...and Network security infrastructure, with network design... ...working with Data Center migrations, server upgrades...OperationsFull timeInternshipWork at officeLocal areaImmediate startShift work$207k - $300k
...a distributed team of engineers.Facilitate alignment and... ...large-scale infrastructure, distributed systems or... ...computing, large-scale data processing.Preferred qualifications... ....Experience with data center architecture and... ...the total cost of operations. In this role, you will...OperationsWorldwide$207k - $300k
...a distributed team of engineers.Scale and performance... ...enabling AI networking infrastructure, enabling high performing... ...of experience with data structures and algorithms... ...Google's data center and AI infrastructure.... ...maintaining the switch operation system deployed in every...OperationsWorldwide- ...What you will do: Design, build, and operate highly performant and scalable batch and stream data processing infrastructure and solutions to support day to day ML operations... ...6+ years of experience as senior/software engineer Experience with Python or Golang or Java or...Operations
$207k - $300k
Lead a team of software engineers to design, build, deploy, and in some cases, operate the critical software systems that directly manage the global data center networking infrastructure.Cultivate a high-performance culture of technical excellence, inspire engineers across...OperationsWorldwide- ...Job Title: Senior Network Engineer - Global Labs Must-Have Skills 5-8 years of experience in Enterprise Networking, Data Center Operations, or Lab Infrastructure. Strong knowledge of TCP/IP, Routing, Switching, VLANs, Subnetting, and Firewalls. Hands‑on experience with...OperationsRemote work
$200k - $322k
NVIDIA is looking for an exceptional Engineering Manager to lead, scale, and innovate our core Data Labeling Platform. This is a highly visible, high-impact role... ...engineering, scalable software systems, and massive-scale operations.In this role, you will lead a team of highly...OperationsFull time- ...a Principal Network Development Engineer within Oracle Cloud Infrastructure (OCI) Network Reliability Engineering... ...reliability, scalability, and operational excellence of one of the world's... ..., Arista, InfiniBand, Firewalls, Data center switching platforms, Circuit management...Operations
- ...Network Deployment Engineer Location: Santa Clara, CA & Ashburn, VA (Onsite)... ...execution for bringing up, expanding, and operating NPI infrastructure in Ashburn and Santa Clara,... ...dependencies and resolve blockers quickly. Data Center Network (DCN) / Fabric Engineering (...OperationsAfternoon shift
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Infrastructure Engineer (Data Center Operations). Be the first to apply!
- data infrastructure engineer Sunnyvale, CA
- infrastructure engineering manager Sunnyvale, CA
- senior infrastructure engineer Sunnyvale, CA
- principal infrastructure engineer Sunnyvale, CA
- remote infrastructure engineer Sunnyvale, CA
- lead infrastructure engineer Sunnyvale, CA
- infrastructure developer Sunnyvale, CA
- infrastructure engineer Sunnyvale, CA
- data center engineer Sunnyvale, CA
- senior cloud data engineer Sunnyvale, CA

