Infrastructure Engineer (Data Center Operations)
Cerebras Systems
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services. This order of magnitude increase in speed is transforming the user experience of AI applications, unlocking real-time iteration and increasing intelligence via additional agentic computation.Cerebras works with the leading model labs, global enterprises, and cutting-edge AI-native startups. OpenAI recently announced a multi-year partnership with Cerebras, to deploy 750 megawatts of scale, transforming key workloads with ultra high-speed inference.About the Role: We are looking for a hands-on Infrastructure Engineer to join our team and support our high-performance, on-premise server and networking infrastructure. You will be responsible for maintaining, provisioning, and troubleshooting hardware and Linux systems, working closely with network and system teams. This is an in-person role, ideal for someone who enjoys working across hardware, networking, and system layers.Key Responsibilities:Physically install, rack, cable, and maintain blade servers and hardware components (CPUs, DIMMs, NICs, storage devices, etc.)Connect servers to high-speed networks (100G/400G), verify optics/DACs, and check link statusConfigure BIOS, firmware, and out-of-band management (IPMI/iDRAC/iLO)Install and provision Linux OS; configure hostnames, IPs, routing, and NFS mount pointsDebug network issues at physical and OS level (VLAN, link issues, routing, etc.)Use Linux tools (e.g., ip, dmesg, netstat, ping) to isolate and fix issuesFollow provisioning playbooks and maintain accurate records of assets and changesUse scripting (Bash, Python) to automate routine tasks and improve efficiencyCollaborate with internal teams (network, systems, storage) and coordinate vendor RMAsDocument procedures and contribute to team knowledge baseTroubleshoot and replace failed server components with minimal downtimeQualifications:3–5+ years of experience in data center, lab, or infrastructure engineering rolesProficient in Linux system administration and network configurationStrong hands-on knowledge of x86 server hardware and enterprise networkingFamiliar with BIOS configuration, firmware updates, and remote management toolsSkilled in physical setup and troubleshooting of high-speed NICs and optical linksExperience with VLANs, static routing, and diagnosing layer 1–3 issuesAbility to write scripts for automation and diagnostics (Bash, Python preferred)Comfortable working on-site daily and lifting/moving server hardwarePreferred Skills:Experience with PXE, NFS, RAID controllers, and monitoring toolsFamiliarity with configuration management tools (e.g., Ansible)Prior experience in a lab or R&D hardware/software environmentThis is a unique opportunity to work with cutting-edge infrastructure and grow into more senior technical roles. If you enjoy bridging hardware and software with hands-on work, we’d love to hear from you.ADD DESCRIPTION HEREWhy Join CerebrasPeople who are serious about software make their own hardware. At Cerebras, we have built a breakthrough architecture that is unlocking new opportunities for the AI industry. With dozens of model releases and rapid growth, we’ve reached an inflection point in our business. Members of our team tell us there are five main reasons they joined Cerebras:Build a breakthrough AI platform beyond the constraints of the GPU.Publish and open source their cutting-edge AI research.Work on one of the fastest AI supercomputers in the world.Enjoy job stability with startup vitality.Our simple, non-corporate work culture that respects individual beliefs.Find out more about what it's like to work at Cerebras here! Apply today and become part of the forefront of groundbreaking advancements in AI!Cerebras Systems is committed to creating an equal and diverse environment and is proud to be an equal opportunity employer. We celebrate different backgrounds, perspectives, and skills. We believe inclusive teams build better products and companies. We try every day to build a work environment that empowers people to do their best work through continuous learning, growth and support of those around them.This website or its third-party tools process personal data. For more details, click here to review our CCPA disclosure notice.LocationSunnyvale, CAEmployment TypeFull timeLocation TypeHybridDepartmentSoftware Engineering
$165.2k - $223.6k
AWS Infrastructure Services owns the design, planning, delivery, and operation of all AWS global infrastructure. In other words, we... ...running. We support all AWS data centers and all of the servers, storage... ..., hardware, and network engineers, supply chain specialists, security...OperationsInternshipLocal areaFlexible hours$200k - $220k
...only vertically integrated AI infrastructure company built from the ground up, we own and operate each layer of the stack — from... ...across energy, manufacturing, data center construction, and cloud services... ...Energy as a Senior Data Engineer, an early and pivotal hire on...OperationsFull timeTemporary work$207k - $300k
...software systems, programming the data plane and control plane for... ...world's largest network infrastructure.Manage individual project... ...:Master’s degree or PhD in Engineering, Computer Science, or a... ...Google Global Networking, Data Center operations, systems research, and much...OperationsWorldwide$165k - $242k
...CoreWeave combines superior infrastructure performance with deep... ...the role: The Data Platforms Team serves... ...solutions, automation and operations of our data platform... ...are seeking a senior engineer with specialization in... ...in our office and data center locations ~ A casual...OperationsPermanent employmentFull timeTemporary workCasual workWork at officeRemote workFlexible hours- ...experienced Senior Customer Quality Engineer (CQE) to lead customer quality engagements for strategic Data Center, Networking and AI Infrastructure, customers. This role will support AMD... ...engineering, product line quality, operations, design, reliability, and...Operations
$115k - $130k
..., and networking solutions for Data Center, Cloud Computing, Enterprise IT... ...talented, passionate, and committed engineers, technologists, and business... ...Engineer to Support daily operations of high-performance networking infrastructure within our AI lab(s) and data center...OperationsWork at officeWorldwideNight shift$147k - $210k
...hardware, network, or service operations and quality.Minimum... ...with developing large-scale infrastructure, distributed systems or networks... ...services, cloud infrastructure, data processing systems or... ...infrastructure consumed by other engineering teams.Experience designing safety...Operations- ...the Team The Monetization Data Platform team builds the trusted... ...data to help Product, Engineering, Finance, and GTM teams make... ...observability, and ongoing operation. This is a hands-on role... ...event streaming, or payment infrastructure. Experience building internal...OperationsFull time
$120.75k - $161k
...the Role/Team Join Eightfold\u0027s core Data Platform team to design, develop, and... ...scalability, and load balancing. Performance Engineering: Drive the scaling, optimization, and... ...evolving business and user needs. System Operations: Diagnose and resolve complex problems...OperationsPermanent employmentFull timeWork at office3 days per week$184k - $287.5k
...forefront of technological advancement. NVIDIA data center systems have become core to NVIDIA's... ...are seeking an excellent Senior System Engineer to work on bring up, integration,... ...-on debugging skills in embedded Linux operating environments.Experience on using simulation...Full timeEarly shift$152k - $241.5k
The NVIDIA DGXC Data Services team builds cloud-native systems... ...hybrid and multi-cloud infrastructure. We are building the next-generation... ...build, train, deploy, and operate AI products at scale without... ...platform teams, and partner engineering teams to understand...OperationsFull time- ...relationships and define technical requirements for AI/ML workloads on GPU infrastructure. You will drive end-to-end program execution across data centers, collaborating with engineering, product, and operations to ensure successful outcomes. You will gather customer feedback,...Operations
- Job Title: GPU Network Engineer Job Location: Sunnyvale, CA (hybrid... ...50k + BenefitsRequirements: Data center, HPC networking, InfiniBand,... ...we are an innovative energy infrastructure company that develops... ...Engineer to design, build, and operate the high-speed fabrics that...OperationsLocal areaRelocation
$135k - $180k
Santa Clara, CAData Engineering - Backend Engineering /Full... ...Silicon Valley with operations in the United States... ...and maintain optimal data processing pipeline architectureBuild the infrastructure required for optimal extraction... ...through multiple data centers and AWS...OperationsFull time$160k - $211k
...Lattice OS, an AI-powered operating system that turns thousands of data streams into a realtime... ...3D command and control center. As the world enters an... ...the TeamThe Thermal Engineering team's work is essential... ...command-and-control infrastructure built to run at forward...OperationsFull timeWork experience placementImmediate start$150k - $217k
...network systems across various engineering teams.Engage in and improve... ..., through deployment, operation, and optimization.Drive innovation... ...and services.The AI and Infrastructure team is redefining what’s... ...Google Global Networking, Data Center operations, systems research...OperationsWorldwide$150k - $217k
...analysis of measurements or data and capacity forecast... ...with other engineering teams.Lead and improve... ...qualification, deployment, operation and optimization.Minimum... ...services.The AI and Infrastructure team is redefining what... ...Global Networking, Data Center operations, systems...OperationsWorldwide$168k - $264.5k
...transforming how the world builds and operates computing infrastructure for accelerated computing and AI. Our Network Deployment Engineering team designs and delivers the global network... ...infrastructure supporting NVIDIA's data centers, offices, labs, and critical business...OperationsFull time$105.3k - $175.21k
...seeking a Network Security Engineer. The candidate chosen... ...to support USG operations.Primary duties and responsibilities... ..., classification of data, etc.• Assist with... ...and Network security infrastructure, with network design... ...working with Data Center migrations, server upgrades...OperationsFull timeInternshipWork at officeLocal areaImmediate startShift work$114.6k - $234.6k
...network behavior across OCI’s data centers and WAN. You will combine... ..., policy controls, and operational safeguards.This position requires... ...brings together the data, infrastructure, applications, and... ...analysis while maintaining engineering, security, and quality standards...OperationsTemporary workFlexible hours$320k
...looking for a Distinguished Engineer to act as a senior technical... ...enthusiastic about cluster operations involving DGX Cloud GPU capacity... ...frameworks that sustain GPU infrastructure health, scalability, and... ...researchers and customers.This role centers on the operational framework...OperationsFull timeRemote work$207k - $300k
...a distributed team of engineers.Facilitate alignment and... ...large-scale infrastructure, distributed systems or... ...computing, large-scale data processing.Preferred qualifications... ....Experience with data center architecture and... ...the total cost of operations. In this role, you will...OperationsWorldwide$207k - $300k
...a distributed team of engineers.Scale and performance... ...enabling AI networking infrastructure, enabling high performing... ...of experience with data structures and algorithms... ...Google's data center and AI infrastructure.... ...maintaining the switch operation system deployed in every...OperationsWorldwide$168k - $270.25k
## Senior Software Engineer, Data PlatformApply: US, CA, Santa Clara: Full time: Posted Today: JR2010934The NVIDIA Operations organization is seeking an experienced software engineering professional for the position of System Data Architect. As a member of our team you...OperationsFull time$200k - $322k
NVIDIA is looking for an exceptional Engineering Manager to lead, scale, and innovate our core Data Labeling Platform. This is a highly visible, high-impact role... ...engineering, scalable software systems, and massive-scale operations.In this role, you will lead a team of highly...OperationsFull time$207k - $300k
Lead a team of software engineers to design, build, deploy, and in some cases, operate the critical software systems that directly manage the global data center networking infrastructure.Cultivate a high-performance culture of technical excellence, inspire engineers across...OperationsWorldwide$163k - $236k
...qualifications:Bachelor's degree in Electrical Engineering, Computer Engineering, Computer... ...AI/ML-driven systems.As an SoC Test Infrastructure Engineer, you will drive the post-... ...Cloud, Google Global Networking, Data Center operations, systems research, and much more....OperationsWorldwide- ...We Are Synopsys is the leader in engineering solutions from silicon to systems,... ...expertise spans the full stack of network infrastructure including routing, switching, wireless, security, and data center technologies, and you have operated at a senior or architecture level...OperationsWork experience placementRemote work
- ...entity.Atlassian’s Trusted Data Platform (TDP) provides secure... ..., backup and recovery, and operation across multiple cloud environments... ...database platform. It gives engineering teams a supported way to... ...to manage the underlying infrastructure, availability, security, scaling...OperationsWork at officeLocal area
$175k - $252k
...connectivity in complex 3D environments.Engineer reliable E2E wireless links in... ....10 years of experience with data center networking architecture, operations, and power distribution systems within... ...gap between complex corporate infrastructure and cutting-edge autonomous...OperationsRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Infrastructure Engineer (Data Center Operations). Be the first to apply!
- principal infrastructure engineer Sunnyvale, CA
- infrastructure engineer Sunnyvale, CA
- infrastructure engineering manager Sunnyvale, CA
- infrastructure developer Sunnyvale, CA
- remote infrastructure engineer Sunnyvale, CA
- senior infrastructure engineer Sunnyvale, CA
- data infrastructure engineer Sunnyvale, CA
- big data cloud engineer Sunnyvale, CA
- big data developer Sunnyvale, CA
- finance data engineer Sunnyvale, CA


