Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Infrastructure Engineer (Data Center Operations)

Cerebras Systems

Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services. This order of magnitude increase in speed is transforming the user experience of AI applications, unlocking real-time iteration and increasing intelligence via additional agentic computation.Cerebras works with the leading model labs, global enterprises, and cutting-edge AI-native startups. OpenAI recently announced a multi-year partnership with Cerebras, to deploy 750 megawatts of scale, transforming key workloads with ultra high-speed inference.About the Role: We are looking for a hands-on Infrastructure Engineer to join our team and support our high-performance, on-premise server and networking infrastructure. You will be responsible for maintaining, provisioning, and troubleshooting hardware and Linux systems, working closely with network and system teams. This is an in-person role, ideal for someone who enjoys working across hardware, networking, and system layers.Key Responsibilities:Physically install, rack, cable, and maintain blade servers and hardware components (CPUs, DIMMs, NICs, storage devices, etc.)Connect servers to high-speed networks (100G/400G), verify optics/DACs, and check link statusConfigure BIOS, firmware, and out-of-band management (IPMI/iDRAC/iLO)Install and provision Linux OS; configure hostnames, IPs, routing, and NFS mount pointsDebug network issues at physical and OS level (VLAN, link issues, routing, etc.)Use Linux tools (e.g., ip, dmesg, netstat, ping) to isolate and fix issuesFollow provisioning playbooks and maintain accurate records of assets and changesUse scripting (Bash, Python) to automate routine tasks and improve efficiencyCollaborate with internal teams (network, systems, storage) and coordinate vendor RMAsDocument procedures and contribute to team knowledge baseTroubleshoot and replace failed server components with minimal downtimeQualifications:3–5+ years of experience in data center, lab, or infrastructure engineering rolesProficient in Linux system administration and network configurationStrong hands-on knowledge of x86 server hardware and enterprise networkingFamiliar with BIOS configuration, firmware updates, and remote management toolsSkilled in physical setup and troubleshooting of high-speed NICs and optical linksExperience with VLANs, static routing, and diagnosing layer 1–3 issuesAbility to write scripts for automation and diagnostics (Bash, Python preferred)Comfortable working on-site daily and lifting/moving server hardwarePreferred Skills:Experience with PXE, NFS, RAID controllers, and monitoring toolsFamiliarity with configuration management tools (e.g., Ansible)Prior experience in a lab or R&D hardware/software environmentThis is a unique opportunity to work with cutting-edge infrastructure and grow into more senior technical roles. If you enjoy bridging hardware and software with hands-on work, we’d love to hear from you.ADD DESCRIPTION HEREWhy Join CerebrasPeople who are serious about software make their own hardware. At Cerebras, we have built a breakthrough architecture that is unlocking new opportunities for the AI industry. With dozens of model releases and rapid growth, we’ve reached an inflection point in our business. Members of our team tell us there are five main reasons they joined Cerebras:Build a breakthrough AI platform beyond the constraints of the GPU.Publish and open source their cutting-edge AI research.Work on one of the fastest AI supercomputers in the world.Enjoy job stability with startup vitality.Our simple, non-corporate work culture that respects individual beliefs.Find out more about what it's like to work at Cerebras here! Apply today and become part of the forefront of groundbreaking advancements in AI!Cerebras Systems is committed to creating an equal and diverse environment and is proud to be an equal opportunity employer. We celebrate different backgrounds, perspectives, and skills. We believe inclusive teams build better products and companies. We try every day to build a work environment that empowers people to do their best work through continuous learning, growth and support of those around them.This website or its third-party tools process personal data. For more details, click here to review our CCPA disclosure notice.LocationHeadquarters/Sunnyvale OfficeEmployment TypeFull timeLocation TypeOn-siteDepartmentSoftware

Vacancy posted 7 hours ago
Similar jobs that could be interesting for youBased on the Infrastructure Engineer (Data Center Operations) in Sunnyvale, CA vacancy
  • $165.2k - $223.6k

    AWS Infrastructure Services owns the design, planning, delivery, and operation of all AWS global infrastructure. In other words, we...  ...running. We support all AWS data centers and all of the servers, storage...  ..., hardware, and network engineers, supply chain specialists, security... 
    Operations
    Internship
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    2 days ago
  • $136.88k - $205k

     ...essential building blocks of the data infrastructure that connects our world....  ...of Ethernet and Data center products from New Production...  ...part of a dynamic product engineering team working on most advanced...  ...product characterization, test operation, qualification, yield ramp,... 
    Operations
    Permanent employment
    Full time
    Internship
    Work from home

    Marvell

    Santa Clara, CA
    3 days ago
  • $200k - $220k

     ...only vertically integrated AI infrastructure company built from the ground up, we own and operate each layer of the stack — from...  ...across energy, manufacturing, data center construction, and cloud services...  ...Energy as a Senior Data Engineer, an early and pivotal hire on... 
    Operations
    Full time
    Temporary work

    Crusoe Energy Systems

    Sunnyvale, CA
    1 day ago
  • $136.88k - $205k

     ...essential building blocks of the data infrastructure that connects our world....  ...and Compute Product Engineering organization at Marvell serves...  ...development and NPI. The team operates across multiple business...  ...scale across hyperscale data centers and emerging AI workloads.... 
    Operations
    Permanent employment
    Full time
    Internship
    Work from home

    Marvell

    Santa Clara, CA
    4 days ago
  • $147k - $210k

     ...devices products entering Google’s data centers focusing on verified boot and embedded...  ...matter expertise to Platform Infrastructure Engineering (PIE) teams designing and developing...  ...works to create and maintain the safest operating environment for Google's users and developers... 
    Suggested

    Google

    Sunnyvale, CA
    1 day ago
  • $307k - $427k

     ...policies for next-generation platforms to minimize both operational and embodied carbon.Drive power optimization...  ...carbon emissions.Partner with chip design, platform engineering, software infrastructure, data center operations, and climate science teams to resolve system... 
    Operations
    Worldwide

    Google

    Sunnyvale, CA
    7 hours ago
  • $100k - $200k

     ...transforming how AI and cloud infrastructure teams manage modern networks...  ...generation of AI and cloud operators. With cutting-edge...  ...business is helping design modern data center architectures and build networks...  ...Center & IT Infrastructure Engineer to manage and support physical... 
    Operations
    Permanent employment
    Work at office
    Local area
    Remote work
    Santa Clara, CA
    more than 2 months ago
  • $184k - $287.5k

     ...forefront of technological advancement. NVIDIA data center systems have become core to NVIDIA's...  ...are seeking an excellent Senior System Engineer to work on bring up, integration,...  ...-on debugging skills in embedded Linux operating environments.Experience on using simulation... 
    Full time
    Early shift

    Nvidia

    Santa Clara, CA
    7 hours ago
  • $174k - $253k

     ...hardware, network, or service operations and quality.Minimum...  ...experience developing large-scale infrastructure, distributed systems or networks...  ...5 years of experience with data structures and algorithms.1...  .... Google's software engineers develop the next-generation... 
    Operations

    Google

    Sunnyvale, CA
    7 hours ago
  • $152k - $241.5k

    The NVIDIA DGXC Data Services team builds cloud-native systems...  ...hybrid and multi-cloud infrastructure. We are building the next-generation...  ...build, train, deploy, and operate AI products at scale without...  ...platform teams, and partner engineering teams to understand... 
    Operations
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  • $147k - $210k

    Develop and Maintain Testing Infrastructure, implement and support...  ...testing systems focused on data storage and processing for large...  ...performance at scale.Enhance engineering productivity, monitor engineering...  ...(QE), Development and Operations (DevOps), engineering productivity... 
    Operations
    Flexible hours

    Google

    Mountain View, CA
    7 hours ago
  • $184k - $287.5k

    NVIDIA’s Hardware Infrastructure organization is seeking a Senior System Software Engineer to lead the evolution of our next-generation Data & Observability Platform. We serve and collaborate...  ...management, and complex pipeline operations.Work in a diverse team to provide... 
    Operations
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $120.75k - $161k

    About the Role/TeamJoin Eightfold’s core Data Platform team to design, develop, and...  ...scalability, and load balancing.Performance Engineering: Drive the scaling, optimization, and...  ...evolving business and user needs.System Operations: Diagnose and resolve complex problems within... 
    Operations
    Permanent employment
    Work at office
    3 days per week

    Eightfold

    Santa Clara, CA
    4 days ago
  • $207k - $301k

    Lead a team software engineers responsible for developing the SDN...  ...across Networking Infrastructure and Cloud with product managers...  ...applications built for the Google data center.Define a goal and long-term...  ...support large scale system operations.Mentor individual engineers... 
    Operations
    Worldwide
    Flexible hours

    Google

    Sunnyvale, CA
    7 hours ago
  • $150k - $217k

     ...analysis of measurements or data and capacity forecast...  ...with other engineering teams.Lead and improve...  ...qualification, deployment, operation and optimization.Minimum...  ...services.The AI and Infrastructure team is redefining what...  ...Global Networking, Data Center operations, systems... 
    Operations
    Worldwide

    Google

    Sunnyvale, CA
    1 day ago
  • $105.3k - $175.21k

     ...seeking a Network Security Engineer. The candidate chosen...  ...to support USG operations.Primary duties and responsibilities...  ..., classification of data, etc.• Assist with...  ...and Network security infrastructure, with network design...  ...working with Data Center migrations, server upgrades... 
    Operations
    Full time
    Internship
    Work at office
    Local area
    Immediate start
    Shift work

    Intel

    Santa Clara, CA
    4 days ago
  • $165k - $242k

     ...CoreWeave combines superior infrastructure performance with deep...  ...the role: The Data Platforms Team serves...  ...solutions, automation and operations of our data platform...  ...are seeking a senior engineer with specialization in...  ...in our office and data center locations ~ A casual... 
    Operations
    Permanent employment
    Full time
    Temporary work
    Casual work
    Work at office
    Flexible hours

    CoreWeave

    Sunnyvale, CA
    16 days ago
  • $207k - $301k

    Lead a team of software engineers to design, build, deploy, and in some cases, operate the critical software systems that directly manage the global data center networking infrastructure.Cultivate a high-performance culture of technical excellence, inspire engineers across... 
    Operations
    Worldwide

    Google

    Sunnyvale, CA
    7 hours ago
  • $163k - $237k

     ...qualifications:Bachelor's degree in Electrical Engineering, Computer Engineering, Computer...  .../ML-driven systems. As an SoC Test Infrastructure Engineer, you will drive the post-...  ...Cloud, Google Global Networking, Data Center operations, systems research, and much more.Individual... 
    Operations
    Worldwide

    Google

    Sunnyvale, CA
    3 days ago
  • $126.1k - $185k

     ...use network and Unix systems engineering to deliver simple,...  ...-generation IP networks?AWS Infrastructure Services owns the design, planning, delivery, and operation of all AWS global infrastructure...  ...running. We support all AWS data centers and all of the servers, storage... 
    Operations
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    2 days ago
  • $126k - $181k

     ...escalating complex issues to senior engineers or managers.Develop and...  ...and growing network is operating at its peak potential. You...  ...perform at scale.The AI and Infrastructure team is redefining what’s...  ...Google Global Networking, Data Center operations, systems research... 
    Operations
    Worldwide

    Google

    Sunnyvale, CA
    7 hours ago
  • $177k - $257k

     ...connectivity in complex 3D environments.Engineer reliable E2E wireless links in...  ....10 years of experience with data center networking architecture, operations, and power distribution systems within...  ...gap between complex corporate infrastructure and cutting-edge autonomous... 
    Operations
    Remote work

    Google

    Mountain View, CA
    7 hours ago
  • $50 - $60 per hour

     ...for advancement Network Infrastructure Engineer Fort Collins, CO (...  ...architecture, security, and operations teams to deliver scalable,...  ...availability issues IPAM & Network Data Modernization Partner...  ...virtual firewalls Data center lab or R&D infrastructure... 
    Operations
    Hourly pay
    Full time
    Contract work
    Work from home
    Flexible hours

    SelectMinds

    Santa Clara, CA
    5 days ago
  •  ...Network Engineer The Network Engineer is responsible for the...  ...company's campus-wide network infrastructure. This role performs...  ...expansions, and reliable business operations. Duties and...  ...infrastructure, including data centers and server services. Education... 
    Operations
    For contractors
    Local area

    Foxconn Technology Group

    Santa Clara, CA
    2 days ago
  • $114.6k - $234.6k

    Oracle Cloud Infrastructure (OCI) is seeking a highly motivated Principal Data Automation Engineer to help accelerate the digital transformation...  ...OCI's global supply chain operations. Reporting to the Senior...  ...Logistics, Capacity Planning, Data Center Operations, and Engineering... 
    Operations
    Temporary work
    Flexible hours

    Oracle Corporation

    Santa Clara, CA
    1 day ago
  •  ...Senior Network Engineer Job Location: Sunnyvale, CA Job...  ...wireless coverage for business operations in R&D facility and office...  ...outstanding WAN/LAN/Data center/ Wireless/firewalls design...  ...and monitoring of network infrastructure ~ Solid knowledge of general... 
    Operations
    Contract work
    Work at office

    InterSources

    Sunnyvale, CA
    1 day ago
  •  ...computing experiences—from AI and data centers, to PCs, gaming and embedded...  ...are hiring AI / ML Platform Engineers to build the platform layer...  .... This role focuses on the infrastructure and platform systems that...  ...RESPONSIBILITIESBuild and operate the shared AI platform for agentic... 
    Operations

    AMD

    Santa Clara, CA
    3 days ago
  • $163k - $237k

     ...Bachelor's degree in Electrical Engineering, Computer Engineering,...  ...machine learning workloads in data centers. As a member of the inter-...  ...environment.The AI and Infrastructure team is redefining what’s possible...  ...Networking, Data Center operations, systems research, and much... 
    Operations
    Worldwide

    Google

    Sunnyvale, CA
    7 hours ago
  • $90k - $150k

     ...Networks is an industry leader in data-driven, client-to-cloud networking for large data center, campus and routing...  ...prestigious awards, such as Best Engineering Team, Best Company for Diversity...  ...concerns raised during installation, operation, maintenance of Arista productsInvestigate... 
    Operations

    Arista Networks

    Santa Clara, CA
    2 days ago
  • $148k - $224.25k

     ...skills for GPU product definition for Data Center. We are a small, dynamic, and motivated...  ...previous product management, AI related engineering, design or development experience highly...  ...including data center total cost of operation and token revenuesDemonstrated ability... 
    Operations
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Infrastructure Engineer (Data Center Operations). Be the first to apply!