Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

HPC Engineer (Biohub Network)

jobright.com

Join to apply for the HPC Engineer (Biohub Network) role at Jobright.ai 2 days ago Be among the first 25 applicants Join to apply for the HPC Engineer (Biohub Network) role at Jobright.ai Jobright is an AI-powered career platform that helps job seekers discover the top opportunities in the US. We are NOT a staffing agency. Jobright does not hire directly for these positions. We connect you with verified openings from employers you can trust. Job Summary: The Chan Zuckerberg Biohub Network is an independent nonprofit research institute that collaborates with leading universities to drive innovative research in biology and disease. They are seeking an experienced Principal High Performance Computing (HPC) Engineer to develop, support, and optimize HPC systems, enhancing the scientific computational capacity of the Biohub. Responsibilities:

  • Manage cluster-level services via the SLURM scheduler as well as user facing services such as Open OnDemand and NoMachine
  • Install, configure and optimize applications and provide user support
  • Work closely with many different science teams simultaneously to translate experimental descriptions into software and hardware requirements and across all phases of the scientific lifecycle, including data ingest, analysis, management and storage, computation, authentication, tool development and many other computational needs expressed by scientific projects
Qualifications: Required:
  • Bachelor’s Degree in Computer Science, Mathematics, Systems Engineering or a related field or equivalent training/experience also acceptable
  • A minimum of 7 years of experience with progressively increasing responsibility in HPC computing environments or complex Linux environments
  • Experience building on-prem HPC infrastructure and capacity planning
  • Experience and expertise working on complex issues where analysis of situations or data requires an in-depth evaluation of variable factors
  • Experience supporting scientific facilities, and prior knowledge of scientific user needs, program management, data management planning or lab-bench IT needs
  • Experience with HPC and cloud computing environments
  • Ability to interact with a variety of technical and scientific personnel with varied academic backgrounds
  • Strong written and verbal communication skills to present and disseminate scientific software developments at group meetings
  • Demonstrated ability to reason clearly about load, latency, bandwidth, performance, reliability, and cost and make sound engineering decisions balancing them
  • Demonstrated ability to quickly and creatively implement novel solutions and ideas
  • Proven ability to analyze, troubleshoot, and resolve complex problems that arise in the HPC production storage hardware, software systems, storage networks and systems
  • Configuring and administering parallel, network attached storage (Lustre, NFS, ESS, Ceph) and storage subsystems (e.g. IBM, NetApp, DataDirect Network, LSI, etc.)
  • Installing, configuring, and maintaining job management tools (such as SLURM, Moab, TORQUE, PBS, etc.)
  • Red Hat Enterprise Linux, CentOS, or derivatives and Linux services and technologies like dnsmasq, systemd, LDAP, PAM, sssd, OpenSSH, cgroups
  • Virtualization (ESXi or KVM/libvirt), containerization (Docker or Singularity), configuration management and automation (tools like xCAT, Puppet, kickstart) and orchestration (Kubernetes, docker-compose, CloudFormation, Terraform.)
  • High performance networking technologies (Ethernet and Infiniband) and hardware (Mellanox and Juniper)
  • Configuring, installing, tuning and maintaining scientific application software
  • Familiarity with source control tools (Git or SVN)
Preferred:
  • Understand and translate researchers' scientific challenges into computational solutions
  • Scientific background, research experience, and/or experience in a University or a research setting
Company: Chan Zuckerberg Biohub Network is a non-profit organisation. Founded in 2015, headquartered in San Francisco, California, USA, team size 201-500 employees, currently Growth Stage. Chan Zuckerberg Biohub Network has a track record of offering H1B sponsorships. Seniority level Seniority level Mid-Senior level Employment type Employment type Full-time Job function Industries Software Development Referrals increase your chances of interviewing at Jobright.ai by 2x Inferred from the description for this job Medical insurance Vision insurance 401(k) Get notified about new Network Engineer jobs in San Francisco, CA . Oakland, CA $120,000.00-$140,000.00 3 weeks ago San Leandro, CA $110,000.00-$125,000.00 2 weeks ago San Francisco, CA $110,000.00-$155,000.00 4 days ago Brisbane, CA $155,000.00-$170,000.00 6 days ago Network Engineer - Data for Autonomous Systems Brisbane, CA $120,000.00-$150,000.00 1 week ago Brisbane, CA $90,000.00-$120,000.00 1 week ago San Mateo, CA $238,520.00-$289,460.00 1 week ago San Francisco, CA $168,000.00-$188,999.00 5 days ago San Francisco, CA $80,000.00-$200,000.00 1 month ago San Francisco, CA $148,512.00-$214,517.00 3 weeks ago Brisbane, CA $112,000.00-$150,000.00 2 weeks ago Alameda, CA $119,000.00-$175,000.00 5 days ago Oakland, CA $135,000.00-$175,000.00 1 week ago Senior System/Network Engineer (Windows-Linux-ProxMox) High-Performance Networking Engineer - Supercomputing Network Engineer - IS Communications Specialist II - Limited Term San Mateo County, CA $123,323.20-$154,148.80 5 months ago Oakland, CA $130,000.00-$160,000.00 1 week ago Foster City, CA $115,000.00-$145,000.00 2 weeks ago San Francisco, CA $16.15-$23.97 5 days ago We’re unlocking community knowledge in a new way. Experts add insights directly into each article, started with the help of AI. #J-18808-Ljbffr jobright.com

Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the HPC Engineer (Biohub Network) in San Francisco, CA vacancy
  • $148.7k - $201.2k

     ...a multi-disciplinary team of scientists, engineers, and technicians, on a mission to develop...  ...quantum computer.We are looking to hire an HPC Platform Engineer to develop, automate,...  ...CloudFormation and Lambda- Experience in network fundamentals (DNS, DHCP, TCP/IP, routing,... 
    Suggested
    Local area
    Flexible hours

    Amazon

    San Francisco, CA
    2 days ago
  • $210k - $240k

     ...reported to LinkedIn. Build and help define the network foundation behind a high-performance AI...  ...enterprise workloads. This is a network engineering role first . We are looking for a deeply...  ...a Plus InfiniBand networking for GPU or HPC environments RDMA or RoCEv2 networking,... 
    Suggested
    Immediate start

    Stratitech

    San Francisco, CA
    4 days ago
  • $300k

     ...training and inference workloads with a global network of compute providers, supporting some of...  .... This company is seeking the first engineer in this function to build and define the...  ...is ideal for someone with deep HPC experience who wants to shape technical standards... 
    Suggested
    Full time
    San Francisco, CA
    23 days ago
  • $224k - $284k

    Atoms is seeking a network engineer based in San Francisco to design and optimize the high-performance network fabric for their GPU and CPU operations...  ...will have substantial experience with network design for HPC environments, familiarity with various switch platforms, and a... 
    Suggested

    ATOMS Careers page

    San Francisco, CA
    4 days ago
  • $224k - $284k

     ...and improve them until they work at scale. We are roboticists, engineers, operators, and builders. We believe the next great technology...  ...real-world impact, join us. What you’ll do We’re seeking a HPC Network Engineer to join our founding team who will design and run the... 
    Suggested
    Full time
    Work at office
    Immediate start
    Flexible hours

    Atoms

    San Francisco, CA
    2 days ago
  •  ...large-scale systems that span data centers, GPUs, networking, and more, ensuring high availability, performance,...  ...unchecked growth. About the role As a software engineer on the Fleet High Performance Computing (HPC) team, you will be responsible for the reliability... 
    Full time

    OpenAI

    San Francisco, CA
    19 hours ago
  • Atoms is hiring an HPC Network Engineer in San Francisco to design and manage high-performance networks connecting GPU compute. The ideal candidate will have expertise in network design, capacity planning, and hands-on experience with various switch platforms. This role... 

    Atoms

    San Francisco, CA
    2 days ago
  • $224k - $284k

    Cssmerge is looking for an HPC Network Engineer to join our founding team in San Francisco, responsible for designing and managing high-performance networking that connects our GPU compute. The ideal candidate has experience with network deployment and scaling in HPC or... 

    Cssmerge

    San Francisco, CA
    1 day ago
  • $226k - $285k

     ...the next generation of AI. The Data Center Engineering team defines the strategy, reference...  ...across electrical, mechanical, controls, network, hardware, construction, commissioning, deployment...  ...data centers, AI infrastructure, HPC environments, colocation, or partner-delivered... 
    Work at office
    Local area
    Flexible hours
    Shift work

    OpenAI

    San Francisco, CA
    3 days ago
  •  ...sized, and leadership is earned by shipping excellence. We seek engineers with strong intrinsic drive, a true passion for advancing the...  ...Angeles. About the Role We’re looking for a systems engineer with HPC or parallel programming experience to help scale AI inference.... 
    Full time
    Work at office

    Vast.ai Inc.

    San Francisco, CA
    4 days ago
  • $250k - $320k

    Gimlet Labs, Inc. is seeking a Network Engineer to design and build network infrastructure for AI workloads at scale. This role involves ensuring robust and reliable networking for production systems across distributed environments, focusing on performance and efficiency... 

    Gimlet Labs, Inc.

    San Francisco, CA
    3 days ago
  • $275k

     ...platform. This company is seeking a Founding Engineer to take ownership of core GPU cloud...  ..., distributed storage, high-bandwidth networking, and inference platforms. This hands-on role...  ...experience operating large-scale AI or HPC infrastructure in production ~ Proven experience... 
    Full time
    Relocation
    San Francisco, CA
    25 days ago
  • Crusoe Cloud is hiring a Senior Cloud Support Engineer to empower customers with sustainable, low-cost GPU compute. You will be the primary...  ..., on-call rotations, and collaboration with SRE, Networking, and Storage teams. The role requires Linux CLI, Kubernetes, Slurm... 

    Crusoe

    San Francisco, CA
    19 hours ago
  • Crusoe Cloud is seeking a Cloud Support Engineer to empower customers leveraging Crusoe Cloud for AI workloads. You will be the primary...  ..., triage, and escalations while collaborating with SRE, Networking, and Storage teams to ensure reliable performance. The role emphasizes... 
    Remote work

    Crusoe

    San Francisco, CA
    19 hours ago
  • $166k - $343k

    Presales Systems Engineer - HPE Networking (Northern California)This role has been designated as ‘Remote/Teleworker’, which means you will primarily work from home.Who We Are:Hewlett Packard Enterprise is the global edge-to-cloud company advancing the way people live and... 
    Full time
    Work experience placement
    Local area
    Immediate start
    Remote work
    Work from home

    Hewlett Packard Enterprise

    San Francisco, CA
    15 hours ago
  • A leading AI research company in San Francisco is seeking a software engineer for its Fleet High Performance Computing team. In this role, you'll ensure the reliability and uptime of the compute fleet, working with automation systems and monitoring tools. Ideal candidates... 

    Jobleads-US

    San Francisco, CA
    19 hours ago
  • $156.86k - $191.72k

     ...(NERSC) is seeking a System Infrastructure / Platform Engineer to help build and manage HPC systems and Linux-based infrastructure. NERSC operates...  ...such as CPU/GPU clusters, parallel storage, high-speed networking, Slurm, and Kubernetes, balancing innovation with reliability... 
    Permanent employment
    Full time
    Remote work
    Flexible hours

    Berkeley Lab

    Berkeley, CA
    1 day ago
  •  ...Greylock, and Conviction. Join us and help build the platform engineers turn to to ship AI products. At Baseten, we are building the...  .... We believe that as LLM and multi-modal workloads scale, the network is the computer. We are looking for foundational engineers to lead... 
    Full time
    Flexible hours

    Baseten

    San Francisco, CA
    19 hours ago
  • $157k - $239k

     ...CA / Golden, COInfrastructure - Cloud Infrastructure /Full time /On-siteWanna join the adventure?As a Site Reliability Engineer with strong networking skills in our Cloud Infrastructure (SRE) team, you help the team own the networks that keep Loft running: cloud networking... 
    Full time
    Temporary work

    Loft Orbital

    San Francisco, CA
    2 days ago
  • StratITech is hiring a senior network engineer to architect and operate secure, high-performance data center and edge networks supporting AI and distributed compute workloads. You will translate loosely defined needs into concrete requirements and plans, and own the end... 

    Stratitech

    San Francisco, CA
    4 days ago
  • $275k - $378k

     ...career-defining work. We're all in on this mission. If you are too, let's talk.The Director, Engineering OpportunityYou will lead the engineering charter for the Okta Integration Network, driving the evolution of a platform that powers 8,000+ integrations across the Okta... 
    Permanent employment
    Work at office
    Local area
    Worldwide
    Flexible hours

    Okta

    San Francisco, CA
    2 days ago
  •  ...of the world’s largest AI infrastructure networks. The team owns day-to-day operations of production...  ...are seeking an Infrastructure Operations Engineer to operate and improve the large-scale...  ...-availability data center, cloud, AI, or HPC networks and can move comfortably from... 
    Permanent employment

    OpenAI

    San Francisco, CA
    1 day ago
  • $109k - $186k

    To achieve our mission, we must ensure our network deployments are designed for performance and scale, as well as executed flawlessly. As a Deployment Engineer, you will be responsible for translating customer requirements to fully designed wired and wireless networks... 

    Meter

    San Francisco, CA
    4 days ago
  • $50 - $70 per hour

     ...along your resume to David Zaragoza - [email protected] Network Engineer Location: San Francisco, California (Onsite) Duration:...  ...Exposure to research or High Performance Computing (HPC) environments. Experience with automation/Infrastructure-as... 
    Hourly pay
    Work experience placement
    Local area
    Remote work

    Apex Systems

    San Francisco, CA
    1 day ago
  • $190k - $280k

     ...Senior Network Engineer San Francisco About the Role Together AI is looking for a Senior Network Engineer to design, deploy, and operate...  ...or InfiniBand fabrics. Experience supporting GPU clusters, HPC environments, distributed storage, or other high-bandwidth and... 
    Full time

    Together AI

    San Francisco, CA
    4 days ago
  • $157.9k - $213.6k

    As System Engineering Lead, you will drive the technical direction for our next-generation wired networking hardware. You will work with our Product team to align business requirements and industrial design with cost, manufacturability, and technical feasibility constraints... 
    Contract work
    Local area
    Work from home
    Flexible hours

    Amazon

    San Francisco, CA
    15 hours ago
  • $135k - $158k

     ...that’s setting the pace for responsible, transformative cloud infrastructure. About This Role: As a Software Engineer II - Software Defined Networking, you will lead the development and execution of our Software Defined Networking strategy. You will work... 
    Full time
    Temporary work

    Crusoe

    San Francisco, CA
    19 hours ago
  • $155k - $183k

     ...setting the pace for responsible, transformative cloud infrastructure. About This Role: As a Senior Software Engineer I - Software Defined Networking, you will lead the development and execution of our Software Defined Networking strategy. You will work... 
    Full time
    Temporary work

    Crusoe

    San Francisco, CA
    19 hours ago
  •  ...About the Team The Platform Networking team is responsible for the collective communication stack used in our largest training jobs....  ...into our training platform.  About the Role As a Software Engineer, Networking you will design and implement custom networking collectives... 
    Full time
    Work at office
    Relocation package

    OpenAI

    San Francisco, CA
    19 hours ago
  •  ...About the Team We’re hiring software engineers to make OpenAI’s networking teams more productive. These teams build and operate the high-performance networking systems that support OpenAI’s training and inference infrastructure at frontier scale. About the Role... 
    Full time

    OpenAI

    San Francisco, CA
    19 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to HPC Engineer (Biohub Network). Be the first to apply!