Sr GPU Cloud South-North Network SRE Expert (SRE SME)
Bitdeer Technologies Group
About Bitdeer Technologies Group
Bitdeer is a world-leading technology company for AI and Bitcoin mining infrastructure.
Bitdeer is committed to providing comprehensive Bitcoin mining solutions for its customers and building AI computational infrastructure to support the AI revolution. Bitdeer handles complex processes involved in computing such as equipment procurement, transport logistics, data center design and construction, equipment management, and daily operations. Bitdeer also offers advanced cloud capabilities to customers with high demand for artificial intelligence.
Headquartered in Singapore, Bitdeer has deployed data centers across multiple countries, including the United States, Norway, Bhutan, and Ethiopia.
To learn more, visit (Job Description
You own the north-south fabric that connects our AI cloud to the world — and drive the automation that turns EVPN/BGP ops from tickets into policy.
NeoCloud is building an AI-operated GPU cloud spanning 4 US DCs, APAC sites, and Iceland. In this role you own the IP/Ethernet infrastructure that connects that fleet — DCI, WAN, internet edge, tenant-facing networks — and you build the automation that lets the AIOps substrate express network intent as code, not as change tickets.
What you'll own
- EVPN-VxLAN data center fabrics for multi-tenant network isolation and workload mobility.
- Multi-site DCI overlay/underlay connecting 4 US DCs with consistent L2/L3 services.
- BGP (eBGP/iBGP), OSPF, ECMP load balancing, and VxLAN across spine-leaf architectures.
- IP transit, peering relationships, and internet edge infrastructure.
- WAN/IP backbone connecting US DCs to APAC and Iceland sites.
- Network equipment: Arista, Cisco, and Palo Alto switches, routers, and firewalls.
- Network automation with Ansible, Terraform, Nautobot/NetBox for IPAM/DCIM.
- Network health, capacity, and SLA monitoring across all sites.
- Network architecture docs, standard configs, and change procedures.
Feed the AIOps substrate
- Every route policy, ACL, VRF, and peering session lives as code in the platform's config-management substrate — no config drift, no snowflakes.
- Every BGP flap, every DCI degradation feeds the network-stability predictor.
- Every routine change becomes a workflow the remediation actuator can execute under a maintenance window.
Job Requirement:
- 5+ years in data center network engineering, with hands-on EVPN-VxLAN deployment experience
- Strong BGP expertise (eBGP/iBGP, route policies, communities, traffic engineering)
- Experience designing and operating multi-site DCI with EVPN multi-homing
- Proficiency with Arista EOS and/or Cisco NX-OS in spine-leaf data center environments
- Experience with firewall platforms (Palo Alto, Fortinet) for perimeter and inter-tenant security
- Hands-on experience with IP transit provisioning and internet peering
- Familiarity with network automation tools (Ansible, Python/Netmiko, Nautobot/NetBox)
- Strong understanding of QoS, traffic engineering, and capacity planning for DC networks
- AIOps aptitude — you've either driven a network-automation program from ticket-driven ops to intent-driven policy, or you have a clear thesis on how you would.
- Runbook-as-code mindset — configs, changes, and rollbacks should be executable by an agent.
--------------------------------------------------------------------
Bitdeer is committed to providing equal employment opportunities in accordance with country, state, and local laws. Bitdeer does not discriminate against employees or applicants based on conditions such as race, color, gender identity and/or expression, sexual orientation, marital and/or parental status, religion, political opinion, nationality, ethnic background or social origin, social status, disability, age, indigenous status, and union.
- .... Bitdeer also offers advanced cloud capabilities to customers with... ...NeoCloud is building an AI-operated GPU cloud. Kubernetes is where all... ..., NVLink domain awareness, network rail affinity. Custom Resource... ...(ArgoCD/Flux) ~ Strong SRE background: SLI/SLO frameworks,...CloudSeniorNetworkFull timeLocal area
- ...operations. Bitdeer also offers advanced cloud capabilities to customers with high demand... .... Bitdeer is building an AI-operated GPU cloud. Storage is where AI workloads either... ...data paths. Deploy and manage storage networking (NFS over RDMA, NVMe-oF, high-speed...CloudSeniorNetworkRemote jobFull timeLocal area
$165k - $225k
...bare-metal performance with cloud-native operational... ...with our systems engineers, network engineers, and platform engineering... ...networking solutions with SR-IOV for high-performance GPU interconnects, multi-tenancy... ...Experience: 5+ years in SRE, DevOps, or infrastructure...CloudSeniorNetworkRemote workFlexible hours- ...traditional infrastructure-only SRE role then this is the... ...scale, and operate the cloud and Kubernetes... ...data scientists, or LLM experts. instead your day to day... ...compute, memory, storage, network, and service dependencies... ...model gateways, AI APIs, GPU-enabled workloads, or distributed...CloudSeniorNetworkPermanent employmentRemote work
$364k
...Google’s technical Infrastructure and Google Cloud products.Architect observability... ..., Google’s Site Reliability Engineering (SRE) is an engineering discipline for building... ...apart so we can rebuild them. We keep our networks up and running, ensuring our users have the...CloudSeniorNetwork- ...Job Title: Senior Principal SRE Architect Location: Dallas or St Louis or Remote Client: Mastercard VISA:... ...provide guidance across front-end, backend, APIs, databases, and network layers. Cloud Platforms & Infrastructure: Hands-on experience with major...CloudSeniorNetworkRemote work
$180.5k - $236.91k
...re hiring a Senior Software Engineer, Cloud Infrastructure / SRE to join our Engineering team.... ...Responsibilities: Become the expert on your team's business and technical... ..., Grafana, or similar. Security & Networking: Knowledge of cloud-native security...CloudSeniorNetworkRemote jobFull timeWork at office- ...scale the infrastructure behind our multi-cloud SaaS platform. Most of our footprint runs... ...subscription and management group hierarchy, and network topology Redesign and enforce least-... ...What You Bring: ~6+ years in SRE, DevOps, or infrastructure engineering roles...CloudSeniorNetworkRemote workFlexible hours
$119.8k - $234.7k
...EngineeringCompany: MicrosoftOverviewMicrosoft Silicon, Cloud Hardware, and Infrastructure Engineering... ...organization. As a Senior Linux SRE (Site Reliability Engineer), your main... ...field (preferred) Solid understanding of networking, systems administration, automation, and...CloudSeniorNetworkOngoing contractPermanent employmentWork at officeLocal areaWorldwide3 days per week$150k - $160k
...brands use our award-winning Conversational Cloud platform to connect with millions of... ...OverviewAs the Senior Cloud Architect (Network & SRE), you will serve as a senior technical lead... ...GCP. This role is ideal for a US-based expert with a deep background in both...CloudSeniorNetworkPermanent employmentTemporary workLocal areaRemote workFlexible hours$150k - $160k
...brands use our award-winning Conversational Cloud platform to connect with millions of... ...Overview As the Senior Cloud Architect (Network & SRE) , you will serve as a senior technical... ...GCP. This role is ideal for a US-based expert with a deep background in both networking...CloudSeniorNetworkPermanent employmentFull timeRemote work- ...scale. What You’ll Do: Cloud Gateways is a complex, multi-cloud... ..., primarily supporting our North America customer base, and be... ...services, API gateways, or similar network infrastructure is highly... ...Contributions to open-source SRE tools or projects. Relevant...CloudSeniorNetworkFull timeShift work
$167k - $196.5k
...between brands, and across its premier global network of top-quality partners. Hundreds of... ...brands and technology platforms.The Global SRE team is responsible for owning and... ...platforms like Kubernetes, Containers and public clouds (GCP or AWS)Experience with deployment...CloudSeniorNetworkFull timeWork at officeRemote workWork from homeFlexible hoursNight shift$83.52k - $125.28k
...ML Platform Engineer (SRE / FTE / Onsite) to join... ...our team in Charlotte, North Carolina (US-NC), United... ...Predictive AI Platform across cloud and on-premises... ...connectivity, identity, network controls, compute, storage... ...Top Employer, we have experts in more than 50...CloudNetworkFull timeTemporary workWork at officeRemote workFlexible hours- ...Citizens Bank is seeking a Dynatrace Subject Matter Expert (SME) to lead a team of Dynatrace engineers across... ..., alerting, and governance in hybrid cloud and on-prem setups. You will partner with development, DevOps, SRE, and infrastructure teams to deliver actionable...CloudRemote job
$109.05 per hour
...Lead AWS Cloud Platform Engineer (IaC / Automation / SRE) Type: Contract (approx. 6 months) Schedule: Mon–Fri... ...environments Troubleshoot AWS networking/connectivity, security, and performance... ...and workforce solutions across North America. We help clients get work...CloudNetworkContract workTemporary workLocal areaRemote work1 day per week- ...SoftPro's Headquarters is in Raleigh, North Carolina. SoftPro has received national... ...-rounded Site Reliability Engineer (SRE) to join our Cloud Operations Team in our Raleigh, NC... ...of the OSI model, software defined networks and public cloud network infrastructure...CloudNetworkHourly payWork at officeRemote work
- ...Administration role. Prior experience building and supporting cloud-based solutions . Experience in a production environment... ...server platforms, Java application platforms, operating systems, network components, virtualization technologies, database platforms)....CloudSeniorNetworkFull time
$148k - $249k
...to monitor the health and performance of cloud and on-prem environments.- Develop and extend... ...and benchmarks (compute, storage, network, ML/AI) and integrate stress, chaos, and regression... ...databases/streaming/batch/ML platforms; GPU/xPU or Arm performance exposure.-...CloudSeniorNetworkFull timeWork at officeWork from homeFlexible hours$207k - $300k
...Experience with Large Language Model.Site Reliability Engineering (SRE) combines software and systems engineering to build and run... ...warranties by taking things apart so we can rebuild them. We keep our networks up and running, ensuring our users have the best and fastest...CloudNetwork- ...infrastructure. The ideal candidate has a strong background in DevOps/SRE practices, cloud infrastructure management, and MLOps tooling — with a passion... ...for ML workloads across AWS and GCP, including GPU/TPU-based training and inference environmentsArchitect and improve...CloudSeniorWork at officeLocal areaRemote workMonday to ThursdayFlexible hours
- ...Machine Learning Engineer/SRE Location: Chicago, IL or Remote Duration: 12 Months... ...including virtual machines, storage, and networking, to support AI model development and... ...deployment processes. Containerization in the Cloud: Strong knowledge of containerization...CloudNetworkRemote work
$120k - $170k
...vertically integrated AI cloud engineered for AI. We... ...—energy, data centres, GPU superclusters, orchestration... ...hardware, east‑west networking, Linux, and data centre... ...traffic differs from north-south. Observability and incident... ..., and backout plans. SRE-style operations. Write...CloudSeniorNetworkRemote workFlexible hours$207k - $300k
...roadmap, and reliability strategy for Home SRE.Minimum qualifications:Bachelor’s degree... ...architecting complex client-to-cloud telemetry systems, CUJ monitoring frameworks... ...apart so we can rebuild them. We keep our networks up and running, ensuring our users have the...CloudNetworkWorldwide$112k - $150k
...AI, Analytics & RPA Support SRE Manager Mizuho AI and Analytics Center Of Acceleration... ..., problems, and changes. Azure Cloud Management: Oversee the deployment and management... ...computing, DevOps, software-defined networking, and Infrastructure-as-a-Service (IaaS)....CloudNetworkWork at officeLocal areaRemote workWorldwide$207k - $300k
...stakeholders.Site Reliability Engineering (SRE) combines software and systems... ...tolerant systems. SRE ensures that Google Cloud's services—both our internally critical and... ...apart so we can rebuild them. We keep our networks up and running, ensuring our users have the...CloudNetwork$207k - $300k
...roadmap, and reliability strategy for AQI Data SRE across the SAGE/SPARK production stack.... ...systems. SRE ensures that Google Cloud's services—both our internally critical and... ...apart so we can rebuild them. We keep our networks up and running, ensuring our users have the...CloudNetwork$160k - $185k
...along the way. And we’re just getting started!OverviewThe Sr. Manager, Site Reliability Engineering (SRE) leads the strategy, execution, and continuous... ...with platform engineering to build scalable, resilient cloud-native architectures.Drive adoption of CI/CD pipelines...CloudSeniorWork at officeLocal areaRemote workWork from home$207k - $300k
...experience.8 years of technical experience in network protocols (TCP/IP, BGP, DNS, routing,... ...engineering VPs, Directors, and enterprise Cloud customers.Extensive background in... ...intelligence pipelines.Site Reliability Engineering (SRE) combines software and systems engineering...CloudNetworkShift work- DescriptionJob Description SummaryThe Digital Site Reliability Engineer (SRE) - GCP Cloud Adoption Engineer is responsible for facilitating the... ...to Office Workplace TypeCertain positions outside our branch network may be eligible for a flexible work arrangement. We’re...CloudNetworkFull timeH1bWork at officeRemote workWork from homeFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Sr GPU Cloud South-North Network SRE Expert (SRE SME). Be the first to apply!




