SRE Monitoring Platform Software Engineer (Entry Level) [Remote]
Bitdeer Technologies Group
- Remote job
Bitdeer is a world-leading technology company for AI and Bitcoin mining infrastructure.
Bitdeer is committed to providing comprehensive Bitcoin mining solutions for its customers and building AI computational infrastructure to support the AI revolution. Bitdeer handles complex processes involved in computing such as equipment procurement, transport logistics, data center design and construction, equipment management, and daily operations. Bitdeer also offers advanced cloud capabilities to customers with high demand for artificial intelligence.
Headquartered in Singapore, Bitdeer has deployed data centers across multiple countries, including the United States, Norway, Bhutan, and Ethiopia.
To learn more, visit (Position Overview
Bitdeer is building an AI-operated GPU cloud — a global fleet of self-built and OEM-rented data centers running the world's most valuable compute, run by a platform that observes, protects, and operates the fleet. The SRE Platform team builds the monitoring and automation substrate that every other squad — storage, network, GPU, K8S, and L1 operators — depends on. Their signals become the system you help build.
As an entry-level Software Engineer on the SRE / Monitoring Platform team, you contribute to the NeoCloud SRE platform — the multi-region system that observes, protects, and operates a GPU rental fleet across self-built and OEM-rented data centers. You join a bounded context led by a senior engineer, take well-scoped components from design to production code that ships through GitOps + the CICD release pipeline, follows the Plugin Framework conventions, meets declared SLOs, and stays drift-free.
This is a build + learn role. You write code, write tests, and operate what you build under the guidance of a senior engineer. You participate in on-call as a shadow before taking primary. Within 12 months, you should be delivering components independently within your assigned area and growing toward owning a sub-context.
Key Responsibilities
Where you'll contribute (guided by a senior engineer)
- Collection + Storage — help build collection-agent, metrics-store / logs-store / traces-store / profiles-store, enrichment-service, collection-monitor. Write ingestion, query, and storage-path code.
- Alert + Correlation + SLO — contribute to alert-engine-framework, alert-correlation, slo-framework; implement and tune default alert rules.
- Topology + Cluster-Health — contribute to topology-service, cluster-health-rollup, OSS-SRE-tool collection plugins for K8s / Slurm / Ray / Volcano / Kueue / KubeRay.
- Remediation + Workflow + Jobs — help build remediation-actuator, orchestration / workflow components, inspection probes, job-scheduler.
- Observability instrumentation — instrument services with metrics, logs, and traces via OpenTelemetry; build dashboards; write runbooks an on-call can follow.
- Test discipline — write unit / integration / contract tests for everything you ship; participate in chaos and soak tests led by senior engineers.
Why this is a great first role
- Greenfield with a well-defined vision. The Plugin Framework, GitOps pipeline, and SLO framework are decided; you build components inside them with a clear blueprint — not from a blank page.
- You learn the full observability stack at production scale — ingest, query, storage — by building it, not just using it.
- Mentorship-heavy. You work directly with senior and principal engineers who own the architecture; their expertise becomes your growth path.
Job Requirements
- 0-2 years of software engineering experience (new graduates with strong projects or internships welcome).
- Solid fundamentals in one programming language — Go (preferred), Python, Java, or Rust. You can write clean, tested, readable code and explain your design choices.
- CS fundamentals — data structures, algorithms, concurrency, basic networking (TCP / and operating-system concepts (processes, threads, I/O). You can reason about correctness and performance.
- Distributed systems basics — you understand the ideas behind idempotency, retries, back-pressure, caching, and eventual consistency, even if you haven't operated them at scale yet. Eagerness to go deep.
- Monitoring / observability exposure — some hands-on with Prometheus, Grafana, Loki, or similar; can write a basic PromQL query and instrument a service. Eagerness to learn the ingest, query, and storage path of a real observability stack.
- Familiarity with Linux and the shell; comfort reading system logs and using standard debugging tools.
- Kubernetes basics — understand Pods, Services, Deployments; have run something on K8s (a project, lab, or internship).
- Git + CI basics — branching, pull requests, and have used a CI pipeline (GitHub Actions, GitLab CI, or similar).
- Test discipline — you write unit and integration tests as a habit, not an afterthought.
- Communication — clear written and verbal English; can write a good PR description and ask good questions.
- Curiosity and a learning mindset — the most important qualifier. You're excited to learn GPU / AI infrastructure, AIOps, distributed systems, and observability at production scale.
Nice-to-Haves
- Internship or project in monitoring / observability, telemetry pipelines, or platform / SRE tooling.
- Exposure to GPU / AI-infra — DCGM, InfiniBand / RoCE, Kubernetes GPU Operator, Slurm / Ray. Interest counts more than depth.
- Exposure to AIOps / ML-adjacent tooling (anomaly detection, alert correlation).
- Contributions to open-source observability or cloud-native projects.
--------------------------------------------------------------------
Bitdeer is committed to providing equal employment opportunities in accordance with country, state, and local laws. Bitdeer does not discriminate against employees or applicants based on conditions such as race, color, gender identity and/or expression, sexual orientation, marital and/or parental status, religion, political opinion, nationality, ethnic background or social origin, social status, disability, age, indigenous status, and union.
$60 - $80 per hour
...Job Title: Senior Site Reliability Engineer (SRE) - Hybrid Duration (Contract)... ...enterprise-scale applications and platforms. This role combines software engineering, infrastructure... ...automation coverage across deployment, monitoring, alerting, and self-healing workflows...SoftwareHourly payContract work- ...looking for a skilled engineer with disciplines... ...aspects of software systems engineering... ...do: • Evangelize SRE mindset and solve... ...alerting to improve platform reliability. • Lead... ...across deployment, monitoring, alerting, and self... ...experience with enterprise-level administration and...Software
- ...knowledgeable Technical Support Engineer with a strong foundation in Site Reliability Engineering (SRE). The ideal candidate will be... ...Utilize SRE methodologies to monitor, troubleshoot, and resolve incidents... .... Familiarity with cloud platforms and services (e.g., AWS, Azure...SuggestedFull time
- ...direct company sponsorship, entry of GM as the immigration... ...Role We are seeking entry-level Software Engineers to join teams across GM’s... ...of reliable, scalable data platforms and cloud services that enable... ...objectives and indicators, monitoring, alerting, troubleshooting,...SoftwareEntry levelFull timeInternshipLocal areaWork from homeFlexible hours
- ...We are currently looking for a Senior SRE / Technical Lead for an onsite role in... ...Ansible Strong experience with Splunk (monitoring, logging, observability) Proven... ...system improvements Work closely with engineering teams to improve system stability This...SuggestedRemote work
- ...JOB SUMMARY The Senior Platform Engineer / Site Reliability Engineer (SRE) will be responsible for designing... ...Kafka, Python or Ansible automation, monitoring platforms, and enterprise-scale... ...and compliance checks into software delivery workflows. • Implement...Software
- ...Job Title: Ai Platform Engineer Duration: 6 Months (possibility of conversion... ...support of AI platforms. Monitor platform health,... ...Define support models, service level objectives, and operational... ...analytics. Knowledge of software deployment and enterprise application...Software
$104.5k - $234.6k
...Description Leads platform projects crossing multiple... ...Platform Software Development: Lead cross... ...guidance and coaching to engineers to drive improvements.... ...software error logging, monitoring, and observability for... ...remains posted. Career Level - IC4 About Us...SoftwareTemporary workFlexible hoursShift work- ...Senior Platform Engineer NODA is a veteran-owned, venture-backed technology... ...and deploy mission-critical software across diverse environments.... ...Design observability and monitoring systems for distributed... ...infrastructure automation ~ Expert-level proficiency with Docker...SoftwarePermanent employmentFor contractorsLocal area
$145k - $260k
...Staff Platform Engineer Raleigh-Cary, NC, Austin, Dallas, TX, Tampa, FL,... ...systems can act with expert-level judgement at enterprise scale... ...assurance teams to streamline the software delivery process and ensure high-quality releases. Monitor, troubleshoot, and optimize...SoftwareWork experience placementFlexible hours- ...Staff Site Reliability Engineer / Cloud SME... ...Summary As the Staff SRE/Cloud SME, you will be... ...across multiple cloud platforms (Azure and AWS) and container... ...implement comprehensive monitoring, logging, and alerting... ...throughout the software development lifecycle,...SoftwareLong term contractRemote work
$155k - $175k
...technology consulting and software development company... ...Title: Senior DevOps Engineer Location: 100%... ...infrastructure, CI/CD platforms, and deployment pipelines... ...automating deployments, monitoring production... ...reliability engineering (SRE), chaos engineering, blue...SoftwareFull timeH1bLocal areaImmediate startRemote workVisa sponsorship- ...consistent data definitions ~ Experience with database performance monitoring, benchmarking, tuning, security, backup, recovery, and... ...and maintain test database environments Support database software installation, configuration, and implementation Provide...SoftwareContract workLocal area
- ...Job Title:- Systems Engineer - Senior / AI Platform Engineer Location:- Austin... ...Enterprise Platform Monitor AI Platform (health, availability... ...platforms. Sr. Level experience is key Enhancements... .... Knowledge of software deployment and enterprise...SoftwareContract workRemote work
- ...experienced Senior or Staff Engineer for our SRE, InfraSec team, to... ...that reinforce the platform’s security posture.... ...CSPM) Automation and Monitoring: - Build automated solutions... ..., including low-level fundamentals, and how... ...industries with software. MongoDB’s unified data...SoftwareFull timeRemote workWorldwide
- ...York City, Take-Two Interactive Software, Inc. is a leading developer,... ...experiences, delivered on every platform relevant to our audience through... ...ChallengeWe are seeking a mid-level Systems Engineer II with a focus on platform and monitoring systems to support and optimize...SoftwareFull timeCasual workFlexible hours
$123.2k - $169.4k
Messaging and Streaming Platform Engineer The ability to drive human progress through technology... ...performance of the following tasks: Software installation, patch installation, upgrades... ..., configuration, security, system monitoring and tuning, disaster recovery planning...Software- ...position serves as the engine of our Acquisitions and... ...reporting packages, and monitoring key metrics: revenue,... ...imports into valuation software and maintaining accurate... ...Technical Mastery: Expert-level Excel proficiency (... ...more than a standard entry-level role; it is a comprehensive...SoftwareEntry levelInternshipWork at office
$136.41k - $168.5k
...passionate about DevOps and Platform Engineering in the age of AI/ML?!... ...implementations, software deployment projects,... ...teams at all levels and play a tactical role... ...You will build monitoring solutions that improve... .... ~ Knowledge of SRE concepts around application...SoftwareWork at officeLocal areaRemote workFlexible hoursShift work- ...Enter codes to create production databases and utility programs to monitor database performance, including distribution of records and... ...required to save, retrieve, and recover databases from hardware and software failures within established procedures. Assist with...Software
$93k - $189k
...posting)The z/OS Platform CICS, MQ & zCEE Lead... ...include z/OS engineering, WLM (Workload Manager... ...skills, and executive level communication in... ...of all system z software zCEE, CICS and MQ.... ...RACF)Proficient in SRE concepts and... ...continuous improvement of monitoring, performance,...SoftwareFull timeWork at officeRemote workWork from homeFlexible hours- ...Selects and enters codes of utility program to monitor database performance, such as... ...and recover databases from hardware and software failures within established procedures. Assists... ...skills 5 Required Extremely high level of analytical ability where problems are...Software
$114.6k - $234.6k
...OCI depends on hardware platforms that are both innovative... ...Lead integration of rack-level hardware platforms for data... ...deployment Partner with software/data teams to support monitoring and analysis Identify... ...Collaborate across OHD, engineering, operations, supply chain...SoftwareTemporary workFlexible hours$131.6k - $210.3k
...Description We started as an SRE team. Now we're becoming something more — a software engineering organization building agentic... ...frameworks, and self-healing platforms that keep Visa's middleware... ...guidance to others. Use advanced monitoring and observability to detect...SoftwareFull timeWork experience placementWork at officeLocal area$60 - $70 per hour
...Selects and enters codes of utility program to monitor database performance, such as... ...and recover databases from hardware and software failures within established procedures. Assists... ...interpersonal skills 5 Required Extremely high level of analytical ability where problems are...Software- ...Design, develop, implement, maintain, and monitor enterprise database systems. Create... ...development teams. Assist with database software installation, configuration, and... ...responsibilities and priorities. High level of analytical ability to solve complex and...Software
- ...databases, including entering codes and monitoring performance metrics such as record distribution... ...Support the installation of database software and assist with analysis, design, and... ...experience applying an extremely high level of analytical ability to unusual and difficult...SoftwareLocal areaRemote workMonday to FridayWeekend workAfternoon shift
- ...Job Description Job Description SRE Support Engineer - Observability While this position is... ...them build, operate, and scale internal platforms used by tens of thousands of engineers... ...IaaS platform, with a focus on monitoring, alerting, telemetry, and operational...Remote work
$107.6k - $198.4k
...Adobe Experience Platform (AEP), you will help... ...enterprise-level AEP data architectures... ...years with AEP Web Software Development Kit (SDK... ..., or data engineering with application programming... ..., production monitoring, and runbooks5+ years... ...development From entry-level employees to...SoftwareEntry levelLocal areaVisa sponsorship- ...enthusiastic Managed IT Services (MSP) Level 1 Technician to join our... ...of IT (hardware & software installation, maintenance, and... ...updates, backups, and system monitoring to ensure optimal performance... ...tools like NinjaRMM or similar platforms. Experience with Ticketing...SoftwareWork at officeLocal areaRemote workFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to SRE Monitoring Platform Software Engineer (Entry Level) [Remote]. Be the first to apply!
- program monitor Austin, TX
- quality assurance monitor Austin, TX
- monitor tech Austin, TX
- pool monitor Austin, TX
- monitor Austin, TX
- security monitor Austin, TX
- clinical research monitor Austin, TX
- site reliability engineer sre Austin, TX
- site reliability engineer Austin, TX
- site reliability engineer remote Austin, TX




