SRE Monitoring Platform Software Engineer (Entry Level)
Bitdeer Technologies Group
Bitdeer is a world-leading technology company for AI and Bitcoin mining infrastructure.
Bitdeer is committed to providing comprehensive Bitcoin mining solutions for its customers and building AI computational infrastructure to support the AI revolution. Bitdeer handles complex processes involved in computing such as equipment procurement, transport logistics, data center design and construction, equipment management, and daily operations. Bitdeer also offers advanced cloud capabilities to customers with high demand for artificial intelligence.
Headquartered in Singapore, Bitdeer has deployed data centers across multiple countries, including the United States, Norway, Bhutan, and Ethiopia.
To learn more, visit (Position Overview
Bitdeer is building an AI-operated GPU cloud — a global fleet of self-built and OEM-rented data centers running the world's most valuable compute, run by a platform that observes, protects, and operates the fleet. The SRE Platform team builds the monitoring and automation substrate that every other squad — storage, network, GPU, K8S, and L1 operators — depends on. Their signals become the system you help build.
As an entry-level Software Engineer on the SRE / Monitoring Platform team, you contribute to the NeoCloud SRE platform — the multi-region system that observes, protects, and operates a GPU rental fleet across self-built and OEM-rented data centers. You join a bounded context led by a senior engineer, take well-scoped components from design to production code that ships through GitOps + the CICD release pipeline, follows the Plugin Framework conventions, meets declared SLOs, and stays drift-free.
This is a build + learn role. You write code, write tests, and operate what you build under the guidance of a senior engineer. You participate in on-call as a shadow before taking primary. Within 12 months, you should be delivering components independently within your assigned area and growing toward owning a sub-context.
Key Responsibilities
Where you'll contribute (guided by a senior engineer)
- Collection + Storage — help build collection-agent, metrics-store / logs-store / traces-store / profiles-store, enrichment-service, collection-monitor. Write ingestion, query, and storage-path code.
- Alert + Correlation + SLO — contribute to alert-engine-framework, alert-correlation, slo-framework; implement and tune default alert rules.
- Topology + Cluster-Health — contribute to topology-service, cluster-health-rollup, OSS-SRE-tool collection plugins for K8s / Slurm / Ray / Volcano / Kueue / KubeRay.
- Remediation + Workflow + Jobs — help build remediation-actuator, orchestration / workflow components, inspection probes, job-scheduler.
- Observability instrumentation — instrument services with metrics, logs, and traces via OpenTelemetry; build dashboards; write runbooks an on-call can follow.
- Test discipline — write unit / integration / contract tests for everything you ship; participate in chaos and soak tests led by senior engineers.
Why this is a great first role
- Greenfield with a well-defined vision. The Plugin Framework, GitOps pipeline, and SLO framework are decided; you build components inside them with a clear blueprint — not from a blank page.
- You learn the full observability stack at production scale — ingest, query, storage — by building it, not just using it.
- Mentorship-heavy. You work directly with senior and principal engineers who own the architecture; their expertise becomes your growth path.
Job Requirements
- 0-2 years of software engineering experience (new graduates with strong projects or internships welcome).
- Solid fundamentals in one programming language — Go (preferred), Python, Java, or Rust. You can write clean, tested, readable code and explain your design choices.
- CS fundamentals — data structures, algorithms, concurrency, basic networking (TCP / and operating-system concepts (processes, threads, I/O). You can reason about correctness and performance.
- Distributed systems basics — you understand the ideas behind idempotency, retries, back-pressure, caching, and eventual consistency, even if you haven't operated them at scale yet. Eagerness to go deep.
- Monitoring / observability exposure — some hands-on with Prometheus, Grafana, Loki, or similar; can write a basic PromQL query and instrument a service. Eagerness to learn the ingest, query, and storage path of a real observability stack.
- Familiarity with Linux and the shell; comfort reading system logs and using standard debugging tools.
- Kubernetes basics — understand Pods, Services, Deployments; have run something on K8s (a project, lab, or internship).
- Git + CI basics — branching, pull requests, and have used a CI pipeline (GitHub Actions, GitLab CI, or similar).
- Test discipline — you write unit and integration tests as a habit, not an afterthought.
- Communication — clear written and verbal English; can write a good PR description and ask good questions.
- Curiosity and a learning mindset — the most important qualifier. You're excited to learn GPU / AI infrastructure, AIOps, distributed systems, and observability at production scale.
Nice-to-Haves
- Internship or project in monitoring / observability, telemetry pipelines, or platform / SRE tooling.
- Exposure to GPU / AI-infra — DCGM, InfiniBand / RoCE, Kubernetes GPU Operator, Slurm / Ray. Interest counts more than depth.
- Exposure to AIOps / ML-adjacent tooling (anomaly detection, alert correlation).
- Contributions to open-source observability or cloud-native projects.
--------------------------------------------------------------------
Bitdeer is committed to providing equal employment opportunities in accordance with country, state, and local laws. Bitdeer does not discriminate against employees or applicants based on conditions such as race, color, gender identity and/or expression, sexual orientation, marital and/or parental status, religion, political opinion, nationality, ethnic background or social origin, social status, disability, age, indigenous status, and union.
$130k - $180k
...seeking a Cloud Site Reliability Engineer (SRE)to work in our Arlington, VA... ...itself. That means writing software and automation that lets... ...systems across our federal cloud platform (AWS GovCloud, IL5 zero-... ...& Incidents Build monitoring, logging, alerting, and tracing...SoftwareFor contractorsWork at officeRemote workShift work$106.5k - $177.5k
...The Site Reliability Engineering discipline at Noctua Technology... ...treat operations as a software engineering challenge,... ...as Code (IaC), monitor it through advanced observability... ...Reliability Engineer (SRE) to join our dynamic... ...and monitoring Service Level Objectives (SLOs), and...SoftwareRemote work$1,000 per month
...financial wellness platform designed to help... ...infrastructure our engineers use every day. When... ...a Senior DevOps / SRE Engineer on this... ...and shipped real software, not only infrastructure... ...observability/monitoring (Datadog, OpenTelemetry... ...PTOYour actual level and base salary...SoftwareTemporary workWork at officeImmediate startRemote workFlexible hours- ...Senior Application Support Engineer, you will help power... ...Trade Processing (ITP) platforms that support cross-... ...Reliability Engineering (SRE) principles, you will support... ...schedules and service-level commitments.Change,... ...operational risk.Enhance monitoring, alerting,...SuggestedRemote workFlexible hours
- ...parameters for hardware/software compatibility.... ...design. Performs engineering studies and... ...functions. Education Level: Bachelor's Degree... ...dynamic AWS Cloud Platform Engineer to provide... ...role Experience with SRE principles and... ...a plus, Platform Monitoring, Observability, &...SoftwareImmediate startRemote work
$207k - $300k
...design consulting, developing software platforms and frameworks, capacity... ...they are live by measuring and monitoring availability, latency and... ...reliability strategy for Home SRE.Minimum qualifications:Bachelor... ...in Computer Science or Engineering, or a related field.Experience...SoftwareWorldwide$207k - $300k
...consulting, developing software platforms and frameworks,... ...live by measuring and monitoring availability, latency... ...Computer Science or Engineering.Experience mentoring... ...Reliability Engineering (SRE) combines software and... ...meet stringent Service Level Objectives. You will...Software$160k - $210k
...category-leading enterprise software that unleashes that... ...team, we build the platforms and systems the entire... ...This is a software engineering role. You will not be... ..., engineer, and build SRE platform systems and capabilities... ...in livesite monitoring rotations, handle escalations...SoftwareWork at officeImmediate startRemote work- ...category-leading enterprise software that unleashes that... ...team, we build the platforms and systems the entire... ...This is a software engineering role. You will not be... ..., engineer, and build SRE platform systems and capabilities... ...in livesite monitoring rotations, handle escalations...SoftwareWork at officeImmediate startRemote work
$180.5k - $236.91k
.... We're hiring a Senior Software Engineer, Cloud Infrastructure / SRE to join our Engineering... ...a full stack technology platform and a relentless focus on... ...development of Service-Level Objectives (SLOs) for systems... ...: Proficiency with monitoring using tools like Prometheus...SoftwareFull timeWork at officeRemote work- ...TVRoku is the #1 TV streaming platform in the U.S., Canada, and... ...talented and experienced Senior Software Engineer, MLOps/DevOps, to join the... ...strong background in DevOps/SRE practices, cloud infrastructure... ...evaluation, deployment, and monitoring — on top of a modern, cloud-...SoftwareWork at officeLocal areaRemote workMonday to ThursdayFlexible hours
- ...Technologies develops software for customers in the... ...looking for an experienced SRE Engineer for a long-term... ...once and reuse it across platforms. It helps confirm age,... ...with CI/CD, monitoring, logging and incident... ...English (Intermediate level and higher). Working...SoftwareFull timeRemote workFlexible hours
- ...POSITION Platform Engineer / SRE-DevOps Engineer (Redshift) REQUIRED SKILLS Strong hands-on experience administering... ...concurrency scaling, and data distribution. Experience with monitoring, alerting, logging, and production incident management....
$40 per hour
A technology solutions provider is seeking a remote Junior SRE/DevOps Engineer. The ideal candidate should have foundational knowledge of Site Reliability Engineering (SRE) and Kubernetes. Responsibilities include gaining experience in a DevOps-driven environment. Applicants...Long term contractInternshipRemote work- ...Senior Site Reliability Engineer to help us mature and... ...our multi-cloud SaaS platform. Most of our footprint... ...treats infrastructure like software. You'll have... ...run by hand Build monitoring, alerting, and observability... ...~6+ years in SRE, DevOps, or infrastructure...SoftwareRemote workFlexible hours
- ...DescriptionThe AI Inference Engineer plays a critical... ..., and monitoring system performance... ...orchestration. Ensure software solutions are optimized... ...against service-level agreements (SLAs).... ..., and cloud platforms such as AWS, GCP,... ...Background in MLOps or SRE roles focused on...SoftwareFull timeLocal areaImmediate start
- ...DevOps Engineer Key Responsibilities Technical Skills • Strong hands-on experience... ...Azure DevOps, etc.) • Strong experience in monitoring & observability tools (CloudWatch,... ...Linux system administration DevOps / SRE Expertise • Strong understanding of DevOps...Remote work
- ...portfolio of data management platforms and mobile offerings in... ...Informatics platform. As a Software Engineer, you must possess world class... ...Site Reliability Engineering (SRE) concepts and practices is a... ...cloud is a plus. Experience monitoring infrastructure with monitoring...SoftwareFull timeTemporary workWork experience placementLive inRemote workFlexible hours
- ...writing, shipping, and running software effortless and efficient. We work closely with product-engineering to identify friction in the... ...Support, and guide engineers on SRE related topics Partner... ...~ Experience with monitoring / alerting (primarily with Prometheus...SoftwareLocal areaRemote work
$185k - $200k
...JobsSenior Site Reliability Engineer (SRE) Dayton, OH (Remote)... ...AFRL's Google Cloud Platform environment. This... .... This is a senior-level position where deep,... ...repository management, monitoring, and platform operations... ...engineering, DevOps, software engineering, or a related...SoftwareFull timeRemote work- ...Senior Kubernetes-focused SRE SRE Ropes assessment required... ...strong cloud automation and software engineering skills who can leverage AI/... ...operations and improve platform reliability at scale. Key... ...Develop observability and monitoring capabilities using tools such...SoftwareImmediate startRemote work
$100k - $130k
...We are looking for a Platform Integration Engineer to join our engineering... ...with DevOps/SRE to define deployment... ...Qualifications ~3+ years of software engineering... ...management and data quality monitoring ~ Working experience... ...-Remote Position Level Associate Country...SoftwareWork experience placementLocal areaRemote workFlexible hours- ...Job title : Platform Engineer / SRE-DevOps Engineer Amazon Redshift Location: Remote Duration: 3+Months Key Responsibilities... ..., database, and data platform deployments. Monitor and optimize Redshift performance, storage, workloads, and...Remote work
- ...closing and title insurance software. A division of Fidelity... ...rounded Site Reliability Engineer (SRE) to join our Cloud Operations... ...in-depth knowledge of the platforms that run our solutions.... ...to refine our service level indicators monitoring capabilities with the goal...SoftwareHourly payWork at officeRemote work
$307k - $427k
...velocity.Ensure that SRE principles (SLOs,... ...into the shipped software so external... ...incident response, monitoring, and debugging paradigms... ...who act as Level 3 support or manage... ...experienced software engineers of large-scale projects... ...of Google platforms, we make Google's...Software$167k - $196.5k
...the data collaboration platform of choice for the... ...platforms.The Global SRE team is responsible for... ...Senior Site Reliability Engineer who is excited about establishing... ...is important (Software Engineer, Site... ...& Product Reliability monitoring and alerting Maintain...SoftwareFull timeWork at officeRemote workWork from homeFlexible hoursNight shift- ...SRE / DevOps Engineer Canada / Remote 6+ Months Contract Position Requirements Collaborate closely with Development teams to improve... ...tools like Sonar. Manage application integration and monitoring using DataDog, SumoLogic, and alert systems like MS Teams....Contract workRemote work
- ...Site Reliability Engineer (SRE) Location: Remote Shift Timings: 5:30 PM to 3:00 AM IST... ...strong background in both log and metrics monitoring stacks, specifically ELK (Elasticsearch... ..., and performance of our customer's platforms and services, bridging the gap between...Remote workShift work
$165k - $225k
...Sr. Site Reliability Engineer (SRE) Chicago, IL or Remote Moonlite... ..., network engineers, and platform engineering team, you'll architect... ...SLIs, SLOs, and monitoring to meet enterprise reliability... ...engineers, network engineers, and software developers. Preferred...SoftwareRemote workFlexible hours$100k - $180k
...Site Reliability Engineer (SRE) - Remote Bright Vision... ...technology consulting and software development company... ...continually pushing the platform toward higher... ...continually refine service-level objectives (SLOs), service... ...comprehensive monitoring, logging, and tracing...SoftwareFull timeH1bLocal areaImmediate startRemote workVisa sponsorship
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to SRE Monitoring Platform Software Engineer (Entry Level). Be the first to apply!


