SRE Monitoring Platform Software Engineer (Early Career / Temporary)
Bitdeer Technologies Group
Bitdeer is a world-leading technology company for AI and Bitcoin mining infrastructure.
Bitdeer is committed to providing comprehensive Bitcoin mining solutions for its customers and building AI computational infrastructure to support the AI revolution. Bitdeer handles complex processes involved in computing such as equipment procurement, transport logistics, data center design and construction, equipment management, and daily operations. Bitdeer also offers advanced cloud capabilities to customers with high demand for artificial intelligence.
Headquartered in Singapore, Bitdeer has deployed data centers across multiple countries, including the United States, Norway, Bhutan, and Ethiopia.
To learn more, visit (Position Overview
Bitdeer is building an AI-operated GPU cloud — a global fleet of self-built and OEM-rented data centers running the world's most valuable compute, run by a platform that observes, protects, and operates the fleet. The SRE Platform team builds the monitoring and automation substrate that every other squad — storage, network, GPU, K8S, and L1 operators — depends on. Their signals become the system you help build.
As an entry-level Software Engineer on the SRE / Monitoring Platform team, you contribute to the NeoCloud SRE platform — the multi-region system that observes, protects, and operates a GPU rental fleet across self-built and OEM-rented data centers. You join a bounded context led by a senior engineer, take well-scoped components from design to production code that ships through GitOps + the CICD release pipeline, follows the Plugin Framework conventions, meets declared SLOs, and stays drift-free.
This is a build + learn role. You write code, write tests, and operate what you build under the guidance of a senior engineer. You participate in on-call as a shadow before taking primary. Within 12 months, you should be delivering components independently within your assigned area and growing toward owning a sub-context.
Key Responsibilities
Where you'll contribute (guided by a senior engineer)
- Collection + Storage — help build collection-agent, metrics-store / logs-store / traces-store / profiles-store, enrichment-service, collection-monitor. Write ingestion, query, and storage-path code.
- Alert + Correlation + SLO — contribute to alert-engine-framework, alert-correlation, slo-framework; implement and tune default alert rules.
- Topology + Cluster-Health — contribute to topology-service, cluster-health-rollup, OSS-SRE-tool collection plugins for K8s / Slurm / Ray / Volcano / Kueue / KubeRay.
- Remediation + Workflow + Jobs — help build remediation-actuator, orchestration / workflow components, inspection probes, job-scheduler.
- Observability instrumentation — instrument services with metrics, logs, and traces via OpenTelemetry; build dashboards; write runbooks an on-call can follow.
- Test discipline — write unit / integration / contract tests for everything you ship; participate in chaos and soak tests led by senior engineers.
Why this is a great first role
- Greenfield with a well-defined vision. The Plugin Framework, GitOps pipeline, and SLO framework are decided; you build components inside them with a clear blueprint — not from a blank page.
- You learn the full observability stack at production scale — ingest, query, storage — by building it, not just using it.
- Mentorship-heavy. You work directly with senior and principal engineers who own the architecture; their expertise becomes your growth path.
Job Requirements
- 0-2 years of software engineering experience (new graduates with strong projects or internships welcome).
- Solid fundamentals in one programming language — Go (preferred), Python, Java, or Rust. You can write clean, tested, readable code and explain your design choices.
- CS fundamentals — data structures, algorithms, concurrency, basic networking (TCP / and operating-system concepts (processes, threads, I/O). You can reason about correctness and performance.
- Distributed systems basics — you understand the ideas behind idempotency, retries, back-pressure, caching, and eventual consistency, even if you haven't operated them at scale yet. Eagerness to go deep.
- Monitoring / observability exposure — some hands-on with Prometheus, Grafana, Loki, or similar; can write a basic PromQL query and instrument a service. Eagerness to learn the ingest, query, and storage path of a real observability stack.
- Familiarity with Linux and the shell; comfort reading system logs and using standard debugging tools.
- Kubernetes basics — understand Pods, Services, Deployments; have run something on K8s (a project, lab, or internship).
- Git + CI basics — branching, pull requests, and have used a CI pipeline (GitHub Actions, GitLab CI, or similar).
- Test discipline — you write unit and integration tests as a habit, not an afterthought.
- Communication — clear written and verbal English; can write a good PR description and ask good questions.
- Curiosity and a learning mindset — the most important qualifier. You're excited to learn GPU / AI infrastructure, AIOps, distributed systems, and observability at production scale.
Nice-to-Haves
- Internship or project in monitoring / observability, telemetry pipelines, or platform / SRE tooling.
- Exposure to GPU / AI-infra — DCGM, InfiniBand / RoCE, Kubernetes GPU Operator, Slurm / Ray. Interest counts more than depth.
- Exposure to AIOps / ML-adjacent tooling (anomaly detection, alert correlation).
- Contributions to open-source observability or cloud-native projects.
--------------------------------------------------------------------
Bitdeer is committed to providing equal employment opportunities in accordance with country, state, and local laws. Bitdeer does not discriminate against employees or applicants based on conditions such as race, color, gender identity and/or expression, sexual orientation, marital and/or parental status, religion, political opinion, nationality, ethnic background or social origin, social status, disability, age, indigenous status, and union.
- ...world's most valuable compute, run by a platform that observes, protects, and operates the fleet. The SRE Platform team builds the monitoring and automation substrate that every... ...you help build. As an entry-level Software Engineer on the SRE / Monitoring Platform team...SoftwareEntry levelFull timeContract workInternship
$83.52k - $125.28k
...We are currently seeking a Lead ML Platform Engineer (SRE / FTE / Onsite) to join our team in Charlotte... ...teams to build, validate, deploy, monitor, and operate predictive models... ...encryption, audit logging, and secure software delivery. ~ Experience implementing...Temporary workSoftwareWork at officeRemote workFlexible hours$105k - $158k
Platform Engineer - Healthcare Cloud PlatformSpecialist... ...that transforms your career—and the lives of... ...Reliability Engineering (SRE) · Establish and monitor Service Level... ...Automate patching, software deployment, inventory... ...matching.Full-time temporary employees receive 6...Temporary workSoftwareFull timePart timeWork at officeLocal areaRemote workFlexible hours- ...Senior DevOps/Platform Engineer (SRE) 100 % Remote (W/M/D) Easybill is the leading provider of cloud-based invoicing software and has over 18 years of successful market presence. Our... ...bring your ideas for optimization. Monitoring and analysis: You improve our monitoring...SoftwareRemote workFlexible hoursNight shift
- ...Job Title: OMS Platform Reliability Lead... ...combines Platform Engineering, Site Reliability Engineering (SRE), and technical leadership... ...bridge between Software Engineering and IT... ...Build advanced monitoring dashboards using tools... ...your career aspirations. As an...Temporary workSoftwareContract workWork experience placementRemote work
- ...skilled Site Reliability Engineer (SRE) to join our... ...an SRE, you’ll blend software engineering with systems... ...improvement across our platform. You’ll be instrumental... ...proactive monitoring, and streamlining incident... ...ensure visibility into temporary workarounds while...Temporary workSoftwareInterim roleRemote workFlexible hours
$1,000 per month
...financial wellness platform designed to help... ...infrastructure our engineers use every day. When... ...a Senior DevOps / SRE Engineer on this... ...and shipped real software, not only infrastructure... ...observability/monitoring (Datadog, OpenTelemetry... ....Experience at an early-stage startup...Temporary workSoftwareWork at officeImmediate startRemote workFlexible hours- ...portfolio of data management platforms and mobile offerings... ...platform. As a Software Engineer, you must possess world... ...Reliability Engineering (SRE) concepts and... ...is a plus. Experience monitoring infrastructure with monitoring... .... It’s discovering a career that’s challenging, supportive...Temporary workSoftwareFull timeWork experience placementLive inRemote workFlexible hours
- Platform Engineer (CI, CD, CT, Ruby, AWS, Azure, Scripting, Monitoring) in VA or NYC AWS, Azure, CD, CI, CT, Docker, Perl, Python, REST API, SQL Location: Virginia... ...initiative and enjoy working with engineers to make the software development process as painless as possible ·...SoftwarePermanent employmentFull timeRemote work
$78k - $185k
...the week.ABOUT THE TEAMThe Platform Engineering team is part of Parametric IT... ...reliability in all phases of the Software Development Lifecycle (SDLC)... ...observability via monitoring tools such as Datadog, Splunk... ...Engineering (i.e., DevOps, SysOps, SRE) experience with a proven...Temporary workSoftwareWork at officeLocal areaRemote workWorldwide3 days per week- ...parameters for hardware/software compatibility. Defines... ...design. Performs engineering studies and analyses,... ...and dynamic AWS Cloud Platform Engineer to provide hands... ...role Experience with SRE principles and transformation... .... is a plus, Platform Monitoring, Observability, &...SoftwareImmediate startRemote work
$87.95k - $162.88k
...Senior AI DevOps Engineer (AI Ops / Platform Engineering) (Hybrid)NTT DATA Services... ...generation of AI-powered software delivery and cloud operations... ..., DevOps, Security, SRE, and AI teams to design secure... ...infrastructure, deployment, monitoring, and operational data to authorized...Temporary workSoftwareWork at officeRemote workFlexible hours3 days per week- ...Jobs > CATI Quality Control Monitor / Telephone Survey Interview... ...appropriately. Proficiency with CATI platforms (e.g. Voxco, Forsta, Blaise,... ...; Quality monitoring software; and Video conferencing platforms... ...have a contingency plan for temporary internet disruptions....Temporary workSoftwareContract workPart timeWork at officeImmediate startRemote workWork from homeHome officeShift workNight shift
$35 - $48 per hour
...by an in-house team of engineers who are just as... ...ownership of our ecommerce platform, working under the guidance... ...to improving monitoring, deployment, and reliability... ...of professional software engineering experience... ...Additional pluses include early technical leadership (...Temporary workSoftwareHourly payFor contractorsCasual workSelf employmentWork at officeRemote workWork from home- ...Senior Platform Engineer We are looking for someone with a strong SaaS... ...a little bit crazy. If bad software keeps you up at night, then... ...database, storage, secrets, and monitoring resources. Build... ...reliability engineering practices (SRE principles, SLAs, SLOs, error...Temporary workSoftwareRemote workNight shift
$1,000 per month
...Demonstrated DevOps/SRE depth and a genuine backend software engineering background with shipped... ...with observability or monitoring, CI/CD, compute, and networking... ...from scratch at an early-stage startup or scaling... ...internal AI developer platforms. Preferred prior SOC...Temporary workSoftwareFull timeWork at officeImmediate startRemote workFlexible hours- ...We are looking for an experienced Software Engineer with a strong background in Platform Engineering to build and scale... ...platform resilience by improving monitoring, logging, and observability using... ...Qualifications:6+ years in DevOps, SRE, or Platform Engineering plus a Bachelor...SoftwareWork at officeWork from homeFlexible hours3 days per week
$5,098.66 - $6,701.75 per month
...encourage you to consider a career with us. Employee... ...opportunities for early career pathways, visit... ...Travel: Up to 5% | Regular/Temporary: Regular | Full Time/... ...revenue projections, monitoring expenditures, and recommending... ...Skills in computer software applications,...Temporary workSoftwareEntry levelFull timeContract workPart timeWork at officeLocal areaRemote workShift work$3,000 - $4,999 per month
Career Opportunities: CPS Heightened Monitoring Specialist IV (19684) Posting ID 19684 -Posted 08/06/2026 - Dept of... ...Travel: Up to 75% Regular/Temporary: Regular Full Time/Part Time: Full... ...documents, Adobe PDFs, webpages, software, training guides, video, and audio...Temporary workSoftwareFull timePart timeWork at officeRemote workFlexible hoursShift work- ...United States As a System Engineer, Sr. on our team, your main role... ...Team (SET) to generate monitoring/observability recommendations... ...Use your hardware and software experience to help strengthen... ...application owners to understand their platform designs and how they operate...SoftwareContract workWork at officeRemote work
$146k - $188k
...recognized “Great Place To Work,” delivers engineering solutions across commercial space and... ...using Government-approved systems, software, and processes. Required Qualifications... ...is required. Ability to travel to temporary work locations up to 10%, as needed....Temporary workSoftwareContract workFor contractorsWork at officeImmediate startRemote work$106.5k - $177.5k
...Software Engineer- Site Reliability Engineering (SRE) DC, MD, VA, CA The Site Reliability Engineering discipline at Noctua Technology, LLC is a strategic... ...; they build it using Infrastructure as Code (IaC), monitor it through advanced observability stacks, and...SoftwareRemote work- ...client seeks an Associate Security Engineer, Cyber Monitoring for a role on the Network Monitoring... ...infrastructure. This is a great opportunity for early‑career cybersecurity professionals looking... ...Technology, Computer Science, Software Engineering, Computer Engineering)...SoftwareEntry levelHourly payLocal areaRemote workEarly shift
$190.9k - $334.1k
...DescriptionIt all started when engineer Fred Luddy wrote code... .... Our ServiceNow AI platform brings together any AI... ...Engineer - SRE & AIOps to drive infrastructure... ...stack, including monitoring platforms, incident management... ...industry.12+ years in software engineering or...SoftwareWork at officeImmediate startRemote workFlexible hours$149.4k - $202k
...Senior Software Engineer- Site Reliability Engineering (SRE) DC, MD, VA, CA The Site Reliability Engineering discipline... ...Infrastructure as Code (IaC), monitor it through advanced observability... ...multiple services or entire platforms, ensuring alignment with business...SoftwareRemote work- ...encourage you to consider a career with us. Employee Benefits DSHS... ...other DSHS opportunities for early career pathways, visit the... ...Functional Title: Contract Monitoring Manager Job Title: Contract... ...Travel: Up to 10% Regular/Temporary: Regular Full Time/Part Time...Temporary workEntry levelFull timeContract workPart timeFor contractorsLive inLocal areaRemote workShift work
- ...customer is seeking a Senior QRadar Platform Engineer to serve as the overall owner of their... ...PRIMARY DUTIES: System health & monitoring : Own the overall health of the QRadar... ...usefulness. Patching & maintenance : Apply software patches and updates to keep the system...Temporary workSoftwareWork at officeImmediate start
- ...Description Koniag IT Systems (KITS) is seeking a Senior Monitoring / SRE Engineer with a minimum of 8 years of experience to lead enterprise... ...candidate has designed and operated enterprise monitoring platforms at scale, has strong incident management experience, and can...Full timeFlexible hours
- ...Our client has an opportunity for a Platform Engineer to design, build, and maintain the AWS... .... Ensure production stability with monitoring, alerting, incident response, and SLA... ...language for automation and tooling. Software engineering fundamentals including Git...Temporary workSoftwareHourly payLocal areaWork from home
$185k - $200k
...JobsSenior Site Reliability Engineer (SRE) Dayton, OH (Remote) full... ...support AFRL's Google Cloud Platform environment. This engineer... ...onboarding, repository management, monitoring, and platform operations.... ...engineering, DevOps, software engineering, or a related discipline...SoftwareFull timeRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to SRE Monitoring Platform Software Engineer (Early Career / Temporary). Be the first to apply!


