SRE Monitoring Platform Software Engineer (Entry Level)
Bitdeer Technologies Group
Bitdeer is a world-leading technology company for AI and Bitcoin mining infrastructure.
Bitdeer is committed to providing comprehensive Bitcoin mining solutions for its customers and building AI computational infrastructure to support the AI revolution. Bitdeer handles complex processes involved in computing such as equipment procurement, transport logistics, data center design and construction, equipment management, and daily operations. Bitdeer also offers advanced cloud capabilities to customers with high demand for artificial intelligence.
Headquartered in Singapore, Bitdeer has deployed data centers across multiple countries, including the United States, Norway, Bhutan, and Ethiopia.
To learn more, visit (Position Overview
Bitdeer is building an AI-operated GPU cloud — a global fleet of self-built and OEM-rented data centers running the world's most valuable compute, run by a platform that observes, protects, and operates the fleet. The SRE Platform team builds the monitoring and automation substrate that every other squad — storage, network, GPU, K8S, and L1 operators — depends on. Their signals become the system you help build.
As an entry-level Software Engineer on the SRE / Monitoring Platform team, you contribute to the NeoCloud SRE platform — the multi-region system that observes, protects, and operates a GPU rental fleet across self-built and OEM-rented data centers. You join a bounded context led by a senior engineer, take well-scoped components from design to production code that ships through GitOps + the CICD release pipeline, follows the Plugin Framework conventions, meets declared SLOs, and stays drift-free.
This is a build + learn role. You write code, write tests, and operate what you build under the guidance of a senior engineer. You participate in on-call as a shadow before taking primary. Within 12 months, you should be delivering components independently within your assigned area and growing toward owning a sub-context.
Key Responsibilities
Where you'll contribute (guided by a senior engineer)
- Collection + Storage — help build collection-agent, metrics-store / logs-store / traces-store / profiles-store, enrichment-service, collection-monitor. Write ingestion, query, and storage-path code.
- Alert + Correlation + SLO — contribute to alert-engine-framework, alert-correlation, slo-framework; implement and tune default alert rules.
- Topology + Cluster-Health — contribute to topology-service, cluster-health-rollup, OSS-SRE-tool collection plugins for K8s / Slurm / Ray / Volcano / Kueue / KubeRay.
- Remediation + Workflow + Jobs — help build remediation-actuator, orchestration / workflow components, inspection probes, job-scheduler.
- Observability instrumentation — instrument services with metrics, logs, and traces via OpenTelemetry; build dashboards; write runbooks an on-call can follow.
- Test discipline — write unit / integration / contract tests for everything you ship; participate in chaos and soak tests led by senior engineers.
Why this is a great first role
- Greenfield with a well-defined vision. The Plugin Framework, GitOps pipeline, and SLO framework are decided; you build components inside them with a clear blueprint — not from a blank page.
- You learn the full observability stack at production scale — ingest, query, storage — by building it, not just using it.
- Mentorship-heavy. You work directly with senior and principal engineers who own the architecture; their expertise becomes your growth path.
Job Requirements
- 0-2 years of software engineering experience (new graduates with strong projects or internships welcome).
- Solid fundamentals in one programming language — Go (preferred), Python, Java, or Rust. You can write clean, tested, readable code and explain your design choices.
- CS fundamentals — data structures, algorithms, concurrency, basic networking (TCP / and operating-system concepts (processes, threads, I/O). You can reason about correctness and performance.
- Distributed systems basics — you understand the ideas behind idempotency, retries, back-pressure, caching, and eventual consistency, even if you haven't operated them at scale yet. Eagerness to go deep.
- Monitoring / observability exposure — some hands-on with Prometheus, Grafana, Loki, or similar; can write a basic PromQL query and instrument a service. Eagerness to learn the ingest, query, and storage path of a real observability stack.
- Familiarity with Linux and the shell; comfort reading system logs and using standard debugging tools.
- Kubernetes basics — understand Pods, Services, Deployments; have run something on K8s (a project, lab, or internship).
- Git + CI basics — branching, pull requests, and have used a CI pipeline (GitHub Actions, GitLab CI, or similar).
- Test discipline — you write unit and integration tests as a habit, not an afterthought.
- Communication — clear written and verbal English; can write a good PR description and ask good questions.
- Curiosity and a learning mindset — the most important qualifier. You're excited to learn GPU / AI infrastructure, AIOps, distributed systems, and observability at production scale.
Nice-to-Haves
- Internship or project in monitoring / observability, telemetry pipelines, or platform / SRE tooling.
- Exposure to GPU / AI-infra — DCGM, InfiniBand / RoCE, Kubernetes GPU Operator, Slurm / Ray. Interest counts more than depth.
- Exposure to AIOps / ML-adjacent tooling (anomaly detection, alert correlation).
- Contributions to open-source observability or cloud-native projects.
--------------------------------------------------------------------
Bitdeer is committed to providing equal employment opportunities in accordance with country, state, and local laws. Bitdeer does not discriminate against employees or applicants based on conditions such as race, color, gender identity and/or expression, sexual orientation, marital and/or parental status, religion, political opinion, nationality, ethnic background or social origin, social status, disability, age, indigenous status, and union.
- ...looking for a skilled engineer with disciplines... ...aspects of software systems engineering... ...do: • Evangelize SRE mindset and solve... ...alerting to improve platform reliability. • Lead... ...across deployment, monitoring, alerting, and self... ...with enterprise-level administration and...Software
- ...Job Role: SRE Engineer Location: Austin, TX//Southlake... ...used for environment monitoring and task automation... ...configurations for software and servers, DB connections... ...transaction including platform, services and tools... ...and vendor service level objectives (SLOs) and...SoftwareFull time
- ...Kafka Site Reliability Engineer to help build, operate... ...enterprise streaming platform ecosystem. This role combines... ...leveraging predictive monitoring and intelligent... ...operational metrics, service-level objectives (SLOs),... ...: Engineering & Software DevelopmentSalary Range...SoftwareFull timeWork at office
- ...We are currently looking for a Senior SRE / Technical Lead for an onsite role in... ...Ansible Strong experience with Splunk (monitoring, logging, observability) Proven... ...system improvements Work closely with engineering teams to improve system stability This...SuggestedRemote work
$116.2k - $229.1k
...Summary AI & Engineering/EaaS - DevOps Engineer... ...cloud, automation, and platform engineering solutions... ...strengthen software delivery performance... ...years of experience with monitoring and observability tools... ...Professional development From entry-level employees to senior...SoftwareEntry levelLocal area$133k - $158k
...technology consulting and software development company... ...Job Title: Azure Platform Engineer Location: 100% Remote... ...development, security, and SRE teams to deliver cloud... .... Production-level experience with infrastructure... ...implementing monitoring, alerting, and observability...SoftwareFull timeH1bLocal areaImmediate startRemote workVisa sponsorship$25 - $30 per hour
Platform Engineering Intern Build a career that matches all your initiative with an impressive dose... ...efficiency. Participate in monitoring, observability, and incident response... ...enhancements and automation efforts that improve software delivery processes. Document platform...SoftwareEntry levelHourly payFull timeSummer workInternshipSummer internshipWork at officeWork from homeMonday to Friday$130k - $200k
...creatives, designers, engineers, and architects. Our... ...Configure and integrate platforms such as NICE CXone,... ...flows, security, monitoring, and scalability.Contribute... ...consulting, software engineering, cloud architecture... ...development From entry-level employees to senior...SoftwareEntry levelLocal area- ...York City, Take-Two Interactive Software, Inc. is a leading developer,... ...experiences, delivered on every platform relevant to our audience through... ...Challenge We are seeking a mid-level Systems Engineer II with a focus on platform and monitoring systems to support and optimize...SoftwareFull timeCasual workFlexible hours
$93k - $189k
...posting). The z/OS Platform CICS, MQ & zCEE... ...experience include z/OS engineering, Workload Manager,... ..., and executive level communication is... ...diagrams of all system z software zCEE, CICS and MQ.... ...). Proficient in SRE concepts and... ...improvement of monitoring, performance, resiliency...SoftwareWork at officeRemote workWork from homeFlexible hours$102.75k - $195.25k
...and test high quality software, acts autonomously,... ...modernizing (including AI-led engineering), and on a team that... ...GlobalAdvantage (GA) platform, you'll own the... ...with monitoring, logging, deployment,... ...Professional development From entry-level employees to senior leaders...SoftwareEntry levelWork at officeLocal areaRemote workVisa sponsorship- ...intelligence of the platforms that power... ...Site Reliability Engineer Are you a skilled... ...passionate about combining software systems... ...to evangelize the SRE mindset, build groundbreaking... ...deployment, monitoring, alerting, and self... ...experience in enterprise-level administration and...Software
- ...ESOM—End Point Support and Operations Monitoring contract. Supporting the Veteran’s... ...roles and responsibilities of enterprise level hardware/software stacks and able to triage, diagnose,... ...in computer science, electronics engineering or other engineering or technical discipline...SoftwareContract work
- ...Enablement organization is seeking a senior Software Development & Engineering Lead to help drive modernization of... ...the next generation of engineering platforms and services supporting strategic... ....Experience with observability, monitoring, and production support practices....SoftwareFull timeWork at office
$104.5k - $234.6k
...state of the art observability platform, powering visibility and... ...workloads on OCI. OCI Logging and Monitoring serve as foundational platforms used by OCI engineering teams to operate and troubleshoot... ....Key ResponsibilitiesPlatform Software Development:Lead cross-team...SoftwareTemporary workFlexible hoursShift work$140k - $224.25k
...seeking a Senior Data Engineer to become part of its... ...throughout DGX Cloud. Our platform supports engineering,... ...and completeness monitoring, actionable alerting,... ...operating production software, data platforms, backend... ...USD - 224,250 USD for Level 3, and 168,000 USD - 2...SoftwareFull timeRemote work$200k - $322k
...organization is seeking a Senior Data Engineer to become part of its data... ...throughout DGX Cloud. Our platform supports engineering,... ...building and operating production software, data platforms, databases, or... ..., including testing, CI/CD, monitoring, alerting, rollback, incident...SoftwareFull timeImmediate startRemote workShift work$155k - $175k
...technology consulting and software development company... ...Title: Senior DevOps Engineer Location: 100% Remote... ...infrastructure, CI/CD platforms, and deployment... ...automating deployments, monitoring production environments... ...reliability engineering (SRE), chaos engineering,...SoftwareFull timeH1bLocal areaRemote workVisa sponsorship$141.2k - $278.3k
...Dynamics 365, Power Platform, Microsoft Fabric, Microsoft... ..., reliability engineering, SecOps, FinOps, data... ...automation, and AI-assisted software delivery.Champion... ..., evaluation, monitoring, and responsible AI controls... ...development From entry-level employees to senior leaders...SoftwareEntry levelFull timeLocal areaFlexible hours- ...innovation partner delivering software, services, and solutions to... ...expertise, proven software platforms, and innovative AI-driven solutions... ...Senior ServiceNow Platform Engineer to manage, enhance, and... ...through proactive monitoring and administration.Plan and...SoftwareMinimum wageFull time
- ...specified location(s).As a Site Reliability Engineer, you will play a critical role in... ...Support, Site Reliability Engineering (SRE), Software Operations, or a related technology support... ....Strong experience with application monitoring and observability tools, including AppDynamics...SoftwarePermanent employmentFull timeWork at officeNight shiftWeekend work
- ...seeking a Site Reliability Engineer (SRE) to improve the reliability,... ...will automate infrastructure, monitor production systems,... ...incidents, and collaborate with software engineering teams to ensure... ...AWS, Azure, or Google Cloud Platform Docker and Kubernetes...Software
- ...developing an end-to-end platform that is secure,... ...seeking a DevOps Engineer to design, build,... ...observability and monitoring frameworks using... ...Mentor junior and mid-level engineers on... ...Infrastructure Engineering, Software Engineering, or a... ...Relevant entry-level or associate...SoftwareEntry levelInternshipLive outWork at officeLocal areaFlexible hours
$198k - $250k
...Site Reliability Engineer role at Bumble Inc... ...Engineers (SRE) are responsible for... ...and performance of software systems while bridging... ...logging, Monitoring, tracing and alerting... ...and observability platforms such as: Grafana,... ...interests. Seniority level ~ Seniority...SoftwareFull timeRemote work- ...of experience with enterprise level administration and support • 6... ...application dashboards for proactive monitoring, setting up alerts for early... ...experience practicing SDLC (Software Development Lifecycle)... ...Agile methodologies Need seasoned SRE veterans for this position. Personal...SoftwareRemote work
- ...Strategy, Experience & Design, Engineering and Managed Services. We... ...build and run an internal agent platform on AWS Bedrock AgentCore for... ...tracing, SLOs, alerting, and cost monitoring for token and compute spend.... ....QUALIFICATIONS6+ years in SRE, platform reliability or...
- Job Title: Ai Platform Engineer Duration: 6 Months (possibility of conversion... ...support of AI platforms. Monitor platform health,... ...Define support models, service level objectives, and operational... ...usage analytics. Knowledge of software deployment and enterprise application...Software
$134.5k - $265.1k
...Our Deloitte AI & Engineering team works to transform technology platforms, drive innovation, and... ...precision, building working software, engaging confidently... ...frameworks, model monitoring, and prompt managementExperience... ...development From entry-level employees to senior...SoftwareEntry levelLocal area- Take-Two Interactive Software, Inc. is seeking a mid-level Systems Engineer II to support and optimize our enterprise data pipeline and monitoring infrastructure. You will be a key administrator... .... In this role you will manage platform configurations, apply scripting to...Software
- ...t mean fewer humans — it means humans focus on judgment calls the platform can't yet make, and every judgment call trains the platform to do it next time. In this L1 role you cover front-line monitoring and incident response for NeoCloud's US GPU DCs during the 8AM-8PM...Local areaShift workNight shift
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to SRE Monitoring Platform Software Engineer (Entry Level). Be the first to apply!
- monitor Austin, TX
- monitor tech Austin, TX
- pool monitor Austin, TX
- security monitor Austin, TX
- quality assurance monitor Austin, TX
- clinical research monitor Austin, TX
- site reliability engineer Austin, TX
- site reliability engineer remote Austin, TX
- site reliability engineer sre Austin, TX
- client platform engineer Austin, TX



