Site Reliability Engineer - Hosting
Are you looking for an exciting opportunity?
Join a specialist technology provider delivering advanced provisioning, management, and security solutions for data centers. The organization helps operators enhance customer experience, streamline day-to-day operations, and stay ahead of the competition through innovative products and services, allowing them to focus on their core strengths in hardware and infrastructure.
If you would like to learn more about this opportunity, feel free to reach out and apply today!
Responsibilities:
- Install and integrate Hydra’s Brokkr software with new datacenters and onboarded servers
- Maintain integrated datacenter and inventory, respond to L2 and L3 requests and alerts, and improve monitoring and other supporting infrastructure
- Monitor system performance and uptimes, ensuring the highest level of systems and infrastructure availability.
- Liaise with vendors and other IT personnel for problem resolution.
- Install, configure, test, and maintain operating systems, application software, and system management tools.
- Maintain security, backup, and redundancy strategies.
- Write and maintain custom scripts to increase system efficiency and lower human intervention time on any tasks.
- Participate in the design of information and operational support systems.
Required Skills/Qualifications:
- BS/MS degree in Computer Science, Engineering, or a related subject. Equivalent experience accepted.
- Proven working experience in installing, configuring, and troubleshooting UNIX/Linux-based environments.
- Solid experience in the administration and performance tuning of application stacks (e.g., Apache, MySQL, NGINX).
- Experience with virtualization and containerization (e.g., QEMU/KVM, Docker).
- Experience with monitoring systems (e.g., Nagios, Zabbix).
- Experience with automation software (e.g., Puppet, Chef, Ansible).
- Solid scripting skills (e.g., shell scripts, Perl, Ruby, Python).
- Solid networking knowledge (OSI network layers, TCP/IP, DNS, DHCP).
Desirable Skills:
- Certification in relevant fields (e.g., Linux Certifications, Cisco Certified Network Associate - CCNA, Microsoft Certified Systems Engineer - MCSE) are a plus.
- Experience with cloud services (AWS, Microsoft Azure) is a plus.
- Strong problem-solving skills and the ability to work under pressure is a must.
- Strong communication skills and the ability to collaborate and be proactive in asking questions is a must.
Benefits :
- Flexible working hours and remote work opportunities.
- A supportive team environment with an emphasis on learning and growth.
- Access to cutting-edge technology and tools.
Salary:
- Competitive salary and comprehensive benefits package.
$250k
...Europe, while now significantly expanding its footprint in the United States. The company is looking for a Senior / Staff Site Reliability Engineer to support and scale large-scale HPC and cloud environments powering GPU-intensive workloads. The role involves working...SuggestedFull timeRemote work$204k - $306k
...mission. If you are too, let's talk.Manager, Site Reliability EngineeringSan Francisco,... ...Francisco Office. The IDaaS Site Reliability Engineering GroupOkta authenticates, authorizes and... ...millions of users a day. The service is hosted on Amazon Web Services (AWS) across multiple...SuggestedPermanent employmentWork at officeLocal areaWorldwideFlexible hours2 days per week$181k - $263k
...line operational support. We are looking for a Senior Staff Site Reliability Engineer who will set the technical direction for reliability... ...collaborative, and friendly people who love what they do. Fun: We host in-person and virtual events such as game nights, happy...SuggestedFull timeWork from homeWorldwideFlexible hoursNight shift$232k - $319k
...millions of users a day. The service is hosted on Amazon Web Services (AWS) across multiple... ...scale the service with great people and reliable, cost-effective, and efficient... ...serviceAccelerate the velocity of SRE and product engineering by developing robust platforms, powerful...SuggestedPermanent employmentLocal areaWorldwideFlexible hours- ...deeply technical systems that must work reliably at massive scale. Our work underpins OpenAI... ...products, and emerging platforms. The Host Assurance team exists to make bare metal... ...the Role OpenAI is seeking a Software Engineer, Host Assurance to build and operate the...SuggestedFull time
$266k
...deeply technical systems that must work reliably at massive scale. Our work underpins OpenAI... ..., products, and emerging platforms.The Host Assurance team exists to make bare metal... ...About the RoleOpenAI is seeking a Software Engineer, Host Assurance to build and operate the...Work at officeLocal areaFlexible hours$113.4k - $162k
...break down barriers to communication and free the flow of conversation for people everywhere.TextNow is looking for motivated Site Reliability Engineer to own infrastructure, monitoring, logging, ci/cd, reliability and everything in between!This role is about impact at...Temporary work$114.3k - $235.32k
...verification who have now purpose-built a CTV performance platform advertisers can trust to grow their business.We are seeking a Site Reliability Engineer to help operate, scale, and continuously improve a cloud-native platform built on AWS, Kubernetes/EKS, and ArgoCD-driven...Work at officeLocal areaRelocationRelocation package$117k - $209.33k
Job Requisition ID #26WD99273Position OverviewWant to help make a better world? As a Senior Site Reliability Engineer at Autodesk, you can help us build and operate reliable, secure, and scalable cloud services for Autodesk GovCloud products.As part of a new SRE team supporting...Full timeFor contractors$152.5k - $205k
...flexible work environment where new ideas are encouraged and everyone is a stakeholder.What you’ll be responsible forThe Site Reliability Engineer builds and maintains shared platform capabilities, common libraries, and infrastructure that help Circle teams ship secure...Flexible hours- About the RoleWe’re looking for an experienced Site Reliability Engineer (SRE) to help us scale our platform with reliability, observability, and operational excellence at the core. You’ll partner with engineers and data scientists to build, automate, and maintain the infrastructure...
$148.5k - $223.9k
...right place! Agentforce is the future of AI, and you are the future of Salesforce.Salesforce is seeking a senior engineering candidate to join the Site Reliability organization in San Francisco. Working closely with counterparts in the Infrastructure and R&D organizations,...Full timeWorldwideWeekend work$167.7k - $245.2k
...very effective.We’re looking for talented engineers with a software or operations background... ...development teams to ensure the reliability, performance and security of our infrastructure... ...insurance. Please see the Cisco careers site to discover more benefits and perks....Full timeTemporary workWork at officeLocal areaFlexible hours1 day per week- ...an SRE to join our infrastructure team. This role will be responsible for building software to ensure the reliability of our back-end systems, working with engineers who develop them, and planning for our future growth. You will work with our existing production...WorldwideHome officeFlexible hours
$165k - $225.6k
...From core infrastructure to enterprise platforms, we partner across functions to drive scale, reliability, and innovation through technology.The Senior Site Reliability Engineer OpportunityReporting to the Manager, Site Reliability Engineering, this role will help build,...Permanent employmentLocal areaWorldwideFlexible hours$190.8k - $267.1k
...while helping Reddit grow its business. The reliability of our Ads systems directly impacts... ...Reliability team partners closely with Ads Engineering teams to improve reliability,... ...advertising ecosystem.We're looking for a Staff Site Reliability Engineer who will define and...For contractorsWork experience placementRemote workFlexible hours- ...let’s build what’s next.About the teamThe Engineering team at Airwallex is a diverse group of... ..., working together to build scalable, reliable, and secure products that empower businesses... ...services.What you’ll doAs a Senior Site Reliability Engineer, you’ll work closely...Temporary workLocal areaWorldwide
$165k - $227k
...opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk.The Engineering OpportunityWe are looking for an experienced Senior Site Reliability Engineer to join Okta's Emerging Products Group (EPG). Our mission is to build highly reliable...Local areaWorldwideFlexible hours$152.5k - $205k
...work environment where new ideas are encouraged and everyone is a stakeholder.What you’ll be responsible for:As a Senior Site Reliability Engineer on Circle’s platform team, you’ll design, build, and operate the secure, scalable platform infrastructure behind critical...Flexible hours$229.9k - $262.4k
Sr. Lead AI Engineer (Inference Optimization, FM hosting, AI Platform) Overview: At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital... ...available through this site. Capital One Financial is made...Full timePart timeLocal area$197.3k - $225.1k
Lead AI Engineer (FM Hosting, LLM Inference) Overview At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an... ...information available through this site. Capital One Financial is made up of...Full timePart timeLocal area- ...Site Reliability Engineer Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling...Flexible hours
- ...The role We're looking for a world-class Site Reliability Engineer to ensure the reliability, performance, and scalability of our AI infrastructure platform. You'll be building and operating the core systems that power agentic AI at scale. Your mission: keep...
$260k - $300k
...software agents. We're the makers of Devin, the first AI software engineer. Our team is extremely talent-dense. Among our founding... ...faster than anyone expects. You will own both the production reliability of our user-facing products and the platform engineering that...$170k - $220k
...Senior Site Reliability Engineer Supio is a trusted AI platform purpose-built for law firms, reshaping how data drives impactful outcomes. Our innovative approach blends technology with deep legal expertise, making us a leader in our field. We go beyond surface-level...Work at officeRemote workFlexible hours- ...enterprise that runs the real economy. Learn more about our vision in our manifesto. About the Role We're looking for a Site Reliability Engineer to take the lead on scaling our operational resilience as we grow. You'll own the stability, observability, and debugging...WorldwideShift work
- ...come shape the future and be part of a truly unique global culture at OutSystems! Hybrid Onsite in Menlo Park, CA Site Reliability Engineering (SRE) is a discipline that incorporates aspects of software engineering and applies them to infrastructure and...Immediate startRemote workWorldwide
- ...Site Reliability Engineer Specter's mission is to help automate the physical world. Today, we build video sensors with state-of-the-art AI agents that answer any question, anywhere in their environments. Our systems can automatically detect and reason about any physical...Remote work
$81.1k - $187k
...Site Reliability Engineer 3 We are looking for a Site Reliability Engineer 3 to support mission-critical cloud services and production operations. The role focuses on improving service reliability, reducing operational risk, automating repetitive tasks, and driving...Temporary workImmediate startFlexible hoursShift work- ...the globe. Join us on this journey to redefine resource management-and change lives along the way. The Role As a Site Reliability Engineer (SRE) at Air Apps, you will be responsible for ensuring the reliability, availability, and scalability of our systems. You...Temporary workWorldwide
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer - Hosting. Be the first to apply!
- site reliability engineer San Francisco, CA
- site reliability engineer remote San Francisco, CA
- site reliability engineer sre San Francisco, CA
- site recruiter San Francisco, CA
- junior website developer San Francisco, CA
- on site coordinator San Francisco, CA
- construction site safety San Francisco, CA
- site services specialist San Francisco, CA
- website content developer San Francisco, CA
- website coordinator San Francisco, CA



