Principal Site Reliability Engineer
Deimos
Principal Site Reliability Engineer
Deimos is a cloud-native developer and security operations technology services company. We help companies of all sizes adopt the cloud for improved service delivery to their clients. We're a fully remote African-based team of engineers who are passionate about implementing engineering best practices. We leverage the latest technologies while building globally competitive solutions for our clients. With Deimos being one of the two moons of Mars, we refer to ourselves as "Martians" who are on a mission to Mars, together.
Our teams value the ability to learn and adapt to technology changes while appreciating solid foundational design and the craft of software engineering. As such our engineers enjoy working with various clients who have different problems to solve. If this sounds like you then you would be an ideal fit for our environment. However, you must be based in one of the countries we currently hire in which are as follows: Kenya, Ghana, Nigeria, South Africa, and Senegal.
Role Overview
We are looking for an experienced Principal Site Reliability Engineer to join our Professional Services team and deliver Software and DevSecOps projects. You will report to a Site Reliability Engineering Manager. As a Principal Site Reliability Engineer you will be expected to fill the role of a technical lead on multiple projects simultaneously, representing the senior technical leadership within our organisation.
SRE / DevOps is one of our core competencies. You will be part of a highly-skilled team that continuously innovates and delivers high value solutions to clients across various industries on all public clouds (AWS, Azure, GCP, etc). Technologies we work with daily include Kubernetes, Helm, Terraform, GitOps, OPA, Calico, Linkerd, just to name a few.
What You Will Be Doing
- Design and build advanced cloud-native infrastructure
- Guide technical discussions with clients and build technical roadmaps
- Collaborate with the Engineering Director(s) to (re)design architecture
- Assist the Site Reliability Manager with resource planning
- Assist engineering managers with building career paths for individuals wishing to be promoted to Principal Engineers
- Teach, mentor, grow, and provide advice to other domain experts, individual contributors, and across several teams.
- Document processes and monitor performance metrics
- Guide conversations to remove blockers and encourage collaboration across teams.
- Constantly improve the stability, scalability, security, cost-effectiveness, and operational excellence of our clients' systems.
- Continuously discover, evaluate, and implement new technologies to maximize development efficiency and security.
- Conduct infrastructure planning, testing, and development
- Provide technical leadership on multiple projects.
What You Must Have
- At least 7 or more years experience working in a DevOps/SRE team
- Extensive experience in DevOps/SRE, team management and collaboration
- Advanced knowledge of best practices related to data encryption and cybersecurity
- Advanced knowledge of the general DevOps/SRE landscape, architectures, and emerging technologies
- Cloud experience, preferably GCP, Azure and AWS
- Experience in Observability Practices and Incident Management
- Extensive experience with Prometheus, Grafana, the Elastic Stack and all versions of Beats, especially within Kubernetes
- Experience with Infrastructure as Code, preferably Terraform
- Experience with general automation and config management, preferably Ansible
- Extensive experience building and maintaining Kubernetes clusters and workloads
- Strong foundation of basic network and security concepts
- Ability to build robust CICD pipelines
- Familiarity with relational and non-relational databases
- Solid understanding of Linux operating systems
Qualities & Behaviours
- Exceptional interpersonal and communication skills
- A zest for automation
- Comfortable working as a remote team member and leader
- Ability to keep up to date with DevOps/SRE best practices, trends and innovation
- Passionate about mentoring and growing technical skills within the team
About You
For us to achieve our ambitious vision together as a team, it is important for our Martians to lead at all levels, be self starters who take initiative and put their hands up for challenging tasks. A growth mindset is important to us and we encourage all our Martians to openly share knowledge, support and help each other, ask questions, get creative with new technologies and learn from setbacks.
Becoming a Martian means:
- Comfortably working and learning from a fully remote, culturally diverse team based predominantly in South Africa, Kenya, Nigeria and Ghana.
- Being an open, honest and respectful communicator.
- You enjoy asking questions, identifying areas of improvement and proposing solutions, no matter your job title or whether you have been with us for a day, a month or years!
- You are comfortable taking initiative and operating independently.
- You thrive in a fast paced environment, where change is constant.
- You find it exciting to work with various clients, from different industries, each with a different problem for you and your team to solve.
- Intentionally sharing tech and industry trends that excite you with your peers.
- Seeking continuous feedback and actively taking steps to continuously grow personally and professionally.
Want To Know What You Get By Joining Us?
- Become a member of a team where we value each individual's contribution from day 1 and empower you to make suggestions, get involved and do what you love most!
- Flexibility and the freedom to work remotely.
- Work-life balance where you are not expected to work over weekends or after hours.
- A forward thinking remote company that knows how important it is to stay connected as one team, by providing virtual social platforms for employee engagement.
- A monthly work from home allowance which you can use to set yourself up to work comfortably from home. Whether that is pens, notebooks, new headphones or work snacks!
- A MacBook or Windows laptop for you to do your best work on.
- Become part of a team of exceptionally clever and talented people who like to share their knowledge and learnings.
- We support your career growth and love to celebrate your successes and advancement!
$139.7k - $232.9k
...designing, implementing, and continuously improving highly reliable, scalable, and resilient platform solutions across the enterprise. Operates as a subject matter expert (SME) in Site Reliability Engineering, driving reliability engineering practices, operational excellence...PrincipalFull timeWork experience placement- ...future sponsorship.Maintain and enhance the reliability, availability, and performance of Navy... ...page of the Navy Federal Career Site.Protect Yourself from Job Scams: Navy Federal... ...Act.Master's degree in computer science, engineering, or the equivalent combination of education...PrincipalInternshipMonday to Friday
$142.8k - $274.8k
...yearEmployment type: Full-TimeWork site: 0 days / week in-office - remoteRole... ...Software EngineeringDiscipline: Site Reliability EngineeringCompany:... ...world’s most demanding workloads. As a Principal Site Reliability Engineer, you will set technical and operational...PrincipalOngoing contractWork at officeLocal area- ...Responsibilities:Defines and leads enterprise-level reliability strategies.Architects resilient... ...senior leadership on reliability engineering best practices.Mentors junior... ...and five (5) years of experience as a Principal Site Reliability Engineer (or closely related...PrincipalFull time
- ...lives. Ready to build the next breakthrough? Join us to start Caring. Connecting. Growing together.We are seeking a Principal Site Reliability Engineer (SRE) to define and scale reliability practices across large-scale cloud platforms.This is a senior individual contributor...PrincipalMinimum wageFull timeWork experience placementWork at officeLocal areaRemote work
- ...Infrastructure Code. Builds reliability into the ecosystem by applying... ...practices in resiliency engineering and observability by developing... ...engineering techniques with site reliability engineering... ...5) years of experience as a Principal Site Reliability Engineer (or...PrincipalFull time
$175.5k - $235.4k
...MyDisneyExperience and Hey, Disney!This role sits in the Commerce Site Reliability Engineering (SRE) specifically supporting Ecommerce , Consumer... ...Products Technology teams from across the company. The Principal of DXT SRE will report to the Director of DXT Commerce SREAbout...PrincipalWorldwide$84.9k - $209.5k
This role combines strategic architecture with practical systems engineering, deployment, automation, patching, troubleshooting, incident response, and compliance support. The Principal Site Reliability Engineer will work across Windows, Linux, Oracle Cloud Infrastructure...PrincipalTemporary workFlexible hours$163.62k - $212.71k
...maintaining the tools, platforms, and processes that improve our engineering teams' productivity and streamline the software... ...Responsibilities:We are seeking a seasoned and strategic Lead/Principal Site Reliability Engineer to drive the reliability, scalability, and...PrincipalFull timePart timeWork experience placementWork at officeLocal areaImmediate startRemote workWork from homeFlexible hoursShift work3 days per week1 day per week$84.9k - $209.5k
.... You’ll partner with customer support, service owners, and engineering teams around the globe to ensure high-quality service for customers... ...posted.Career Level - IC4Escalation points for junior site reliability engineers during complex or high-impact incidents.Manage and...PrincipalTemporary workMonday to FridayFlexible hoursShift workNight shift- ...Principal Site Reliability Engineer location- Washington, DC -Onsite Remote- No 6+ Months Job Summary At Amtrak, we're seeking a seasoned Principal Site Reliability Engineer with a strong focus on pipelines as code, CI/CD, and IaC. You...PrincipalRemote work
$159k - $272k
...generosity. Join us for the opportunity to grow and make a difference in ways that matter to you. Role SummaryIn this role as Principal Site Reliability Engineer, Infrastructure Observability you will help formulate, develop, and implement a team of Site Reliability Engineers (...PrincipalFull timePrivate practiceLocal areaRemote workWork from home3 days per week$96.3k - $264.1k
...infrastructure and service, ensuring alignment with reliability and functionality standards. Takes full... ...tools and provides expertise in site reliability trends.Only Oracle brings... ...LeadershipDefine and drive the site reliability engineering strategy for large-scale, distributed,...PrincipalTemporary workFlexible hours- ...disrupt, and thrive! KēSTA I.T. is actively seeking a Principal Engineer for an immediate full-time opportunity with our industry creating... ...An innovative technology company is seeking experienced Site Reliability Engineers to take ownership of building reliable, scalable...PrincipalPermanent employmentFull timeTemporary workImmediate start
$272k - $431.25k
NVIDIA is looking for a Cloud Site Reliability Engineering Architect to work in IPP's (Infrastructure, Planning and Process) Cloud Infrastructure Team. IPP is a global organization within NVIDIA. This group works with various other groups within NVIDIA such as Graphics...PrincipalFull timeWork experience placementWorldwide- ...Principal Site Reliability Engineer About ShipperHQ: ShipperHQ is a trusted leader in the e-commerce shipping space, with over 15 years of experience helping merchants deliver better checkout experiences. Founded in 2009, we power shipping logic and checkout optimization...PrincipalFull timeWork at office
$240k - $250k
...MattersSaviynt’s platform is mission-critical for our customers. As we scale globally, reliability, availability, and performance are not optional—they are core product features.As a Principal Engineer, you will define and drive the reliability strategy for our SaaS platform....PrincipalFull time- ...products and services that help people, businesses and governments realize their greatest potential. Title and Summary Principal Site Reliability Engineer Who is Mastercard? At Mastercard technology, we work to connect and power an inclusive, digital economy that...PrincipalFull timeWorldwideFlexible hours
$147k - $237.5k
...builds and delivers the industry's most advanced SecOps platform, consisting of XDR, XSIAM, XSOAR, and XPANSE. As a Principal Site Reliability Engineer within the Cortex DevOps team, you will serve as a technical leader responsible for driving the reliability,...Principal$160k - $180k
...global, with headquarters in Denver, Colorado, and offices across the U.S., Canada, and India. We are seeking a Principal Site Reliability Engineer to define the strategic vision and own the enterprise-wide reliability, scalability, and performance of our critical...PrincipalContract workWork from homeFlexible hours$151.6k - $245.3k
...About the Role Palo Alto Networks runs a large infrastructure and is one of the largest GCP customers. As a Principal Site Reliability Engineer for the ADEM (Autonomous Digital Experience Management) team, you will be part of a team supporting the services that...PrincipalFull timeWork at officeVisa sponsorshipWork visa$142.8k - $274.8k
...yearEmployment type: Full-TimeWork site: 3 days / week in-officeRole type:... ...Software EngineeringDiscipline: Site Reliability EngineeringCompany:... ...demanding workloads.We are seeking a Principal Site Reliability Engineering Manager to lead a team responsible...PrincipalOngoing contractTemporary workFixed term contractLocal areaImmediate start3 days per week$84.9k - $209.5k
...Job Description As a Principal Site Reliability Engineer (IC4), you will be responsible for designing, building, and operating highly available, scalable, secure, and resilient cloud services. You will combine software engineering with infrastructure expertise to improve...PrincipalTemporary workFlexible hours- Role Description Symmetrio is recruiting a Principal Site Reliability Engineer (SRE) for our customer, a rapidly growing healthcare technology organization focused on advanced healthcare technology solutions. This individual will play a critical role in ensuring the reliability...PrincipalFull time
$194k - $237k
...at the date of hire. This position is ineligible for employment Visa sponsorship. Overall Purpose The Principal Site Reliability Engineer partners with development teams by designing availability and resiliency patterns in applications and infrastructure....PrincipalHourly payWork at officeImmediate startVisa sponsorshipWork visaFlexible hours- ...serve.The Information Technology group delivers secure, reliable technology solutions that enable DTCC to be the trusted... ...scalability, and performance of enterprise platforms.As a Principal Site Reliability Engineer (SRE), you will drive operational excellence across...PrincipalRemote workFlexible hours
- ...Invisible Technologies is looking for a Principal Software Engineer (SRE/DevOps) to work remotely. The ideal candidate will possess dual expertise in application engineering and infrastructure, contributing to a variety of technical initiatives. This role includes overseeing...PrincipalRemote work
$96k - $163k
...products and services that help people, businesses and governments realize their greatest potential. Title and Summary Senior Site Reliability Engineer Who is Mastercard? At Mastercard technology, we work to connect and power an inclusive, digital economy that...Full timePart timeWorldwideFlexible hours- ...services that help people, businesses and governments realize their greatest potential. Title and Summary Senior Site Reliability Engineer Overview-The ProCOM team is looking for a Site Reliability Engineering (SRE) who can help us solve problems, build our...Full timePart timeImmediate startWorldwideFlexible hours
$96k - $163k
...products and services that help people, businesses and governments realize their greatest potential. Title and Summary Senior Site Reliability Engineer, Performance Engineering Senior Site Reliability Engineer, Performance Engineering Payment Optimization unifies...Full timePart timeWorldwideFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Principal Site Reliability Engineer. Be the first to apply!
- associate director engineering United States
- principal network engineer United States
- senior director engineering United States
- director mechanical engineering United States
- civil engineer project manager United States
- aerospace engineering director United States
- principal developer United States
- clinical engineering director United States
- chief design engineer United States
- principal test engineer United States




