Principal Site Reliability Engineer
KēSTA I.T.
Job Description
Job Description
Come build, innovate, disrupt, and thrive!
KēSTA I.T. is actively seeking a Principal Engineer for an immediate full-time opportunity with our industry creating client.
Are you on the lookout for a unique career opportunity that offers leadership, responsibility, and the chance to make a significant impact? If you're eager to contribute to a thriving and stable organization while maintaining your confidentiality, continue reading.
The Opportunity
An innovative technology company is seeking experienced Site Reliability Engineers to take ownership of building reliable, scalable platforms that deliver advanced 3D/4D spatial content to global users across AR/VR environments.
This is a high-impact role focused on ensuring system reliability at scale, requiring deep expertise in observability, multi-tenant architectures, and data-driven operational decision-making. You will play a key role in designing and maintaining infrastructure that supports high-volume streaming workloads while meeting enterprise-grade security and compliance standards.
This role partners closely with web services and platform engineering teams to implement SRE best practices, establish robust monitoring, and build infrastructure capable of supporting rapid growth and global distribution.
What You’ll Do
- Design, configure, and maintain cloud infrastructure using infrastructure-as-code tools (e.g., Terraform), with a focus on optimizing content delivery and CDN performance
- Develop and execute capacity planning strategies and performance optimization initiatives for large-scale streaming platforms
- Instrument services to monitor system health, building dashboards and alerting systems that provide actionable insights into performance and user experience
- Define and implement observability strategies, including SLI/SLO frameworks and error budget management
- Establish escalation protocols and participate in on-call rotations to ensure 24/7 system availability
- Lead incident response efforts and conduct post-incident reviews to drive continuous improvement
- Implement and promote reliability engineering practices, including deployment safety, code review standards, and operational readiness
- Mentor engineering teams on best practices for reliability, scalability, and production operations
What You’ll Bring
- 7+ years of experience in Site Reliability Engineering, DevOps, or related roles, with a track record of improving system reliability and operational maturity
- Strong expertise in cloud platforms and modern infrastructure environments (e.g., AWS, containerized workloads, or similar ecosystems)
- Experience with infrastructure automation and container orchestration (e.g., Terraform, Kubernetes or equivalent technologies)
- Deep understanding of multi-tenant architecture, security principles, and data protection practices
- Hands-on experience with observability tools and monitoring frameworks (e.g., Prometheus, Grafana or similar)
- Experience implementing automated compliance and governance practices (e.g., SOC 2, GDPR, ISO 27001 or similar standards)
- Strong leadership and mentoring capabilities, with the ability to influence engineering teams and drive adoption of reliability-focused practices
About KēSTA I.T.:
Our name says it all; KēSTA I.T. (Keys-to-I.T.) AND our people are our keys to our success!
KēSTA I.T. is a premier Utah-based technical staffing and consulting services firm. We specialize in temporary and permanent placement of Software, Hardware, Network, Cloud, CRM/ERP, Data, End-User support, Web and Executive / leadership-based positions on a full time and consulting basis. If you're interested in a role where top performance is rewarded, personal time is valued, and excellence is demanded at every level we want to talk to you today!
Where do you want to go? We've got the keys! ~ KēSTA I.T.
- ...Company Description Aerospace / Finance Job Description PRINCIPAL SOFTWARE ENGINEER (ARCHITECT) Visa Candidates Welcome Job Description Responsibilities: • Maintain and improve the functionality and performance of existing Windows and WCF services...PrincipalFull time
$165k - $265k
...is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER - TOP SECRET CLEARANCE (STARLINK)At SpaceX we’re leveraging our experience in building rockets and spacecraft to deploy Starlink...SuggestedPermanent employmentTemporary workWorldwideWeekend work$165k - $230k
...is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER (STARSHIELD)Starshield leverages SpaceX’s Starlink technology and launch capability to support national security efforts....SuggestedPermanent employmentTemporary workImmediate startWeekend work$165k - $265k
...is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER (STARLINK)At SpaceX we’re leveraging our experience in building rockets and spacecraft to deploy Starlink, the world’s most...SuggestedPermanent employmentTemporary workWorldwideWeekend work$142.5k - $190k
A prominent entertainment agency is seeking a Principal Architect to lead the design of technology infrastructure spanning on-premises and cloud environments. This role focuses on Microsoft Azure, driving zero-trust security models, and creating a multi-year infrastructure...Principal$153k - $185k
...infrastructure, from in-orbit pharmaceutical processing to reliable and economical reentry capsules. Varda’s W-Series... ...and materials science — and we’re looking for bold engineers to help us get there. As a Senior Site Reliability Engineer, you'll be critical in building,...Permanent employmentFull timeImmediate startRelocation packageFlexible hoursWeekend work$145k - $195k
...SpaceX is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SITE RELIABILITY ENGINEER (TOP SECRET CLEARANCE)As a member of the Classified IT Systems Engineering team, the Site Reliability Engineer is involved...Permanent employmentTemporary workWeekend work$125k - $150k
...SpaceX is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SITE RELIABILITY ENGINEER (RAPTOR)SpaceX is looking for a Site Reliability Engineer with a strong drive to solve challenging problems in the Raptor...Permanent employmentTemporary work- ...your big ideas, and your desire to team up with some of the best and brightest in technology and entertainment. The RoleThe Site Reliability Engineer (SRE) II is responsible for designing, implementing, and maintaining scalable and reliable systems and applications. Focus...Full timeLocal areaWorldwideFlexible hours
$125k - $145k
...SpaceX is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SITE RELIABILITY ENGINEER - TOP SECRET CLEARANCEAs a Site Reliability Engineer, you will design, develop, and test key aspects of an in-house...Permanent employmentTemporary workWeekend work$155k - $195k
...you to join us on our mission of providing humankind access to the galaxy beyond our planet. About the RoleWe are seeking a Site Reliability Engineer to join our Ground Software team. As a Site Reliability Engineer, you will design, build, and operate the ground and site...Full timeWork at office$125k - $145k
...SpaceX is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SITE RELIABILITY ENGINEER, GNCSpaceX’s mission is to make humanity multiplanetary by developing fully and rapidly reusable launch systems capable of...Permanent employmentTemporary workFlexible hoursWeekend work$107.8k - $162k
...expectation of a minimum of three (3) days per week working in the office and flexibility to work remotely on the remaining days. On-site expectations may evolve over time to support business needs, with clear communication provided in advance. JOB DESCRIPTION...Full timeWork at officeLocal areaRemote work3 days per week- ALO is seeking a Planner to join our Merchandising Planning team. You will collaborate with Planning, Buying and Merchandising to develop and communicate financial plans and merchandise strategies for International Owned Retail Stores and Web, driving top-down and bottom...Principal
$150k - $180k
...Senior Cloud Reliability EngineerIrvine, California, United States; Los Angeles, California, United StatesThe Senior Cloud Reliability Engineer will be responsible for writing and integrating various open source and closed sources tools. The ideal candidate will possess...Work experience placementLocal area$164k - $270k
...for the 21st century and beyond.The Role What You’ll DoOwn the reliability of our robotics systems, from PLCs through ROS2/middleware to... ...remediation.Partner with controls, robotics, and platform engineering teams to bake reliability in early. Review designs, develop SLOs...Permanent employmentFull timeLocal areaFlexible hours$30.53 - $56.48 per hour
Job Title:Associate Site Reliability EngineerRequisition ID:R027696Job Description:Job Title: Associate Site Reliability EngineerReporting... ...TechnologyLocation: Santa Monica, CaOverviewThe Associate Site Reliability Engineer helps keep Marketing Technology services reliable, observable...Hourly payFull timeTemporary workPart timeInternshipLocal areaWorldwideRelocation package$181k - $225k
...systems across all product teams. You will collaborate closely with engineering leadership, product managers, and cross-functional teams to... ...and Helm Understand the importance of performant and reliable systems Education - Ideally looking for a B.A. / B.S....Full timeWork at officeImmediate start3 days per week$164k - $270k
Hadrian - Manufacturing the FutureHadrian is building autonomous factories that help aerospace and defense companies manufacture rockets, satellites, jets, and ships up to 10x faster and up to 2x cheaper. By combining advanced software, robotics, and full-stack manufacturing...Permanent employmentFull timeLocal areaRemote workFlexible hours$222k - $287k
...fast-growing technology startup with the unique opportunity to help build out a critical function for the company. As a Principal Software Engineer, Financial Systems , you will own the cross-cutting architecture of BuildOps’ financial platform and the systems that...PrincipalFull timeContract workFor contractorsWork at officeLocal areaWork from homeFlexible hours$122.8k - $184.2k
...Defense Systems is seeking an Environments Capability Lead Sr. Principal Systems Engineer (level 4) to join its team. This position is in Roy UT,... ...various stakeholder communitiesProven experience leading Reliability, Availability, Maintainability, and Cost (RAM-C)...PrincipalFull timeRelocation packageMonday to ThursdayShift work$220k - $320k
...About the team At Q-CTRL, Quantum Computing Engineering is a global team of software engineers and infrastructure experts,combining... ...quantum advantage worldwide. About the role We are seeking a Principal Software Engineer to lead the architectural evolution of Q-...PrincipalFull timeWorldwideFlexible hours$150k - $200k
...our CEO's funding announcement: The Reliability team owns the availability, performance,... ...enforcing reliability standards across engineering Designing incident response processes... ...ownership of production systems. As a Site Reliability Engineer on the Reliability...Remote workVisa sponsorshipWork visaFlexible hours$240k - $250k
...BE DOING Build and deploy AI agents that automate customer escalation workflows — reducing resolution times and improving engineering efficiency. Integrate agentic capabilities into identity platforms, translating customer needs into engineering roadmap contributions...PrincipalFull time- ...advantage.About The Role:The Mission Systems Engineering team develops the Mission Management... ...and ground systems. As a Senior/Principal Mission Software Engineer, you will architect... ...decisions involving performance, reliability, latency, fault tolerance, and extensibility...PrincipalWeekly payPermanent employmentFull timeWork at office
- ...What you will do: Partner with a team of high-performing engineers and developers who are focused on delivering best in class software... ...our shift to a SecDevOps culture, solving for security, reliability, cost-effectiveness, and observability Building Zero trust...Full timeContract workLocal areaFlexible hoursShift work
- Amazon.com Services LLC is seeking a Principal UX Designer to lead frontier design and experimentation in an AI-first entertainment world... ...’s evolving product landscape. You’ll partner with Product, Engineering, and Science teams, mentor designers, and push the industry...Principal
$200k - $285k
...possible, with the ultimate goal of enabling human life on Mars.PRINCIPAL SOFTWARE ENGINEER (PLATFORM TEAM) The Platform Team builds the foundational... ...prompts, skills, and infrastructure, this team unlocks reliable, high-impact AI capabilities across the entire...PrincipalPermanent employmentTemporary work- Northrop Grumman Defense Systems (NGDS) is seeking a Software Engineer / Principal Software Engineer in Manhattan Beach, CA to work on state-of-... ..., with the ability to obtain a DoD Secret Clearance. On-site work is required in Manhattan Beach. #J-18808-Ljbffr Northrop...Principal
$200k - $285k
...possible, with the ultimate goal of enabling human life on Mars.PRINCIPAL SOFTWARE ENGINEER, CONTINUOUS INTEGRATION (STARSHIP)As a Principal Software... ...of engineers to integrate their changes and deliver reliable software to all Starship vehicles and launch pads on a continuous...PrincipalPermanent employmentTemporary workWeekend work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Principal Site Reliability Engineer. Be the first to apply!
- on-site clinical research associate (traveling/remote) Beverly Hills, CA
- construction site safety Beverly Hills, CA
- IT site lead Beverly Hills, CA
- official site Beverly Hills, CA
- site leader Beverly Hills, CA
- chief structural engineer
- director protein engineering
- project engineer assistant project manager
- principal data engineer
- principal cloud engineer



