Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Principal Site Reliability Engineer

KēSTA I.T.

Job Description

Job Description

Come build, innovate, disrupt, and thrive!

KēSTA I.T. is actively seeking a Principal Engineer for an immediate full-time opportunity with our industry creating client.

Are you on the lookout for a unique career opportunity that offers leadership, responsibility, and the chance to make a significant impact? If you're eager to contribute to a thriving and stable organization while maintaining your confidentiality, continue reading.

The Opportunity

An innovative technology company is seeking experienced Site Reliability Engineers to take ownership of building reliable, scalable platforms that deliver advanced 3D/4D spatial content to global users across AR/VR environments.

This is a high-impact role focused on ensuring system reliability at scale, requiring deep expertise in observability, multi-tenant architectures, and data-driven operational decision-making. You will play a key role in designing and maintaining infrastructure that supports high-volume streaming workloads while meeting enterprise-grade security and compliance standards.

This role partners closely with web services and platform engineering teams to implement SRE best practices, establish robust monitoring, and build infrastructure capable of supporting rapid growth and global distribution.

What You’ll Do

  • Design, configure, and maintain cloud infrastructure using infrastructure-as-code tools (e.g., Terraform), with a focus on optimizing content delivery and CDN performance
  • Develop and execute capacity planning strategies and performance optimization initiatives for large-scale streaming platforms
  • Instrument services to monitor system health, building dashboards and alerting systems that provide actionable insights into performance and user experience
  • Define and implement observability strategies, including SLI/SLO frameworks and error budget management
  • Establish escalation protocols and participate in on-call rotations to ensure 24/7 system availability
  • Lead incident response efforts and conduct post-incident reviews to drive continuous improvement
  • Implement and promote reliability engineering practices, including deployment safety, code review standards, and operational readiness
  • Mentor engineering teams on best practices for reliability, scalability, and production operations

What You’ll Bring

  • 7+ years of experience in Site Reliability Engineering, DevOps, or related roles, with a track record of improving system reliability and operational maturity
  • Strong expertise in cloud platforms and modern infrastructure environments (e.g., AWS, containerized workloads, or similar ecosystems)
  • Experience with infrastructure automation and container orchestration (e.g., Terraform, Kubernetes or equivalent technologies)
  • Deep understanding of multi-tenant architecture, security principles, and data protection practices
  • Hands-on experience with observability tools and monitoring frameworks (e.g., Prometheus, Grafana or similar)
  • Experience implementing automated compliance and governance practices (e.g., SOC 2, GDPR, ISO 27001 or similar standards)
  • Strong leadership and mentoring capabilities, with the ability to influence engineering teams and drive adoption of reliability-focused practices

About KēSTA I.T.:

Our name says it all; KēSTA I.T. (Keys-to-I.T.) AND our people are our keys to our success!

KēSTA I.T. is a premier Utah-based technical staffing and consulting services firm. We specialize in temporary and permanent placement of Software, Hardware, Network, Cloud, CRM/ERP, Data, End-User support, Web and Executive / leadership-based positions on a full time and consulting basis. If you're interested in a role where top performance is rewarded, personal time is valued, and excellence is demanded at every level we want to talk to you today!

Where do you want to go? We've got the keys! ~ KēSTA I.T.



Vacancy posted 12 days ago
Similar jobs that could be interesting for youBased on the Principal Site Reliability Engineer in Beverly Hills, CA vacancy
  •  ...Company Description Aerospace / Finance Job Description PRINCIPAL SOFTWARE ENGINEER (ARCHITECT)  Visa Candidates Welcome Job Description Responsibilities: • Maintain and improve the functionality and performance of existing Windows and WCF services... 
    Principal
    Full time

    Direct Staffing Inc

    Culver City, CA
    15 hours ago
  • $165k - $265k

     ...is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER - TOP SECRET CLEARANCE (STARLINK)At SpaceX we’re leveraging our experience in building rockets and spacecraft to deploy Starlink... 
    Suggested
    Permanent employment
    Temporary work
    Worldwide
    Weekend work

    SpaceX

    Hawthorne, CA
    3 days ago
  • $165k - $230k

     ...is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER (STARSHIELD)Starshield leverages SpaceX’s Starlink technology and launch capability to support national security efforts.... 
    Suggested
    Permanent employment
    Temporary work
    Immediate start
    Weekend work

    SpaceX

    Hawthorne, CA
    1 day ago
  • $165k - $265k

     ...is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER (STARLINK)At SpaceX we’re leveraging our experience in building rockets and spacecraft to deploy Starlink, the world’s most... 
    Suggested
    Permanent employment
    Temporary work
    Worldwide
    Weekend work

    SpaceX

    Hawthorne, CA
    4 days ago
  • $142.5k - $190k

    A prominent entertainment agency is seeking a Principal Architect to lead the design of technology infrastructure spanning on-premises and cloud environments. This role focuses on Microsoft Azure, driving zero-trust security models, and creating a multi-year infrastructure... 
    Principal

    Jobleads-US

    Beverly Hills, CA
    15 hours ago
  • $153k - $185k

     ...infrastructure, from in-orbit pharmaceutical processing to reliable and economical reentry capsules. Varda’s W-Series...  ...and materials science — and we’re looking for bold engineers to help us get there. As a Senior Site Reliability Engineer, you'll be critical in building,... 
    Permanent employment
    Full time
    Immediate start
    Relocation package
    Flexible hours
    Weekend work

    Varda Space Industries

    El Segundo, CA
    15 hours ago
  • $145k - $195k

     ...SpaceX is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SITE RELIABILITY ENGINEER (TOP SECRET CLEARANCE)As a member of the Classified IT Systems Engineering team, the Site Reliability Engineer is involved... 
    Permanent employment
    Temporary work
    Weekend work

    SpaceX

    Hawthorne, CA
    4 days ago
  • $125k - $150k

     ...SpaceX is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SITE RELIABILITY ENGINEER (RAPTOR)SpaceX is looking for a Site Reliability Engineer with a strong drive to solve challenging problems in the Raptor... 
    Permanent employment
    Temporary work

    SpaceX

    Hawthorne, CA
    4 days ago
  •  ...your big ideas, and your desire to team up with some of the best and brightest in technology and entertainment. The RoleThe Site Reliability Engineer (SRE) II is responsible for designing, implementing, and maintaining scalable and reliable systems and applications. Focus... 
    Full time
    Local area
    Worldwide
    Flexible hours

    AXS Group

    Los Angeles, CA
    4 days ago
  • $125k - $145k

     ...SpaceX is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SITE RELIABILITY ENGINEER - TOP SECRET CLEARANCEAs a Site Reliability Engineer, you will design, develop, and test key aspects of an in-house... 
    Permanent employment
    Temporary work
    Weekend work

    SpaceX

    Hawthorne, CA
    4 days ago
  • $155k - $195k

     ...you to join us on our mission of providing humankind access to the galaxy beyond our planet. About the RoleWe are seeking a Site Reliability Engineer to join our Ground Software team. As a Site Reliability Engineer, you will design, build, and operate the ground and site... 
    Full time
    Work at office

    Apex Technology

    Los Angeles, CA
    11 hours ago
  • $125k - $145k

     ...SpaceX is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SITE RELIABILITY ENGINEER, GNCSpaceX’s mission is to make humanity multiplanetary by developing fully and rapidly reusable launch systems capable of... 
    Permanent employment
    Temporary work
    Flexible hours
    Weekend work

    SpaceX

    Hawthorne, CA
    1 day ago
  • $107.8k - $162k

     ...expectation of a minimum of three (3) days per week working in the office and flexibility to work remotely on the remaining days. On-site expectations may evolve over time to support business needs, with clear communication provided in advance. JOB DESCRIPTION... 
    Full time
    Work at office
    Local area
    Remote work
    3 days per week

    Green Dot Corporation

    Los Angeles, CA
    5 days ago
  • ALO is seeking a Planner to join our Merchandising Planning team. You will collaborate with Planning, Buying and Merchandising to develop and communicate financial plans and merchandise strategies for International Owned Retail Stores and Web, driving top-down and bottom...
    Principal

    Alo

    Beverly Hills, CA
    2 days ago
  • $150k - $180k

     ...Senior Cloud Reliability EngineerIrvine, California, United States; Los Angeles, California, United StatesThe Senior Cloud Reliability Engineer will be responsible for writing and integrating various open source and closed sources tools. The ideal candidate will possess... 
    Work experience placement
    Local area

    Viant

    Los Angeles, CA
    5 days ago
  • $164k - $270k

     ...for the 21st century and beyond.The Role What You’ll DoOwn the reliability of our robotics systems, from PLCs through ROS2/middleware to...  ...remediation.Partner with controls, robotics, and platform engineering teams to bake reliability in early. Review designs, develop SLOs... 
    Permanent employment
    Full time
    Local area
    Flexible hours

    Hadrian

    Los Angeles, CA
    3 days ago
  • $30.53 - $56.48 per hour

    Job Title:Associate Site Reliability EngineerRequisition ID:R027696Job Description:Job Title: Associate Site Reliability EngineerReporting...  ...TechnologyLocation: Santa Monica, CaOverviewThe Associate Site Reliability Engineer helps keep Marketing Technology services reliable, observable... 
    Hourly pay
    Full time
    Temporary work
    Part time
    Internship
    Local area
    Worldwide
    Relocation package

    Activision

    Santa Monica, CA
    1 day ago
  • $181k - $225k

     ...systems across all product teams. You will collaborate closely with engineering leadership, product managers, and cross-functional teams to...  ...and Helm Understand the importance of performant and reliable systems Education - Ideally looking for a B.A. / B.S.... 
    Full time
    Work at office
    Immediate start
    3 days per week

    Altruist

    Los Angeles, CA
    3 days ago
  • $164k - $270k

    Hadrian - Manufacturing the FutureHadrian is building autonomous factories that help aerospace and defense companies manufacture rockets, satellites, jets, and ships up to 10x faster and up to 2x cheaper. By combining advanced software, robotics, and full-stack manufacturing...
    Permanent employment
    Full time
    Local area
    Remote work
    Flexible hours

    Hadrian

    Los Angeles, CA
    4 days ago
  • $222k - $287k

     ...fast-growing technology startup with the unique opportunity to help build out a critical function for the company. As a Principal Software Engineer, Financial Systems , you will own the cross-cutting architecture of BuildOps’ financial platform and the systems that... 
    Principal
    Full time
    Contract work
    For contractors
    Work at office
    Local area
    Work from home
    Flexible hours

    Buildops

    Los Angeles, CA
    15 hours ago
  • $122.8k - $184.2k

     ...Defense Systems is seeking an Environments Capability Lead Sr. Principal Systems Engineer (level 4) to join its team. This position is in Roy UT,...  ...various stakeholder communitiesProven experience leading Reliability, Availability, Maintainability, and Cost (RAM-C)... 
    Principal
    Full time
    Relocation package
    Monday to Thursday
    Shift work

    Northrop Grumman

    Manhattan Beach, CA
    2 days ago
  • $220k - $320k

     ...About the team At Q-CTRL, Quantum Computing Engineering is a global team of software engineers and infrastructure experts,combining...  ...quantum advantage worldwide. About the role We are seeking a Principal Software Engineer to lead the architectural evolution of Q-... 
    Principal
    Full time
    Worldwide
    Flexible hours

    Q-ctrl

    Los Angeles, CA
    15 hours ago
  • $150k - $200k

     ...our CEO's funding announcement: The Reliability team owns the availability, performance,...  ...enforcing reliability standards across engineering Designing incident response processes...  ...ownership of production systems. As a Site Reliability Engineer on the Reliability... 
    Remote work
    Visa sponsorship
    Work visa
    Flexible hours

    GrabJobs

    Los Angeles, CA
    1 day ago
  • $240k - $250k

     ...BE DOING Build and deploy AI agents that automate customer escalation workflows — reducing resolution times and improving engineering efficiency. Integrate agentic capabilities into identity platforms, translating customer needs into engineering roadmap contributions... 
    Principal
    Full time

    Saviynt

    Los Angeles, CA
    15 hours ago
  •  ...advantage.About The Role:The Mission Systems Engineering team develops the Mission Management...  ...and ground systems. As a Senior/Principal Mission Software Engineer, you will architect...  ...decisions involving performance, reliability, latency, fault tolerance, and extensibility... 
    Principal
    Weekly pay
    Permanent employment
    Full time
    Work at office

    Hermeus

    Los Angeles, CA
    2 days ago
  •  ...What you will do: Partner with a team of high-performing engineers and developers who are focused on delivering best in class software...  ...our shift to a SecDevOps culture, solving for security, reliability, cost-effectiveness, and observability Building Zero trust... 
    Full time
    Contract work
    Local area
    Flexible hours
    Shift work

    DISQO

    Los Angeles, CA
    12 days ago
  • Amazon.com Services LLC is seeking a Principal UX Designer to lead frontier design and experimentation in an AI-first entertainment world...  ...’s evolving product landscape. You’ll partner with Product, Engineering, and Science teams, mentor designers, and push the industry... 
    Principal

    Prime Video & Amazon MGM Studios

    Culver City, CA
    1 day ago
  • $200k - $285k

     ...possible, with the ultimate goal of enabling human life on Mars.PRINCIPAL SOFTWARE ENGINEER (PLATFORM TEAM) The Platform Team builds the foundational...  ...prompts, skills, and infrastructure, this team unlocks reliable, high-impact AI capabilities across the entire... 
    Principal
    Permanent employment
    Temporary work

    SpaceX

    Hawthorne, CA
    2 days ago
  • Northrop Grumman Defense Systems (NGDS) is seeking a Software Engineer / Principal Software Engineer in Manhattan Beach, CA to work on state-of-...  ..., with the ability to obtain a DoD Secret Clearance. On-site work is required in Manhattan Beach. #J-18808-Ljbffr Northrop... 
    Principal

    Northrop Grumman

    Manhattan Beach, CA
    1 day ago
  • $200k - $285k

     ...possible, with the ultimate goal of enabling human life on Mars.PRINCIPAL SOFTWARE ENGINEER, CONTINUOUS INTEGRATION (STARSHIP)As a Principal Software...  ...of engineers to integrate their changes and deliver reliable software to all Starship vehicles and launch pads on a continuous... 
    Principal
    Permanent employment
    Temporary work
    Weekend work

    SpaceX

    Hawthorne, CA
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Principal Site Reliability Engineer. Be the first to apply!