Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Principal Site Reliability Engineer

KēSTA I.T.

Job Description

Job Description

Come build, innovate, disrupt, and thrive!

KēSTA I.T. is actively seeking a Principal Engineer for an immediate full-time opportunity with our industry creating client.

Are you on the lookout for a unique career opportunity that offers leadership, responsibility, and the chance to make a significant impact? If you're eager to contribute to a thriving and stable organization while maintaining your confidentiality, continue reading.

The Opportunity

An innovative technology company is seeking experienced Site Reliability Engineers to take ownership of building reliable, scalable platforms that deliver advanced 3D/4D spatial content to global users across AR/VR environments.

This is a high-impact role focused on ensuring system reliability at scale, requiring deep expertise in observability, multi-tenant architectures, and data-driven operational decision-making. You will play a key role in designing and maintaining infrastructure that supports high-volume streaming workloads while meeting enterprise-grade security and compliance standards.

This role partners closely with web services and platform engineering teams to implement SRE best practices, establish robust monitoring, and build infrastructure capable of supporting rapid growth and global distribution.

What You’ll Do

  • Design, configure, and maintain cloud infrastructure using infrastructure-as-code tools (e.g., Terraform), with a focus on optimizing content delivery and CDN performance
  • Develop and execute capacity planning strategies and performance optimization initiatives for large-scale streaming platforms
  • Instrument services to monitor system health, building dashboards and alerting systems that provide actionable insights into performance and user experience
  • Define and implement observability strategies, including SLI/SLO frameworks and error budget management
  • Establish escalation protocols and participate in on-call rotations to ensure 24/7 system availability
  • Lead incident response efforts and conduct post-incident reviews to drive continuous improvement
  • Implement and promote reliability engineering practices, including deployment safety, code review standards, and operational readiness
  • Mentor engineering teams on best practices for reliability, scalability, and production operations

What You’ll Bring

  • 7+ years of experience in Site Reliability Engineering, DevOps, or related roles, with a track record of improving system reliability and operational maturity
  • Strong expertise in cloud platforms and modern infrastructure environments (e.g., AWS, containerized workloads, or similar ecosystems)
  • Experience with infrastructure automation and container orchestration (e.g., Terraform, Kubernetes or equivalent technologies)
  • Deep understanding of multi-tenant architecture, security principles, and data protection practices
  • Hands-on experience with observability tools and monitoring frameworks (e.g., Prometheus, Grafana or similar)
  • Experience implementing automated compliance and governance practices (e.g., SOC 2, GDPR, ISO 27001 or similar standards)
  • Strong leadership and mentoring capabilities, with the ability to influence engineering teams and drive adoption of reliability-focused practices

About KēSTA I.T.:

Our name says it all; KēSTA I.T. (Keys-to-I.T.) AND our people are our keys to our success!

KēSTA I.T. is a premier Utah-based technical staffing and consulting services firm. We specialize in temporary and permanent placement of Software, Hardware, Network, Cloud, CRM/ERP, Data, End-User support, Web and Executive / leadership-based positions on a full time and consulting basis. If you're interested in a role where top performance is rewarded, personal time is valued, and excellence is demanded at every level we want to talk to you today!

Where do you want to go? We've got the keys! ~ KēSTA I.T.



Vacancy posted 14 days ago
Similar jobs that could be interesting for youBased on the Principal Site Reliability Engineer in Beverly Hills, CA vacancy
  •  ...Principal Sre For Disney Experiences "We Power the Magic!" That's our motto at Disney Experiences (DX). Our team creates...  ...Disney Vacation Club. This role sits in the Commerce Site Reliability Engineering (SRE) specifically supporting Ecommerce, Consumer Products... 
    Principal
    Work experience placement
    Worldwide

    The Walt Disney Studios

    Glendale, CA
    1 day ago
  •  ...Company Description Aerospace / Finance Job Description PRINCIPAL SOFTWARE ENGINEER (ARCHITECT)  Visa Candidates Welcome Job Description Responsibilities: • Maintain and improve the functionality and performance of existing Windows and WCF services... 
    Principal
    Full time

    Direct Staffing Inc

    Culver City, CA
    18 hours ago
  •  ...systems that power personalization, content generation, computer vision, and decision automation. We’re looking for a hands-on AI Engineer who loves to build. You’ll design and deploy bespoke, high-impact AI solutions that stretch across product design, marketing,... 
    Principal
    Full time

    Alo Yoga

    Beverly Hills, CA
    18 hours ago
  • $165k - $265k

     ...is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER (STARLINK)At SpaceX we’re leveraging our experience in building rockets and spacecraft to deploy Starlink, the world’s most... 
    Suggested
    Permanent employment
    Temporary work
    Worldwide
    Weekend work

    SpaceX

    Hawthorne, CA
    3 days ago
  • $165k - $230k

     ...is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER (STARSHIELD)Starshield leverages SpaceX’s Starlink technology and launch capability to support national security efforts.... 
    Suggested
    Permanent employment
    Temporary work
    Immediate start
    Weekend work

    SpaceX

    Hawthorne, CA
    1 day ago
  • $165k - $265k

     ...is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER - TOP SECRET CLEARANCE (STARLINK)At SpaceX we’re leveraging our experience in building rockets and spacecraft to deploy Starlink... 
    Permanent employment
    Temporary work
    Worldwide
    Weekend work

    SpaceX

    Hawthorne, CA
    2 days ago
  • $153k - $185k

     ...infrastructure, from in-orbit pharmaceutical processing to reliable and economical reentry capsules. Varda’s W-Series...  ...and materials science — and we’re looking for bold engineers to help us get there. As a Senior Site Reliability Engineer, you'll be critical in building,... 
    Permanent employment
    Full time
    Immediate start
    Relocation package
    Flexible hours
    Weekend work

    Varda Space Industries

    El Segundo, CA
    18 hours ago
  • $125k - $145k

     ...SpaceX is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SITE RELIABILITY ENGINEER - TOP SECRET CLEARANCEAs a Site Reliability Engineer, you will design, develop, and test key aspects of an in-house... 
    Permanent employment
    Temporary work
    Weekend work

    SpaceX

    Hawthorne, CA
    4 days ago
  •  ...your big ideas, and your desire to team up with some of the best and brightest in technology and entertainment. The RoleThe Site Reliability Engineer (SRE) II is responsible for designing, implementing, and maintaining scalable and reliable systems and applications. Focus... 
    Full time
    Local area
    Worldwide
    Flexible hours

    AXS Group

    Los Angeles, CA
    4 days ago
  • $125k - $145k

     ...SpaceX is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SITE RELIABILITY ENGINEER, GNCSpaceX’s mission is to make humanity multiplanetary by developing fully and rapidly reusable launch systems capable of... 
    Permanent employment
    Temporary work
    Flexible hours
    Weekend work

    SpaceX

    Hawthorne, CA
    1 day ago
  • $125k - $150k

     ...SpaceX is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SITE RELIABILITY ENGINEER (RAPTOR)SpaceX is looking for a Site Reliability Engineer with a strong drive to solve challenging problems in the Raptor... 
    Permanent employment
    Temporary work

    SpaceX

    Hawthorne, CA
    4 days ago
  • $145k - $175k

     ...SpaceX is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SITE RELIABILITY ENGINEER (TOP SECRET CLEARANCE)As a member of the Classified IT Systems Engineering team, the Site Reliability Engineer is involved... 
    Permanent employment
    Temporary work
    Weekend work

    SpaceX

    Hawthorne, CA
    4 days ago
  •  ...A leading livestream shopping platform is seeking a Senior Software Engineer for the Logistics Platform team. This role focuses on improving logistical data systems, enhancing buyer and seller experience, and fostering collaboration across departments. Ideal candidates... 
    Remote work

    Whatnot

    Los Angeles, CA
    4 days ago
  • $164k - $270k

     ...for the 21st century and beyond.The Role What You’ll DoOwn the reliability of our robotics systems, from PLCs through ROS2/middleware to...  ...remediation.Partner with controls, robotics, and platform engineering teams to bake reliability in early. Review designs, develop SLOs... 
    Permanent employment
    Full time
    Local area
    Flexible hours

    Hadrian

    Los Angeles, CA
    3 days ago
  • $30.53 - $56.48 per hour

    Job Title:Associate Site Reliability EngineerRequisition ID:R027696Job Description:Job Title: Associate Site Reliability EngineerReporting...  ...TechnologyLocation: Santa Monica, CaOverviewThe Associate Site Reliability Engineer helps keep Marketing Technology services reliable, observable... 
    Hourly pay
    Full time
    Temporary work
    Part time
    Internship
    Local area
    Worldwide
    Relocation package

    Activision

    Santa Monica, CA
    1 day ago
  • $141k

     ...be a part of our journey! About the role We are committed to providing our customers with reliable and secure services so we are expanding our central Site Reliability Engineering team. You will be responsible for building and leading processes to ensure the reliability... 
    Local area
    Remote work
    Home office
    Flexible hours

    GrabJobs

    Los Angeles, CA
    1 day ago
  • $140k - $180k

     ...fundamentally different class of spacecraft. Engineered to survive the harshest radiation...  ...create highly available, deployable, and reliable products Reduce operational toil through...  ...experience in Software Engineering, Site Reliability Engineering or DevOps ~... 
    Permanent employment
    Shift work

    K2 Space

    Los Angeles, CA
    18 hours ago
  • $164k - $270k

    Hadrian - Manufacturing the FutureHadrian is building autonomous factories that help aerospace and defense companies manufacture rockets, satellites, jets, and ships up to 10x faster and up to 2x cheaper. By combining advanced software, robotics, and full-stack manufacturing...
    Permanent employment
    Full time
    Local area
    Remote work
    Flexible hours

    Hadrian

    Los Angeles, CA
    4 days ago
  • $122.8k - $184.2k

     ...Systems is seeking a Communications Capability Lead Sr. Principal Systems Engineer (level 4) to join its team. This position is in Roy UT, Bellevue...  ...various stakeholder communitiesProven experience leading Reliability, Availability, Maintainability, and Cost (RAM-C)... 
    Principal
    Full time
    Relocation package
    Monday to Thursday
    Shift work

    Northrop Grumman

    Manhattan Beach, CA
    2 days ago
  • $142.2k - $213.4k

     ...enabling solutions for global security. Our Engineering and Sciences (E&S) organization pushes...  ...Mission Systems is searching for a Sr. Principal Engineer Embedded Software to support...  ...a reality. This position will serve on-site on the Sentinel Program, at our Woodland... 
    Principal
    Full time
    Relocation package
    Shift work

    Northrop Grumman

    Los Angeles, CA
    18 hours ago
  •  ...scale, we invite you to bring your talents to Zscaler to help shape the future of cybersecurity. Role We are looking for a Site Reliability Engineer-SkillBridge Intern (San JosA Ca or Bellevue WA) to join our Zero Trust Exchange team. This is a remote role based in San... 
    Internship
    Work at office
    Local area
    Remote work
    Worldwide

    GrabJobs

    Glendale, CA
    4 days ago
  • $130k - $180k

     ...alongside some of the most experienced and innovative leaders and engineers in the field. Where we work Headquartered in Amsterdam and...  ...an in-house AI R&D team. The role Nebius is looking for a Site Reliability Engineer in Hardware Infrastructure team. You’re welcome to... 
    Temporary work
    Work at office
    Immediate start
    Remote work
    Flexible hours

    GrabJobs

    Los Angeles, CA
    1 day ago
  • $32 - $35 per hour

     ...and assignment.) Key Responsibilities: In this role, you will help ensure the reliability, performance, and stability of key restaurant-facing platforms by working closely with engineering and infrastructure teams. You will use observability tools such as DataDog, Grafana... 
    Contract work
    Local area
    Immediate start

    Pyramid Consulting

    Los Angeles, CA
    2 days ago
  • $135.8k - $213.4k

     ...think is impossible. Our employees are not only part of history, they're making history.Northrop Grumman is seeking a Sr Principal Software Engineer. This position will be located at our Defense Systems Sector in Manhattan Beach (CA), Hollywood (MD), Linthicum (MD), or... 
    Principal
    Full time
    Relocation package
    Shift work

    Northrop Grumman

    Manhattan Beach, CA
    3 days ago
  • $200k - $285k

     ...possible, with the ultimate goal of enabling human life on Mars.PRINCIPAL SOFTWARE ENGINEER (PLATFORM TEAM) The Platform Team builds the foundational...  ...prompts, skills, and infrastructure, this team unlocks reliable, high-impact AI capabilities across the entire... 
    Principal
    Permanent employment
    Temporary work

    SpaceX

    Hawthorne, CA
    2 days ago
  • $200k - $240k

     ...new era demands a fundamentally different class of spacecraft. Engineered to survive the harshest radiation environments and to fully...  ...attitude control systems, and power systems to ensure safe and reliable operation of the vehicle. In your first 6 months you will developcore... 
    Principal
    Permanent employment
    Full time
    Shift work

    K2 Space

    Los Angeles, CA
    18 hours ago
  • $200k - $260k

     ...functionality that is part of any Demand Side Platform (DSP) such as Viant DSP. You will serve as the technical lead working with a team of engineers to oversee the design and implementation of these interconnected services, ensuring that the system evolves to keep pace with the... 
    Principal
    Full time
    Work experience placement
    Local area

    Viant Technology

    Los Angeles, CA
    18 hours ago
  • $228.7k - $306.7k

    Job Posting Title:Senior Principal Machine Learning Engineer, Ad PlatformsReq ID:10150390Job Description:Technology is at the heart of Disney’s past, present, and future. Disney Entertainment and ESPN Product & Technology is a global organization of engineers, product developers... 
    Principal
    Full time

    Hulu

    Santa Monica, CA
    1 day ago
  • $200k - $285k

     ...possible, with the ultimate goal of enabling human life on Mars.PRINCIPAL SOFTWARE ENGINEER, CONTINUOUS INTEGRATION (STARSHIP)As a Principal Software...  ...of engineers to integrate their changes and deliver reliable software to all Starship vehicles and launch pads on a continuous... 
    Principal
    Permanent employment
    Temporary work
    Weekend work

    SpaceX

    Hawthorne, CA
    3 days ago
  • $114k - $171k

     ...are not only part of history, they're making history.Northrop Grumman Aeronautics Systems Sector has an opening for a Principal/Sr Principal Engineer Software onsite in Rancho Bernardo or El Segundo, CaliforniaAccomplishJoin our Research and Advanced Design (R&AD) Division... 
    Principal
    Full time
    Local area
    Relocation package
    Flexible hours
    Shift work

    Northrop Grumman

    El Segundo, CA
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Principal Site Reliability Engineer. Be the first to apply!