Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Site Reliability Engineer (SRE) - SecOps

Arkenstone Defense

About Us

At Arkenstone Defense, we empower defense tech startups with the tools, infrastructure, and compliance solutions they need to become successful prime contractors. Our mission is to remove barriers and help innovators grow - from day one to becoming a trusted prime for the U.S. Government.

We're early, we're lean, and we're building something that actually matters. The people who do well here aren't waiting to be told what to do; they see a gap and fill it.

Overview

We are seeking a highly motivated Systems Reliability Engineer (SRE) to lead the design and implementation of operational excellence across our multi-cloud environments. This role is central to ensuring the scalability, reliability, and performance of our products running in AWS, Azure, and GCP infrastructure. 

As the lead SRE, you will own the uptime, observability, and system resilience for our critical services. This includes driving architecture decisions, automation practices, and incident response strategies—working closely with the product owner(s), developer teams, and security operations teams. 

What You’ll Do

  • Design, implement, and own the infrastructure reliability strategy across AWS, Azure, and GCP
  • Champion observability by developing and maintaining effective logging, monitoring, and alerting systems 
  • Define and enforce SLOs/SLAs for critical systems and services 
  • Lead efforts in performance tuning, system hardening, capacity planning, and disaster recovery 
  • Own the incident management lifecycle: from detection to postmortem and root cause analysis 
  • Automate deployment, scaling, and recovery workflows to reduce manual toil
  • Contribute to infrastructure as code (Terraform, ARM templates, CloudFormation, etc.) 
  • Act as a mentor and technical leader to junior engineers and cross-functional partners 
  • Drive a culture of accountability, ownership, and continuous improvement 
  • Perform any other related duties as required or assigned.

Requirements

  • 5+ years of experience in SRE, DevOps, or infrastructure engineering roles
  • Proven track record of operating large-scale systems in multi-cloud environments, with hands-on expertise in AWS and GCP 
  • Strong knowledge of cloud-native architecture, container orchestration with Kubernetes, and CI/CD pipelines 
  • Proficient in scripting (Python, Bash, etc.) and infrastructure automation tools (e.g., Terraform) 
  • Experience with monitoring and observability platforms (e.g., Prometheus, Grafana, Datadog, ELK) 
  • Excellent problem-solving skills with the ability to manage incidents and make sound decisions under pressure 
  • Clear communicator capable of translating technical concepts to mixed audiences and participating in customer discussions 

Who You Are

  • The Security Builder: You don't just consume security tools — you extend and improve them. You're energized by the opportunity to make analysts faster and compliance more automated. 
  • Quality Over Speed: You write code that lasts. You push for clean interfaces, good documentation, and tests — even in a fast-moving environment. 
  • Security-Minded Developer: You treat security as a first-class requirement, not an afterthought. You're comfortable reading CVEs, threat models, and compliance controls. 
  • Cross-Functional Partner: You can work fluidly with security analysts, engineers, and compliance professionals, translating needs into reliable software. 

Mission Alignment

We are a Defense-focused company supporting sensitive and cleared workforces. The Site Reliability Engineer (SRE) - SecOps will embrace our commitment to operational excellence, compliance rigor, and a world-class employee experience.

Physical Requirements

  • Prolonged periods of sitting at a desk and working on a computer
  • Must be able to lift up to 15 pounds at times
  • May require occasional travel to office locations or client sites
  • Ability to communicate effectively in written and verbal form

Benefits for working with us!

We are committed to supporting our employees both professionally and personally. Our robust benefits package is designed to promote your well-being, growth, and work-life balance

  • Competitive Salary: Recognizing your hard work with attractive compensation and rewarding excellence.
  • Health and Wellness Programs: Including medical, dental, & vision insurance options, along with mental health support & wellness initiatives.
  • Retirement Planning: Secure your future with our flexible 401(k) plan and matching company contributions.
  • Paid Time Off & Holidays: Generous PTO, sick leave, and holiday pay to help you recharge and enjoy life outside of work.
  • Employee Assistance Program: Confidential resources for personal and professional support.
  • Professional Development: Access to training, certifications, and continuing education to foster your career growth.

We are an Equal Opportunity Employer. We celebrate diversity and are committed to creating an inclusive environment for all employees. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex (including pregnancy, gender identity, and sexual orientation), national origin, age, disability, genetic information, veteran status, or any other characteristic protected under applicable law.
Vacancy posted 9 days ago
Similar jobs that could be interesting for youBased on the Site Reliability Engineer (SRE) - SecOps in Menlo Park, CA vacancy
  • $100k - $200k

    OPPO US Research Center is seeking a skilled and proactive Site Reliability Engineer (SRE) to join our team. In this role, you will be responsible for ensuring the stability, scalability, and performance of our application systems. The ideal candidate is passionate about... 
    Suggested
    Full time

    OPPO

    Palo Alto, CA
    3 days ago
  • $61k - $101k

     ...,000 per year Requirements: We require formal training or certification in site reliability engineering, along with 5+ years of hands-on experience. We need advanced knowledge of SRE culture and principles, with proven ability to apply them in an application or... 
    Suggested
    Full time

    J.P. Morgan

    Palo Alto, CA
    16 days ago
  •  ...development, cloud infrastructure, DevOps, SRE, and platform engineering. You will test AI-generated commands,...  ...workflows for accuracy and reliability. Work with AWS, Azure, GCP, Kubernetes...  ...DevOps Cloud Infrastructure Site Reliability Engineering (SRE) Platform... 
    Suggested
    Remote job
    For contractors

    YO AI Labs

    Palo Alto, CA
    26 days ago
  • $70 - $100 per hour

     ...Job Title: Cloud SRE Engineer - Mandarin Bilingual Position Type: Contract (12 months) Location: Palo Alto, CA Salary Rate: $7...  ...team is looking for a skilled Cloud SRE Engineer to own the reliability, stability, and continuous improvement of core cloud services... 
    Suggested
    Hourly pay
    Contract work
    Temporary work
    Work experience placement

    IntelliPro Group Inc.

    Palo Alto, CA
    more than 2 months ago
  •  ...this role. JOB DESCRIPTION Elevate your engineering prowess to unprecedented levels by...  ...position yourself among the top echelon in site reliability. As a Senior Lead Site Reliability Engineer...  ...open-source projects, particularly in SRE, observability, or AI/ML domains, and... 
    Suggested

    J.P. Morgan

    Palo Alto, CA
    2 days ago
  •  ...Overview Title: Site Reliability Engineer SRE – ML platform Location: Austin, TX or Sunnyvale, CA Employment type: Full-time • Seniority: Mid-Senior level • ONLY W2 Responsibilities Continuous Deployment using GitHub Actions, Flux, Kustomize Design and implement cloud... 
    Full time

    Saransh

    Sunnyvale, CA
    2 days ago
  • $165k - $280k

     ...actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars. SR. SITE RELIABILITY ENGINEER (STARLINK) At SpaceX we’re leveraging our experience in building rockets and spacecraft to deploy Starlink, the world’s... 
    Permanent employment
    Temporary work
    Worldwide
    Weekend work

    InvestedintheMission

    Palo Alto, CA
    3 days ago
  • $276.1k - $311.4k

     ...that defines your career! The role As SRE Manager, you'll build the Vehicle...  ...defining its charter, hiring its founding engineers, establishing the operating model, and creating...  ...the technical strategy that makes reliability a first-class property of the software running... 
    Permanent employment
    Full time
    Work at office
    Work from home

    Lindus Health

    Sunnyvale, CA
    5 days ago
  •  ...About the RoleWe are building a high-performance SRE function to support one of the world’s fastest-growing...  ...inference services, powered by the Wafer-Scale Engine (WSE). This team will help deliver world-class, ultra-reliable inference infrastructure for leading model... 
    Shift work

    CEREBRAS SYSTEMS INC.

    Sunnyvale, CA
    3 days ago
  • $170k - $200k

     ...We are seeking a talented and motivated Site Reliability Engineer to join our engineering team. You will be responsible for building, maintaining...  ...background in infrastructure automation, system reliability, and a SRE mindset of continuous improvement.Key Responsibilities:... 
    Full time
    Worldwide

    Fortinet

    Sunnyvale, CA
    15 hours ago
  • $207k - $300k

    Lead a team of Software/Systems Engineers on projects for users and be directly responsible for uptime.Own end-to-end availability...  ...or Engineering.Experience with Large Language Model.Site Reliability Engineering (SRE) combines software and systems engineering to build and... 

    Google

    Sunnyvale, CA
    3 days ago
  •  ...service — and as we scale toward launch, our engineering infrastructure has to scale with us. We're hiring a Staff Site Reliability Engineer to own source control at Zoox,...  ...developer-facing tooling that reduce toil for SRE and product engineering alike, measured by build... 

    Zoox

    Foster, CA
    3 days ago
  •  ...Google Cloud is seeking a Manager, Software Engineer, Site Reliability Engineering in Sunnyvale, CA. You will lead a team focused on uptime, reliability, and scalable infrastructure, delivering automated solutions and architecting resilient systems. The role requires... 

    Jobleads-US

    Sunnyvale, CA
    2 days ago
  •  ...Google is seeking a Software Engineering Manager II in Site Reliability Engineering, based in Sunnyvale, California. This onsite role leads a team to ensure reliability and performance of critical systems, partnering with product and engineering teams to deliver scalable... 

    Epic Games

    Sunnyvale, CA
    2 days ago
  • $150.4k - $277.6k

     ...Technical Operations & Site Reliability Engineer, Customer SystemsAt Apple, Customer Experience is at the forefront of everything we do. The Customer Systems Operations team is looking for a highly skilled and motivated TechOps Engineer (Technical Operations & Site Reliability... 
    Work experience placement
    Relocation

    Apple

    Sunnyvale, CA
    3 days ago
  • $130k - $200k

     ...IXL Learning, developer of personalized learning products used by millions of people globally, is seeking a Senior Site Reliability Engineer to join our team, and help maintain the reliability and optimal performance of our products. We are seeking engineers with a passion... 
    Full time
    Work at office
    Immediate start

    IXL Learning

    San Mateo, CA
    15 hours ago
  • $169k - $338k

     ...systems that can autonomously handle complex reliability engineering workflows, predictive failure analysis,...  ...operations across any Walmart system.Site Reliability Engineering Technical...  ...Innovate in agentic AI technologies for SRE including large language models (LLMs)... 
    Full time
    Temporary work
    Part time

    Walmart

    Sunnyvale, CA
    a month ago
  •  ...Site Reliability Engineer, Data Platform - USDS Responsibilities Engage in and improve the whole lifecycle of service, from inception and design, through to deployment, operation and refinement. Ensure reliable, fault-tolerant, efficiently scalable and cost-effective data... 

    Tik Tok

    Mountain View, CA
    2 days ago
  •  ...We are seeking a Senior Database Reliability Engineer (DBRE) to design, operate, and improve reliable...  ...The role combines database engineering, site reliability engineering, Linux systems administration...  ...data systems and will work closely with SRE, application, security, and... 

    Neshent Technologies

    Los Altos, CA
    1 day ago
  •  ...design by customizing MES tool per business needs Education Requirements, Ideal Experience: Associate’s degree in Industrial Engineering or IT related field Minimum of 0-3 years’ relevant experience Experience in C#, Delphi desired Knowledge of the... 
    Work at office

    Foxconn Industrial Internet - FII

    Sunnyvale, CA
    25 days ago
  • $243.29k - $295.25k

     ...shared experiences for everyone.The Infrastructure Compute Site Reliability Engineering mission is to own and manage the successful operation of our...  ...a proven track record including at least 6 years as an SRE or Software Engineer.Fluency with high-level programming languages... 
    Full time
    Work experience placement
    H1b
    Work at office
    Local area
    Visa sponsorship
    Monday to Friday

    Roblox

    San Mateo, CA
    9 hours ago
  •  ...Senior Lead Software Engineer Be an integral part of an agile team...  ...concepts and howto apply those in reliability settings Understanding of...  ...systems and how they work in SRE environments to solve...  ...comprehensive health care coverage, on-site health and wellness centers, a... 
    For contractors

    Hackajob

    Palo Alto, CA
    1 day ago
  •  ...sides. Founded by ex-Meta product and engineering leaders, we've raised over $30M in total...  ...Engineer, Infrastructure to own the reliability, scalability, and operational excellence...  ...growing customer base, and we need a seasoned SRE to help us scale these systems safely... 
    Work at office
    Remote work
    Flexible hours
    Shift work

    Nectar Social

    Palo Alto, CA
    3 days ago
  • $135k - $200k

     ..., forecast supply chain disruptions, locate missing children, and more. The Role We are seeking a Forward Deployed Software Engineer to join a newly-formed team focused on developing advanced Command and Control (C2) Software for Autonomous Systems for use in operational... 
    Full time
    Work experience placement
    Work at office
    Remote work
    Work from home
    Relocation package

    Palantir Technologies

    Palo Alto, CA
    9 hours ago
  • $174k - $252k

     ...by pushing for changes that improve reliability and velocity.Practice sustainable...  ...Bachelor’s degree in Computer Science, Engineering, a related field, or equivalent...  ...in Computer Science or Engineering.Site Reliability Engineering (SRE) is what you get when you treat operations... 

    Google

    Sunnyvale, CA
    2 days ago
  • $115.5k - $189.75k

     ...driving by turning on-road signals and incidents into actionable engineering insights.  The Release & Triage Tooling sub-team builds AI-...  ...end-to-end workflows. Design and implement scalable, reliable internal services used by release and triage teams, ensuring maintainability... 
    Full time
    Temporary work
    Work at office
    Flexible hours

    Woven

    Palo Alto, CA
    9 hours ago
  •  ...Technologies is seeking a Streaming Data Engineer to architect, deploy, and operate large-scale...  ...with cross-functional teams to deliver a reliable streaming platform. The ideal candidate...  ...Kafka internals knowledge, strong DevOps/SRE practices, and excellent #J-18808-Ljbffr... 
    Remote job

    Bright Vision Technologies

    Redwood City, CA
    4 days ago
  • $262k - $364k

     ...within the AViD ecosystem have reliability and uptime appropriate to...  ...and performance.Build creative engineering solutions to operations and infrastructure...  ...and contribute to the cross-SRE AI Ops program, driving the...  ...in a strategic way.Site Reliability Engineering (SRE)... 

    Google

    Mountain View, CA
    2 days ago
  • $222k - $300.5k

     ...OverviewAbout the TeamIntuit's Infrastructure and Site Reliability organization owns the operational...  .... The Fintech Platform Systems Engineering team builds and operates the AWS-based...  ...engineering, product, security, and other SRE/infrastructure leaders across Intuit to... 
    Worldwide
    Shift work

    Intuit

    Mountain View, CA
    9 hours ago
  •  ...the right place.As a Principal Engineer at JPMorganChase within the...  ...with performance engineering, SRE, and quality engineering to ensure...  ...the platform operates reliably across hybrid infrastructure (...  ...comprehensive health care coverage, on-site health and wellness centers, a... 
    Shift work

    JP Morgan Chase

    Palo Alto, CA
    a month ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Site Reliability Engineer (SRE) - SecOps. Be the first to apply!