Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Staff Site Reliability Engineer - Splunk

$194k - $267k

Okta

Secure Every Identity, from AI to Human

Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence.

This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk.

Position Overview:

We are seeking a highly technical Staff Observability Site Reliability Engineer with a specialty in Splunk to own and evolve our Splunk ecosystem. In this role, you will move beyond simple monitoring to delivering a world class, comprehensive, scalable Observability Platform that enables our SRE teams and business partners. You will treat infrastructure as code —utilizing Terraform and strong coding proficiency in Go, Python, or Ruby —to automate the deployment of agents and collectors across complex distributed systems.

Key Responsibilities

  • Automated Infrastructure: Design, build, and maintain scalable observability infrastructure using tools like Terraform.
  • Splunk Engineering: Optimize the collection, processing, and storage of log data to ensure high reliability and low latency of our Splunk services
  • Incident Response: Participate in on-call rotations and lead post-incident reviews to drive systemic improvements and "observability-driven development."
  • Automation: Eliminate "toil" by automating the deployment and scaling of observability agents and collectors.

Required Skills & Experience (The Essentials)

Log Management: Minimum 5+ Experience scaling and managing Splunk Cloud at scale (1000+ SVCs), including Workload Management (WLM) and HEC optimization. Visualization: Expertise in creating intuitive, actionable Splunk dashboards that correlate data across multiple sources.
SRE Mindset: Minimum 5+ years of experience in an SRE, DevOps, or Systems Engineering role with a focus on high-availability systems.

  • Programming Proficiency: Strong coding skills in SPL , Go for building internal tools and automating workflows.
  • Distributed Systems: Deep understanding of Linux internals, networking (TCP/IP, DNS, Load Balancing), and container orchestration (Kubernetes/EKS).
  • Problem Solving: A data-driven approach to debugging complex, cross-service performance bottlenecks.

Bonus Skills (The "Nice-to-Haves")

  • Telemetry Standards: Hands-on experience with OpenTelemetry (OTel), Vector, or similar frameworks for instrumenting applications.
  • Charge-back app: Experience in implementing Splunk charge-back app for usage reporting 

Cloud Platforms: Experience managing observability native tools within AWS or GCP.

Additional requirements:

  • This position requires the ability to access federal environments and/or have access to protected federal data.  As a condition of employment for this position, the successful candidate must be able to submit documentation establishing U.S. Person status (e.g. a U.S. Citizen, National, Lawful Permanent Resident, Refugee, or Asylee. 22 CFR 120.15) upon hire.
  • This person must attend in person onboarding in our San Francisco office the first week of employment. 

#LI-MM

#LI-Hybrid
P14596_3372199

The annual base salary range for this position for candidates located in the San Francisco Bay area is between: $194,000—$267,000 USD

Below is the annual base salary range for candidates located in California (excluding San Francisco Bay Area), Colorado, Illinois, New York and Washington. Your actual base salary will depend on factors such as your skills, qualifications, experience, and work location. In addition, Okta offers equity (where applicable), bonus, and benefits, including health, dental and vision insurance, 401(k), flexible spending account, and paid leave (including PTO and parental leave) in accordance with our applicable plans and policies. To learn more about our Total Rewards program please visit: .   

The annual base salary range for this position for candidates located in California (excluding San Francisco Bay Area), Colorado, Illinois, New York, and Washington is between: $174,000—$239,000 USD

The Okta Experience

We are intentional about connection. Our global community, spanning over 20 offices worldwide, is united by a drive to innovate. Your journey begins with an immersive, in-person onboarding experience designed to accelerate your impact and connect you to our mission and team from day one.

Okta is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, ancestry, marital status, age, physical or mental disability, or status as a protected veteran. We also consider for employment qualified applicants with arrest and convictions records, consistent with applicable laws.

If reasonable accommodation is needed to complete any part of the job application, interview process, or onboarding please  use this Form to request an accommodation.

Notice for New York City Applicants & Employees: Okta may use Automated Employment Decision Tools (AEDT), as defined by New York City Local Law 144, that use artificial intelligence, machine learning, or other automated processes to assist in our recruitment and hiring process. In accordance with NYC Local Law 144, if you are an applicant or employee residing in New York City, please  click here to view our full NYC AEDT Notice.
Vacancy posted 22 days ago
Similar jobs that could be interesting for youBased on the Staff Site Reliability Engineer - Splunk in Washington DC vacancy
  •  ...7 environment Technical expertise: Analysis of issues via Splunk (including Splunk APM and Splunk O11y), AppDynamics, Grafana...  ...a new job is posted. Sign in to set job alerts for “Senior Site Reliability Engineer” roles. Bellevue, WA $204,000.00-$259,000.00 1 day ago Seattle... 
    Splunk
    Contract work
    Remote work

    Signature IT World Inc

    Washington DC
    5 days ago
  •  ...protect our country from threats. Job Description SITE RELIABILITY ENGINEER (SRE) Own your opportunity. Make your impact As a Site...  ...enterprise monitoring platforms (e.g., SolarWinds, SCOM, Splunk, Nagios, ELK) ~ Strong understanding of ITIL/ITSM workflows... 
    Splunk

    General Dynamics Information Technology

    Washington DC
    a month ago
  • $100k - $110k

     ...Instructional Design and Training, Software Engineering and IT Support Services to improve the...  ...with MKS2. Position Title: Site Reliability Systems Engineer (REMOTE) Program:...  ...Tools like SolarWinds, Dynatrace, and Splunk will be part of your daily workflow, giving... 
    Splunk
    Work at office
    Remote work

    MKS2 Technologies

    Washington DC
    23 days ago
  • $194k - $267k

     ...s talk. We are seeking a highly technical Observability Site Reliability Engineer with a specialty in Google Cloud, to own and expand our Observability...  ...data to ensure high reliability and low latency of our Splunk and Grafana services Incident Response: Participate in... 
    Splunk
    Permanent employment
    Local area
    Worldwide
    Flexible hours

    Okta

    Washington DC
    4 hours ago
  • $174k - $239k

     ...we partner across functions to drive scale, reliability, and innovation through technology. The Staff Site Reliability Engineer Opportunity Okta Federal, Inc. is looking for...  ...with monitoring tools, especially Splunk, CloudWatch, and the Grafana stack. ~ Experience... 
    Splunk
    Local area
    Worldwide
    Flexible hours

    Okta

    Washington DC
    10 days ago
  • $175k - $250k

     ...Senior Cloud Infrastructure Engineer Location: San Francisco, CA. Remote unavailable. Modality: On‑Site only. Must live within commuting distance of San Francisco or...  ...while ensuring scalability, performance, and reliability across environments. What You’ll Do Design, build... 
    Full time
    Remote work
    Relocation
    Relocation package

    The Recruiting Guy

    Washington DC
    5 days ago
  •  ...A leading security infrastructure firm in Washington, D.C. is seeking a hands-on Site Reliability Engineer (SRE) with expertise in Kubernetes and cloud infrastructure. The role emphasizes total ownership of security infrastructure while defending against advanced threats... 

    Cyrad Solutions LLC

    Washington DC
    14 hours ago
  • $149.4k - $202k

     ...Senior Software Engineer- Site Reliability Engineering (SRE) DC, MD, VA, CA The Site Reliability Engineering discipline at Noctua Technology, LLC is a strategic force driving digital transformation. We treat operations as a software engineering challenge, focusing... 
    Remote work

    Noctua Technology

    Washington DC
    2 days ago
  • $95k - $171k

     .... Opportunities exist to focus on GPU infrastructure, Kubernetes, and ensuring reliability for AI workloads within Akamai's serverless inference platform. As an Site Reliability Engineer II, you will be responsible for: Building and maintaining dashboards, alerts... 
    Permanent employment
    Work experience placement
    Work at office
    Remote work
    Work from home
    Worldwide
    Flexible hours

    Akamai

    Washington DC
    4 days ago
  • $165k - $230k

     ...actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars. SR. SITE RELIABILITY ENGINEER (STARSHIELD) Starshield leverages SpaceX’s Starlink technology and launch capability to support national security efforts... 
    Permanent employment
    Temporary work
    Immediate start
    Weekend work

    SpaceX

    Washington DC
    1 day ago
  • $106.3k - $221.1k

     ...more. Join us to drive positive, lasting change that moves missions and the government forward! Job Description The Site Reliability Engineer will ensure the reliability, performance, and scalability of the Client System. The engineer will define and track Key... 
    Live in
    Work at office
    Local area

    Accenture

    Arlington, VA
    1 day ago
  • $131k - $227.13k

     ...Description: The 1LMX MES COE is seeking an engineer who will own infrastructure‑as‑code, cloud platform, and reliability for the Apriso environment on AWS. This role blends full‑stack development, DevOps, and Site Reliability Engineering (SRE) practices to deliver... 
    Full time
    Temporary work
    Work experience placement
    Work at office
    Remote work
    Relocation
    Flexible hours
    Shift work
    3 days per week

    Lockheed Martin Corporation

    Bethesda, MD
    6 days ago
  • $135k - $150k

    Senior Site Reliability Engineer Job number: 884 This is a remote position. Ad Hoc is a technology company that empowers organizations to deliver scalable, impactful digital services. Using modern, agile methods, our team creates products that meet people's... 
    Remote work
    Flexible hours

    Ad Hoc LLC

    Silver Spring, MD
    4 days ago
  • $153k - $185k

     ...Senior Site Reliability Engineer El Segundo, California, United States About Varda Low Earth orbit is open for business. Varda is accelerating...  ...suited to accomplishing this goal, with leadership and staff comprised of veterans from SpaceX, Blue Origin, major... 
    Permanent employment
    Full time
    Immediate start
    Relocation package
    Flexible hours
    Weekend work

    Varda Space Industries

    Washington DC
    2 days ago
  • $121.4k - $218.6k

     ...solve complex challenges? Do you have a passion for automation and building systems that scale? Join our highly skilled Site Reliability Engineering team! Our team designs, develops, and manages applications and infrastructure that support Akamai Cloud's products and... 
    Work experience placement
    Work at office

    Akamai

    Washington DC
    2 days ago
  •  ...Qualifications: ~10+ years of overall experience in IT including, with hands-on Development and Systems engineering background ~3-5 years of experience in a Site Reliability Engineering role ~ Experience with Enterprise Cloud transformation efforts ~ Experience with... 
    Temporary work
    Immediate start

    Samprasoft

    Washington DC
    1 day ago
  •  ...Join to apply for the Lead Site Reliability Engineer role at Bridge Defense About Bridge Defense Bridge Defense is redefining how modern defense technology is delivered. Based in Washington, D.C., we are built for the dynamic mission environment facing the Department of... 
    Full time
    Contract work
    Remote work
    Relocation

    Bridge Defense

    Washington DC
    2 days ago
  • Log Management Engineer Looking for a log management engineer. The candidate will be responsible for log standardization and optimization. Must have in depth knowledge of Splunk, Cribl, syslog, HEC, Azure Eventhub, AWS Kinesis, or similar.
    Splunk

    Samprasoft

    Washington DC
    13 hours ago
  • $51.9 per hour

     ...This job is responsible for the reliability, availability, and...  ...efficiency. This role blends software engineering, clinical engineering, and...  ...cross-functionally with AHN site leaders and teams to navigate...  ..., performance management and staff productivity.Plan, organize,... 
    For contractors
    Local area

    Highmark Health

    Washington DC
    5 days ago
  •  ...protocols Tracking and documenting on-site incident response activities and providing...  ..., WireShark, Sleuth Kit/ Autopsy, Snort, Splunk or other EDR Tools (Crowdstrike, Carbon...  ...Computer Science, Cybersecurity, Computer Engineering, or related degree; or HS Diploma and 10+... 
    Splunk
    Immediate start
    Remote work

    Solutions³ LLC

    Arlington, VA
    2 days ago
  •  ...Solutions Engineer – Arlington, VA We deliver essential technology services in support of missions critical to national security and economic...  ...Experience integrating and operating SIEM/SOAR platforms (Splunk, Cortex, QRadar, Sentinel). Familiarity with EDR, IDS/IPS,... 
    Splunk
    Shift work

    Chenega Corporation

    Arlington, VA
    4 days ago
  •  ...protocols Tracking and documenting on-site incident response activities and providing...  ..., WireShark, Sleuth Kit/ Autopsy, Snort, Splunk or other EDR Tools (Crowdstrike, Carbon...  ...Computer Science, Cybersecurity, Computer Engineering, or related degree; or HS Diploma and 10+... 
    Splunk
    Full time
    For contractors
    Immediate start
    Remote work

    Solutions³ LLC

    Arlington, VA
    9 days ago
  • $100k

     ...Maximus is currently seeking a Cloud Platform Engineer. This is a remote position....  ...real-time communications systems, ensuring reliability, performance, and operational continuity...  ...tools (e.g., Azure Monitor, Log Analytics, Splunk, or equivalent). ~ Experience... 
    Splunk
    Contract work
    Remote work

    MAXIMUS

    Washington DC
    2 days ago
  •  ...Java Engineer Project Details: Fed now – Developing for the payment industry, enterprise money movement. Deploying to the internal...  ...GraphQL Caching experience Experience with tools like Splunk, AppDynamics, Kibana Required Skills : Core Java with... 
    Splunk
    Contract work

    Software Technology Inc

    Washington DC
    5 days ago
  •  ...EnCase, FTK, X-Ways, SIFT, Volatility, Sleuth Kit/Autopsy Wireshark, Splunk, Snort, or EDR tools (CrowdStrike, Carbon Black, SentinelOne) Experience conducting malware reverse-engineering and all-source research Understanding of threat actor TTPs and advanced... 
    Splunk
    Remote work

    Argo Cyber Systems

    Arlington, VA
    9 days ago
  • $190k - $240k

    Belay Technologies is seeking a Senior Software Engineer (SWE3). This position is designed for a Fullstack Developer with required skillset...  ...Java StE/StN Desired Skills AWS familiarity Splunk Apache (Hadoop, Spark) Benefits 8 weeks paid leave - 4 weeks... 
    Splunk

    Belay Technologies

    Laurel, MD
    2 days ago
  • $191k - $253k

     ...Senior Software Engineer, Developer Platform Anduril Industries is a defense technology...  ...downstream pays for it - but when it's reliable and zippy, life is awesome. In this role...  ...Datadog, Prometheus, Grafana, ELK stack, Splunk). US Salary Range $191,000 - $253... 
    Splunk
    Full time
    Work experience placement
    Immediate start

    anduril

    Washington DC
    13 hours ago
  •  ...Ford Pro Software Engineer While others are taking a technology-first approach, Ford Pro is putting people first. We are building a...  ...Programming (XP) practices, JIRA, Flyway, Azure, Jenkins, MS SQL, Splunk Preferred Qualifications: Deep understanding of... 
    Splunk

    Software Technology Inc

    Washington DC
    5 days ago
  •  ...Forward Deployed Software Engineer iboss is a cloud delivered Zero Trust network security...  .... Key Responsibilities On-Site Deployment & Integration Deploy, configure...  ...Experience with SIEM/SOAR integrations (Splunk, Microsoft Sentinel, etc.) Relevant certifications... 
    Splunk
    Worldwide

    iboss

    Washington DC
    5 days ago
  •  ...Expertise in Pega PRPC 8.6 and above, including case management, rules engine, and UI development. Strong proficiency in configuring and...  ...integrations. Optional Skills (Nice-to-Have): Experience with Splunk for monitoring and analytics. Knowledge of MongoDB for NoSQL database... 
    Splunk

    TechDigital Group

    Washington DC
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Staff Site Reliability Engineer - Splunk. Be the first to apply!