Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior Software Engineer- Site Reliability Engineering (SRE)

$149.4k - $202k

Noctua Technology

The Site Reliability Engineering discipline at Noctua Technology, LLC is a strategic force driving digital transformation. We treat operations as a software engineering challenge, focusing on the seamless integration, scalability, and long-term reliability of cloud native systems. Our SREs don’t just manage infrastructure; they build it using Infrastructure as Code (IaC), monitor it through advanced observability stacks, and protect it by engineering for failure. We work closely with clients to bridge the gap between development and operations.

We are seeking a highly experienced and autonomous Senior Site Reliability Engineer (SRE) to join our dynamic team. As a technical leader, you will define the strategy and apply advanced software engineering principles to operations, focusing on the architecture, reliability, and long-term performance of large-scale production systems. You will play a crucial role in reducing toil through automation, defining and monitoring Service Level Objectives (SLOs), and implementing best practices for system stability and incident response. This role requires working with modern cloud technologies to ensure the high availability and efficiency of applications and infrastructure.

  • Location : Primarily Remote. Candidates must be based in CA or DC Metro Area for proximity to project and client teams.
  • Security Clearance Requirement: Applicants must be US citizens and eligible to obtain and maintain an active Secret security clearance or above.

Key Responsibilities

Site Reliability Engineering

  • Drive the definition and adoption of SLIs and SLOs across multiple services or entire platforms, ensuring alignment with business goals.
  • Design and architect Infrastructure as Code (IaC) solutions for large-scale, complex environments, establishing standards and best practices.
  • Implement and manage containerized and serverless architectures using Docker, Kubernetes, and cloud-native services, focusing on performance and error budgets.
  • Build and maintain reliable and self-healing CI/CD pipelines to automate deployments and improve development workflows.

Toil Reduction and Incident Management 

  • Implement and refine comprehensive monitoring, alerting, and logging to detect and address performance and availability issues proactively.
  • Lead the strategic effort to eliminate toil, identifying and championing major automation projects that deliver significant organizational efficiency.
  • Lead high-severity incident response and coordinate blameless postmortems for major outages, driving the resulting remediation and systemic improvements.

Testing and Service Resiliency

  • Implement cloud security best practices, including identity and access management (IAM), encryption, and compliance controls.
  • Proactively identify and address system weaknesses and ensure performance under stress.
  • Support disaster recovery and high availability strategies through backup and failover planning.

Collaboration and Knowledge Sharing

  • Serve as a primary SRE liaison for development teams, influencing application architecture and design to meet reliability and scalability targets from inception.
  • Create and maintain documentation for cloud architectures, deployment processes, and best practices.
  • Contribute to internal knowledge-sharing initiatives, ensuring continuous learning within the team.

Stakeholder Communication

  • Act as a subject matter expert and trusted advisor to clients and internal leadership on cloud infrastructure, reliability strategy, and Service Level Agreement (SLA) negotiations.
  • Act on client feedback to refine and enhance cloud solutions.
  • Conduct training and knowledge-sharing sessions to help clients manage their cloud environments effectively.

Continuous Learning and Innovation

  • Stay updated on the latest developments in cloud infrastructure and technology trends.
  • Drive innovation by proposing and implementing new techniques and technologies.

Qualifications

  • 5+ years of experience in site reliability engineering, cloud engineering, or related fields.
  • Strong software engineering skills with an emphasis on writing clean, modular, and maintainable code, specifically for automation and system management.
  • Deep experience with Infrastructure as Code (IaC) tools like Terraform or CloudFormation.
  • Deep experience with containerization and orchestration tools like Docker and Kubernetes.
  • Deep knowledge of networking concepts, cloud security best practices, and identity management.
  • Experience with programming or scripting languages such as Python, Bash, or Go.
  • Experience with CI/CD pipelines and DevOps methodologies.
  • Strong problem-solving skills and the ability to troubleshoot complex cloud environments.
  • Demonstrated ability to influence technical decision-making across organizational boundaries

Preferred qualifications:

  • Bachelor's or advanced degree in Computer Science or a related field.
  • Any of the below cloud certifications:
    • Google Cloud Professional Cloud Architect
    • Google Cloud Professional Cloud DevOps Engineer
    • AWS Certified Solutions Architect
    • AWS Certified Developer
    • AWS Certified SysOps Administrator
  • CompTIA Security+ certification or an equivalent DoD 8140/8570 IAT Level II baseline certification.

Salary Range : $149,400 - $202,000

Vacancy posted 5 days ago
Similar jobs that could be interesting for youBased on the Senior Software Engineer- Site Reliability Engineering (SRE) in United States vacancy
  • $101k - $161k

     ...artificial intelligence, and software-defined networking to...  ...awards, such as Best Engineering Team, Best Company for...  ...WithWe’re looking for Site Reliability Engineers to join our...  ...(CVaaS) global SRE team. SREs at Arista combine...  ...level: Mid-Senior LevelIndustry: Computer... 
    Senior

    Arista Networks

    Santa Clara, CA
    1 day ago
  •  ...Opportunity: We are looking for a skilled engineer with disciplines that incorporate aspects of software systems engineering and operations. We...  ...-driven approaches to observability and reliability. What you’ll do: • Evangelize SRE mindset and solve problems through systematization... 
    Senior

    Mindlance

    Austin, TX
    3 days ago
  •  ...Senior Site Reliability Engineer (SRE) Our client is a global technology consulting and digital solutions company that enables enterprises across industries to reimagine business models, accelerate innovation, and maximize growth by harnessing digital technologies.... 
    Senior
    Local area

    E-Solutions

    New York, NY
    3 days ago
  •  ...Description The Senior Site Reliability Engineer (SRE) will implement, secure, and operate the cloud infrastructure that supports CenCore Group...  ..., and reliability best practices. Partner with software engineering and product teams to improve application performance... 
    Senior
    Work at office
    Remote work

    CenCore

    United States
    4 days ago
  • $185k - $200k

    Back to All JobsSenior Site Reliability Engineer (SRE) Dayton, OH (Remote) full time Top Secret (TS)...  ...Position Overview Metronome is seeking a Senior Site Reliability Engineer (SRE) to...  ...cloud/platform engineering, DevOps, software engineering, or a related discipline.... 
    Senior
    Full time
    Remote work

    Metronome LLC

    Dayton, OH
    5 days ago
  • $120k - $175k

     ...Senior Site Reliability Engineer (SRE) Atlanta, GA preferred, Remote At PrizePicks, we are the fastest-growing sports company in North America, as recognized by Inc. 5000. As the leading platform for Daily Fantasy Sports, we cover a diverse range of sports leagues... 
    Senior
    Full time
    Remote work
    Work visa
    Flexible hours

    PrizePicks

    Atlanta, GA
    4 days ago
  •  ...Senior Site Reliability Engineer (SRE) We are looking for a highly experienced and driven Senior Site Reliability Engineer to join our forward-thinking...  ...members and Mirantis customers to deliver high-quality software and services. As a senior engineer, you will work... 
    Senior
    Remote work

    Mirantis

    United States
    4 days ago
  •  ...build safer, more resilient organizations. The Role: As a Senior Site Reliability Engineer (SRE) at Dune Security, you will play a critical role in...  ...and prevent bot attacks. Establish best practices for software reliability, incident response, and fault tolerance. Lead... 
    Senior
    Full time
    Work at office

    Dune Security

    New York, NY
    1 day ago
  • $119.8k - $234.7k

     ...yearEmployment type: Full-TimeWork site: 3 days / week in-officeRole type: Individual...  ...Cloud Hardware, and Infrastructure Engineering (SCHIE) is the team behind Microsoft...  ...engineering organization. As a Senior Linux SRE (Site Reliability Engineer), your main focus will be... 
    Senior
    Ongoing contract
    Permanent employment
    Work at office
    Local area
    Worldwide
    3 days per week

    Microsoft

    Hillsboro, OR
    1 day ago
  •  ...About The Role: We're looking for a Senior Site Reliability Engineer to help us mature and scale the...  ..., and who treats infrastructure like software. You'll have significant ownership over...  ...What You Bring: ~​​6+ years in SRE, DevOps, or infrastructure... 
    Senior
    Remote work
    Flexible hours

    Dental Intelligence

    United States
    3 days ago
  • $60 - $80 per hour

     ...Job Title: Senior Site Reliability Engineer (SRE) - Hybrid Duration (Contract): 6 Months Client Location: Austin, TX Location Preference...  ...-scale applications and platforms. This role combines software engineering, infrastructure operations, and AI/ML-... 
    Senior
    Hourly pay
    Contract work

    Smart IMS Inc

    Austin, TX
    3 days ago
  • $175k - $215k

     ...exciting experiences.Sr. Manager, Site Reliability Engineer provides strategic leadership across multiple SRE teams and their managers,...  ...and innovation. Influences senior internal and external stakeholders...  ...of how SRE integrates with software development, security, and... 
    Senior

    Disney Interactive

    Orlando, FL
    5 days ago
  •  ...unwavering security to responsibly propel the global lottery industry ever forward.Position SummaryWe are looking for a skilled Site Reliability Engineer (SRE) to enhance the stability, performance, and reliability of our production systems. The SRE will work closely with... 
    Senior
    Permanent employment
    Full time
    Work experience placement
    Local area

    Scientific Games Corporation

    Alpharetta, GA
    4 days ago
  • $160k - $185k

     ...And we’re just getting started!OverviewThe Sr. Manager, Site Reliability Engineering (SRE) leads the strategy, execution, and continuous improvement...  ...support. This role operates at the intersection of software engineering and infrastructure, driving automation, reducing... 
    Senior
    Work at office
    Local area
    Remote work
    Work from home

    Planet Fitness

    Hampton, NH
    5 days ago
  • $1,000 per month

     ...building the AI infrastructure our engineers use every day. When this team...  ...controls instead of gaps.As a Senior DevOps / SRE Engineer on this team, you'll own reliability and deployments across our AWS...  ...'ve written and shipped real software, not only infrastructure code.... 
    Senior
    Temporary work
    Work at office
    Immediate start
    Remote work
    Flexible hours

    Creditly Corp

    Philadelphia, PA
    4 days ago
  • $140k - $170k

    We are looking for a Senior Site Reliability Engineer to work as part of a lean, product‑focused engineering organization. This role is about building...  ...and work experience 8+ years of experience in DevOps, SRE, platform engineering, or similar roles supporting application... 
    Senior
    Full time
    Work experience placement
    Flexible hours

    SEI Investments Developments

    Chicago, IL
    4 days ago
  • $165k - $225k

     ...Sr. Site Reliability Engineer (SRE) Chicago, IL or Remote Moonlite delivers high-performance AI infrastructure for organizations running intensive...  ...including systems engineers, network engineers, and software developers. Preferred Qualifications... 
    Senior
    Remote work
    Flexible hours

    Moonlite AI

    Chicago, IL
    2 days ago
  • $300k

     ...experimentation, full-scale model training, or inference. As a Platform Engineer/Senior Site Reliability Engineer, you’ll own the reliability, performance, and...  .... Skills / Must Have: ~7+ years of experience in SRE, DevOps, or Infrastructure Engineering roles supporting... 
    Senior
    Permanent employment
    San Francisco, CA
    more than 2 months ago
  •  ...We are seeking an experienced Site Reliability Engineer (SRE) – Microsoft Hyper-V & Private Cloud to operate highly available private cloud and...  ...Storage Spaces Direct enables you to build highly available, software-defined storage by pooling local disks (SSDs, NVMe drives... 
    Senior
    Contract work
    Local area

    2T Consulting

    Jersey City, NJ
    a month ago
  •  ...Netherlands. On behalf of Feeld , GT is looking for a Senior Site Reliability Engineer (SRE) to join a fast-growing consumer mobile product in the...  ...The ideal profile is someone who started in backend/software engineering and has moved into SRE or reliability-focused... 
    Senior
    Remote job
    Full time

    GT

    Remote
    17 days ago
  •  ...Responsibilities: Provide technical leadership to a growing team focused on applying software engineering practices to operations at scale. Monitor and report on service level objectives for a given applications services. Work with business and product owners... 
    Senior
    Full time

    PointClickCare

    Remote
    21 days ago
  •  ...Role: Senior SRE / Cloud Engineer Backup & Cyber Recovery Location: Chicago, IL Hybrid (3 days/week) Must Locals Rate: $60/hr on C2C/1099 Max Exp : 10+ Key Requirements ~7+ years in Infrastructure, Backup Engineering, or SRE. ~ Strong... 
    Senior
    Local area
    3 days per week

    Nexwave Inc

    Chicago, IL
    3 days ago
  • $85 - $90 per hour

     ...Role:  Senior SRE Engineer  Location: Dallas / Fort Worth, Texas Rate: up to $85-$90 per hour...  ...Structure: 8 Month contract *** 4 days on-site *** -- We have a great new...  ...infrastructure. ~6+ years of overall software engineering experience in a development... 
    Senior
    Hourly pay
    Contract work
    Work experience placement

    CorGTA

    Dallas, TX
    more than 2 months ago
  •  ...provider of enterprise-scale context engines capable of analyzing trillions of real...  ...a highly skilled and motivated Site Reliability Engineer (SRE) to join our growing team. As an SRE...  ...infrastructure. You will bridge the gap between software development and operations, applying... 
    Full time

    Lovelace Ai

    Pittsburgh, PA
    1 day ago
  •  ...organization is expanding, and we are seeking a Senior Site Reliability Engineer to help drive a major architectural...  ...we build adheres to rigorous SRE principles. Mission & Impact...  ...on-prem operational environment as a software project by using Infrastructure as... 
    Senior
    Permanent employment
    Full time
    H1b
    Local area
    Remote work
    Shift work

    Jack Henry & Associates

    New York, NY
    1 day ago
  • We are looking for an experienced Site Reliability Engineer (SRE) to strengthen observability and operational resilience across a Microsoft Azure...  ...expand end-to-end visibility.• Work alongside DevOps and software engineering teams to strengthen platform stability, incident... 
    Long term contract

    Robert Half

    Maumee, OH
    4 days ago
  • $170k - $220k

     ...re looking for a hands-on, high-agency Site Reliability Engineer to help shape and scale the...  ...systems, and GitHub workflows. This is a software engineering role, deeply embedded in DevOps...  ...a Great Fit If YouHave 3-6+ years in SRE, DevOps, or infrastructure roles with... 
    Senior

    Supio

    Seattle, WA
    3 days ago
  •  ...TechMContact: Meghana GorusuCompany: SRI Tech SolutionsJob Title: Senior Site Reliability EngineerLocation: Plano , TX (remote)Years of Experience:...  ...are seeking a highly skilled Senior Site Reliability Engineer (SRE) to join our dynamic team. The ideal candidate will have... 
    Senior
    Remote work

    SRI Tech

    Plano, TX
    3 days ago
  • About the RoleWe’re looking for an experienced Site Reliability Engineer (SRE) to help us scale our platform with reliability, observability, and operational excellence at the core. You’ll partner with engineers and data scientists to build, automate, and maintain the... 
    Senior

    Alembic

    San Francisco, CA
    3 days ago
  • Inspire Brands is hiring two Senior Site Reliability Engineers to help build and scale reliable, resilient,...  ...digital platforms. These role blends software engineering, systems thinking, and operational...  ...experience applying and implementing SRE principles — not just supporting... 
    Senior
    Worldwide

    Inspire Brands

    Atlanta, GA
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior Software Engineer- Site Reliability Engineering (SRE). Be the first to apply!