Site Reliability Engineer
$145.02k - $180kArctic Wolf Incident Response
Senior Reliability DeveloperAt Arctic Wolf, you won't just watch the cybersecurity industry evolve – you'll help lead the change. Our global Pack is made up of people who thrive on solving hard problems, moving fast, and building technology that protects organizations around the world. We're proud to be recognized by Forbes, CNBC, Fortune, CRN, Gartner Peer Insights and IDC MarketScape – but what matters most is the work behind it: delivering real outcomes for customers through award winning innovation like our Aurora Platform.If you're looking for meaningful work, smart teammates and the chance to make a real impact in a high-growth company that's redefining security operations, Arctic Wolf is the right place for you!Our mission is simple: End Cyber Risk. We're looking for a Senior Reliability Developer to be part of making this happen.About the RoleResponsible for changes and updates to public/private cloud infrastructure. Owns all monitors that support AWN business services and our customers. Implements effective automation using programming languages such as Bash, Python, and JavaScript. Ensure our full infrastructure stack is resilient with day-to-day care and feeding. Write and manage Terraform modules and configuration trees for deployments into AWS and OpenStack. Manage automated patching solutions for instances deployed across multiple regions to ensure compliance with strict security requirements. Monitor internal and external TLS certificates for expiration, renewals, and re-deployment into their respective environments. Create and improve service monitoring solutions using Prometheus, Grafana, Zabbix, and various Prometheus exporters (e.g., postgres, node-exporter). Migrate monitoring for legacy production services to a new monitoring system and recreated existing monitoring within Prometheus + Grafana. Collaborate frequently with development teams to implement new features, fixes, or troubleshoot services in lab/production environments. Troubleshoot problems with microservices running within containers, administer containers in Kubernetes clusters, manage cloud components deployed in multiple AWS regions, and maintain CI/CD pipelines. Manage the Cylance AWS Amazon cloud platform through regular administration of Kubernetes cluster upgrades, cluster module upgrades, logging configurations, and resizing of EBS volumes and instances. Troubleshoot issues between interconnected services by working with various teams and resolving issues related to API calls and service unavailability. Create, improve, and maintain detailed documentation for MOPs, service architecture, runbooks, monitoring configurations, and other general resources. Plan and execute service decommissioning in both lab and production environments upon customer or service owner request. Work on an on-call rotation to support business-critical services outside of business hours, respond to escalations, and conduct root cause analysis during incident reviews. Telecommuting permissible from any location within US.Requirements: Bachelor's degree or foreign degree equivalent in Computer Information Systems, or related field and five (5) years of progressive, postbaccalaureate experience in a Technology related role or job offered or related role.Experience and/or education must include:Utilizing Infrastructure-as-Code (IaC) tools including Terraform and configuration management tools including Puppet and Chef to design, implement, and manage multi-region cloud infrastructure deployments across AWS and OpenStack environments.Utilizing containerization and orchestration technologies including Docker, Kubernetes, ECS, and EKS to deploy, manage, and troubleshoot microservices architectures in production environments.Utilizing Python, Bash, JavaScript, Groovy, and PowerShell scripting languages for automation of infrastructure provisioning, certificate lifecycle management, and automated patching solutions.Utilizing monitoring and observability tools including Prometheus, Grafana, Zabbix, AlertManager, and PagerDuty to create dashboards, configure alerts, and implement service health monitoring solutions.Utilizing AWS cloud services including EC2, ECS, EKS, ELB, S3, RDS, IAM, Lambda, CloudFormation, VPC, Route53, and CloudWatch to architect, deploy, and manage highly available cloud-based systems.Utilizing CI/CD pipeline tools including Jenkins, GitLab, GitHub, Bitbucket, and Gaia to automate testing, deployment processes, and maintain continuous integration workflows.Managing Kubernetes cluster operations including cluster upgrades, module upgrades, logging configurations, node scaling, EBS volume management, and pod orchestration across multiple AWS regions.Utilizing certificate management and security tools including Keeper for secret management, TLS certificate monitoring, renewal coordination, and automated deployment across production and development environments.Troubleshooting distributed systems and microservices communication issues including API failures, service mesh architectures, container networking, inter-service dependencies, and deployment failures.Creating and maintaining technical documentation using Atlassian tools including Jira, Confluence, Backstage, and LucidChart for MOPs (Method of Procedures), runbooks, service architecture diagrams, and monitoring configurations.All wolves receive compelling compensation and benefits packages, including:Equity for all employeesFlexible time off and paid volunteer daysRRSP and 401k matchTraining and career development programsComprehensive private benefits plan including medical, mental health, dental, disability, life and AD&D, and value-added servicesRobust Employee Assistance Program (EAP) with mental health servicesFertility support and paid parental leaveArctic Wolf is an Equal Opportunity Employer and considers applicants for employment without regard to race, color, religion, sex, orientation, national origin, age, disability, genetics, or any other basis forbidden under federal, provincial, or local law. Arctic Wolf is committed to fostering a welcoming, accessible, respectful, and inclusive environment ensuring equal access and participation for people with disabilities. As such, we strive to make our entire experience as accessible as possible and provide accommodations as required for candidates and employees with disabilities and/or other specific needs where possible. Please let us know if you require any accommodation by emailing View email address on click.appcast.io. View our Hiring Page to learn more about our application process.Security RequirementsConducts duties and responsibilities in accordance with AWN's Information Security policies, standards, processes, and controls to protect the confidentiality, integrity and availability of AWN business information (in accordance with our employee handbook and corporate policies).Background checks are required for this position.This position may require access to information protected under U.S. export control laws and regulations, including the Export Administration Regulations ("EAR"). Please note that, if applicable, an offer for employment will be conditioned on authorization to receive software or technology controlled under these U.S. export control laws and regulations.The base salary range for this job family is 145,018 to 180,000 USD annually. This range reflects the base pay the company reasonably expects to offer for this position, aligned to the broader job family base pay structure. Actual base pay may vary based on skills, experience, and location, including job family level. In addition to base pay, Arctic Wolf offers variable incentive compensation, new hire equity grants, and a comprehensive benefits package.
$91.7k - $163.7k
...modernization program aimed at updating and enhancing enterprise technology systems in accordance with modern design standards.As a Site Reliability Engineer (SRE), you will play a key role in ensuring the reliability, scalability, and performance of our systems, applications,...SuggestedMinimum wageFull timeWork experience placementLocal areaRemote work$91.7k - $163.7k
...potential to change lives. Ready to build the next breakthrough? Join us to start Caring. Connecting. Growing together. The Site Reliability Engineer will architect, develop, and maintain Optum Serve's cloud environment in both the commercial and government cloud. The...SuggestedMinimum wageFull timeWork experience placementWork at officeLocal areaRemote work- ...lives. Ready to build the next breakthrough? Join us to start Caring. Connecting. Growing together.We are seeking a Principal Site Reliability Engineer (SRE) to define and scale reliability practices across large-scale cloud platforms.This is a senior individual contributor...SuggestedMinimum wageFull timeWork experience placementWork at officeLocal areaRemote work
$91.7k - $163.7k
...serve as you help us advance health equity on a global scale. Join us to start Caring. Connecting. Growing together. As a Senior I O Engineer specializing in CICS System Programming, you will play a critical role in maintaining and optimizing the CICS environment on IBM z...SuggestedMinimum wageFull timeWork experience placementLocal areaRemote work- ...Arctic Wolf Networks, Inc. is seeking a Senior Reliability Developer to own cloud infrastructure automation and monitoring across AWS and OpenStack. You will manage Terraform modules, patching, TLS certificates, and multi-region deployments while ensuring security compliance...Suggested
$91.7k - $163.7k
...system stability; experienced in on-call support for critical incidentsLeverage enterprise-approved AI tools to streamline daily engineering workflows, automate routine operational tasks, and drive continuous process improvementEvaluate emerging technical trends and AI-...Minimum wageFull timeWork experience placementWork at officeLocal areaRemote work$112.7k - $193.2k
...the next breakthrough? Join us to start Caring. Connecting. Growing together.As an Experienced z System ISV Infrastructure Software Engineer within Optum Technology's Z Systems Application Hosting team, you will play a critical role in supporting and optimizing our high-...Minimum wageFull timeWork experience placementWork at officeLocal areaRemote workWeekend work$72.8k - $130k
...as you help us advance health optimization on a global scale.Join us to start Caring. Connecting. Growing together. As a Software Engineer within our OI Provider DevOps Enablement - Revenue Cycle Automation team, you will play a crucial role in modernizing and simplifying...Minimum wageFull timeTemporary workWork experience placementLocal areaRemote work$91.7k - $163.7k
...centric UI experiences using React, HTML5, CSS3, and modern UI frameworks Collaborate with Product Managers, UX Designers, and backend engineering teams to translate business requirements into efficient, scalable, maintainable UI solutionsIdentify and resolve front-end...Minimum wageFull timeWork experience placementWork at officeLocal areaRemote work$91.7k - $163.7k
...Web SSO etc. It is a highly available, reliable and scalable service hosted in public cloud... ...supportCollaborate with solution engineering, development teams, partners, and vendors... ...services (Azure AI, AWS SageMaker) Exposure to Site Reliability Engineering concepts and...Minimum wageFull timeWork experience placementWork at officeLocal areaRemote work$112.7k - $193.2k
...ensure timely resolution of incidents and service requests, lead root cause analysis and problem management, and partner closely with engineering, cybersecurity, infrastructure, SRE, product, and line-of-business operations teams to support high availability, resilience,...Minimum wageFull timeWork experience placementLocal areaRemote work$119.8k - $202.3k
C.H Robinson is seeking a Software Engineer III on our Global Forwarding engineering team where you will collaborate with product managers, fellow engineers, and business partners to design and deliver solutions that transform how we serve our global customers. You’ll...Hourly payFull timeWork experience placementRemote workWorldwide$91.7k - $163.7k
...revenue growth. Ready to help us deliver results that improve lives? Join us to start Caring. Connecting. Growing together. Software engineering is the application of engineering to the design, development, implementation, testing and maintenance of software in a systematic...Minimum wageFull timeWork experience placementWork at officeLocal areaRemote work$134.6k - $230.8k
...platforms and infrastructure to new levels of performance and reliability? Do you want to work alongside talented teammates who value collaboration... ...powers healthcare for millions of people. Every day, our engineers design, build, and operate some of the most advanced...Minimum wageFull timeWork experience placementWork at officeLocal areaRemote work$115k - $180k
ABOUT ADVANCED ENERGYAdvanced Energy (Nasdaq: AEIS) is a global leader in the design and manufacturing of highly engineered, precision power conversion, measurement and control solutions for mission-critical applications and processes. AE’s power solutions enable customer...Work at officeFlexible hours$112.7k - $193.2k
...Connecting. Growing together. We are seeking a Lead Software Engineer to drive development and delivery of the next generation of scalable... ..., including compatibility, scalability, security, reliability, data handling, observability, supportability, and long-term maintainabilityIndependently...Minimum wageFull timeWork experience placementLocal areaRemote work$200.4k - $343.5k
...lives. Ready to build the next breakthrough? Join us to start Caring. Connecting. Growing together.As the Vice President of Software Engineering for Utilization Management, you will lead the strategic technology vision, platform engineering execution, and AI workforce...Minimum wageFull timeWork experience placementWork at officeLocal areaRemote work$112.7k - $193.2k
...and serverless architecture3+ years of experience with GenAI services (Bedrock, Amazon Q, AI Agents)Driver's License and access to reliable transportationPreferred Qualifications: Proven excellent communication and stakeholder managementProven solid analytical and...Minimum wageFull timeWork experience placementWork at officeLocal areaRemote work$112.7k - $193.2k
Job ID: 96537911920Posted: 2026-08-01Location: Eden Prairie, Minnesota, United StatesSalary: $112,700 - $193,200/yearJob Area: TechnologyCompany: UnitedHealth GroupOptum is a global organization that delivers care, aided by technology to help millions of people live healthier...Minimum wageFull timeWork experience placementLocal areaRemote work$91.7k - $163.7k
...challenges. Position SummaryAs a Senior DevOps Engineer within Optum Technology, you will lead... ...of governance, quality, and operational reliability.Primary ResponsibilitiesLead the end-to-... ...in Release Engineering, DevOps, Site Reliability Engineering, or Software Engineering...Minimum wageFull timeWork experience placementLocal areaRemote work$134.6k - $230.8k
...change lives. Ready to build the next breakthrough? Join us to start Caring. Connecting. Growing together.As a Principal Software Engineer on the Identity and Profile team within Optum Technology, you will own the governance processes for enterprise-wide consumer identity...Minimum wageFull timeWork experience placementWork at officeLocal areaRemote work$112.7k - $193.2k
...an experienced and highly motivated Backend Tech Lead Software Engineer to join our team and help drive the design, development, and evolution... ..., and ensure the delivery of high-quality, scalable, and reliable software solutions. You will set engineering standards, mentor...Minimum wageFull timeWork experience placementLocal areaRemote work$159.75k - $225.5k
...who they uniquely are.We are seeking a Senior Principal Software Engineer to lead the design and development of next-generation Agentic AI... ...commitment to Diversity, Equity, and Inclusion on our Career Site.The compensation package for this role is based on multiple factors...Remote work$112.7k - $193.2k
Job ID: 97895518096Posted: 2026-08-01Location: Eden Prairie, Minnesota, United StatesSalary: $112,700 - $193,200/yearJob Area: TechnologyCompany: UnitedHealth GroupOptum Insight is improving the flow of health data and information to create a more connected system. We remove...Minimum wageFull timeWork experience placementWork at officeLocal areaRemote work$106k - $186k
...As an AI Engineer at 2nd Swing, you'll build production-grade AI systems that automate real work across the business, especially multi... ...Code, Cursor, Windsurf, etc.) to move faster with control and reliability. Come work with us, not for us! 2nd Swing is a one-of-a-...Full timeFlexible hours- Starkey Laboratories in Eden Prairie, MN is seeking a Failure Analysis & Reliability Engineer to investigate electrical and physical failures in microelectronic circuits, flex assemblies, and components. You will conduct reliability testing, perform root-cause analyses...Flexible hours
$84.8k - $191k
...structured problem-solving and root cause analysis to eliminate inefficiencies and improve operational processes.Ensure solution reliability through testing, monitoring, documentation, governance, and ongoing maintenance.Create reusable automation frameworks, standards,...Hourly payFull timeWorldwide$112.7k - $193.2k
...platforms and infrastructure to new levels of performance and reliability? Do you want to work alongside talented teammates who value collaboration... ...intent? UnitedHealth Group is seeking a Linux Platform Engineer who is passionate about Linux, virtualization, automation, and...Minimum wageFull timeWork experience placementWork at officeLocal areaRemote work$72.8k - $130k
...Software Engineer - Ai & Security Technology Optum is a global organization that delivers care, aided by technology, to help millions of people live healthier lives. The work you do with our team will directly improve health outcomes by connecting people with the care...Minimum wageFull timeTemporary workWork experience placementLocal areaRemote work$85.2k - $127.6k
...expected.Role OverviewAs an Application Engineer, you play a critical role in supporting... ...tailor solutions accordinglyConduct on-site and remote training sessions for customers... ...the basis of individual skill, ability, reliability, productivity, and other factors important...Local areaRemote workMonday to Friday
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!
- on-site clinical research associate (traveling/remote) Eden Prairie, MN
- site safety Eden Prairie, MN
- junior website developer Eden Prairie, MN
- construction site safety Eden Prairie, MN
- IT site lead Eden Prairie, MN
- historic site Eden Prairie, MN
- site leader Eden Prairie, MN
- site reliability engineer
- site reliability engineering manager
- junior site reliability engineer


