SRE
Software Technology Inc
Job Title
Works with Software Engineering teams to drive continuous improvement of system reliability. Establishes a system of monitoring and alerting to measure reliability over time and identify customer impacting issues in a timely manner to help teams operate within their error budgets. Requires the ability to understand a complex architecture across a broad range of technologies. Performs maintenance and provides technical assistance and advice on existing software solutions.
Major Duties And Responsibilities
- Actively and consistently supports all efforts to simplify and enhance the customer experience
- Ability to understand programs according to functional and non-functional requirements
- Develops and maintains technical documentation
- Serves as an escalation point to resolve incidents and problems for production applications and web services supported by the team in accordance with identified Service Level Objectives
- Collaborates with internal customers, technical, and architecture teams to solve complex software problems
- Provides general system users and management with system analysis and feedback
- Influences system design by identifying and recommending design and requirements needs for software enhancements
- Mentors and coaches less experienced staff
- Participates in continuous performance improvement and root cause analysis sessions to discuss opportunities to improve processes, system reliability, and standards
- Analyzes and resolves computer related problems by coordinating with in-house personnel to diagnose and fix operational issues, as well as consulting, advising and training on specialized features and functions
- Follows established configuration/change control processes
- Helps to reduce toil
- Implements internal automation systems, including code development
Required Qualifications
Skills/Abilities and Knowledge
- Ability to read, write, speak, and understand English
- Ability to identify measures or indicators of system reliability and the actions needed to improve or correct reliability, relative to the goals of the system
- Ability to deal with ambiguity, uncertainty, and incomplete information when evaluating alternatives and making recommendations
- Ability to work seamlessly within a team as well as manage individual tasks
- Creative and abstract thinking skills to troubleshoot and resolve complex problems
- Proven ability to work independently; designing, developing and deploying solutions, and to deliver projects on time with minimal direction
- Ability to listen and evaluate all opinions without bias, and contribute to a common culture of excellence
- Extensive technical knowledge of Information Technology field and computer systems
- Strong communication skills (written, interpersonal, presentation), with the ability to easily and effectively interact and negotiate with business stakeholders
- Strong ability to pick up complex concepts and processes quickly
- Proven leadership abilities including ability to share knowledge, resolve conflict and create consensus
- Demonstrated ability to take the lead on the most complex projects
- Servant-leader with ability to establish a blameless culture focused on continuously improving
- 1-2 years of experience with RESTful services
- Working knowledge of SQL and NoSQL databases
- 1 year working in Splunk to write Reports, Alerts, and basic queries
Education: BA/BS in Information Technology, Computer Science, related field or equivalent work experience
Related Work Experience: 3-5 years telecommunication Experience, 3-5 years IT Experience, 3-5 years of experience with one of the following: Software Development methodologies, Architecture, Networking, or Systems Administration
Required Skills: Splunk and SQL with an analytical mindset is the most important thing we're looking for in candidates.
$119k - $170k
About ZscalerZscaler accelerates digital transformation to ensure our customers can be more agile, efficient, resilient, and secure. As an AI-forward enterprise, we are constantly pushing the envelope, leveraging the world’s largest security data lake to power our cloud...SuggestedFull timeWork at officeLocal areaRemote work3 days per week- Raft is seeking a Software Reliability Engineer to analyze test results, coordinate bug fixes with development teams, and develop automated performance tests. You will design test environments, review procedures, and contribute to performance plans across Web services. ...Suggested
- ...understand the friction, and remove it. What you will be doing at UiPath As a Senior Software Engineer, you will Design, engineer, and build SRE platform systems and capabilities with cutting-edge AI, treating them as products that other engineering teams depend on in their...SuggestedWork at officeImmediate startRemote work
- PNC is seeking a Site Reliability Engineer to join the Information Technology Group, based at one of its IT hubs including Denver, CO. The role focuses on stabilizing complex production environments, building robust monitoring, and driving incident resolution with a strong...Suggested
- UiPath is seeking a Senior Software Engineer for the Site Reliability team. You will design, build, and ship SRE platform systems and capabilities that other engineering teams depend on, with a focus on reliability, scalability, and AI-powered improvements. You will own...Suggested
- ...Platform Engineer / SRE (Site Reliability Engineer) Cloud & AI Automation Location: Colorado (Mon to Thu wfo) Role Overview: We are seeking a highly skilled and proactive **Senior Platform Engineer / SRE** to lead the design, automation, and operational...Flexible hours
- ...creative development firm that addresses clients most pressing needs and challenges. We currently looking for Platform Engineer / SRE (Site Reliability Engineer) Cloud & AI for a client based out in Denver Colorado . Please see the job description below for...
$100k - $130k
...delivery.• Lead cloud modernization and platform transformation initiatives for enterprise clients.• Mentor engineering teams on DevOps, SRE, cloud-native, and platform engineering best practices.• Implement Git branching and release management processes that support...Temporary workLocal area- ...BringRequired:5+ years of experience in consulting, solutions architecture, platform engineering, DevOps, Site Reliability Engineering (SRE), or related customer-facing technical roles.Proven experience advising customers on DevSecOps, software delivery modernization,...Local areaImmediate startRemote work
- ...testable and performant with limited oversight and guidance, following best practices and approved code patterns. • Working with the SRE teams establishes a system of monitoring and alerting to measure reliability over time and identify customer-impacting issues in a...
$165k - $216.56k
...concepts to both technical and non-technical audiences.PreferredExperience selling to Developer, DevOps, and Site Reliability Engineering (SRE) personas.Prior experience in a fast-paced, high-growth Observability or Application Performance Monitoring (APM) company.Every...Contract workRemote work$141k - $221k
...participate in penetration tests of our cloud servicesWe are looking for people who have:5+ years hands-on-keyboard in Cloud Security, SRE, DevOps, DevSecOps, or Infra Engineering.Strong working knowledge of Kubernetes and ecosystem tools such as helm, ArgoCD.Production...Contract workLocal areaImmediate startRemote workWorldwideHome officeShift work$98.5k - $206.8k
...technical discipline.10+ years of experience supporting software development, DevOps, platform engineering, site reliability engineering (SRE), or system administration activities.Extensive experience designing and maintaining CI/CD pipelines using Jenkins, GitLab CI/CD,...Contract workWork experience placementFlexible hours$90k - $105k
...the Role Acquire Learning is hiring its first dedicated Site Reliability Engineer. This is a mid‑level role with a clear path to Lead SRE at Acquire as the company grows. We are looking for someone with real‑world SRE, DevOps, infrastructure, and production‑engineering...Work at officeImmediate start3 days per week$192.4k - $275.8k
...of related experience, or Masters + 8 years of related experience, or PhD + 5 years of related experience.7+ years of experience in SRE, cloud operations, and systems engineering with Linux administration6+ years of hands-on experience across AWS, GCP, or Azure; 5+ years...Full timeTemporary workLocal areaFlexible hours- ...understand the friction, and remove it.What you will be doing at UiPathAs a Senior Software Engineer, you willDesign, engineer, and build SRE platform systems and capabilities with cutting-edge AI, treating them as products that other engineering teams depend on in their...Work at officeImmediate startRemote work
$130k - $155k
...Maintenance: Manage alert triage, prepare for maintenance windows, and conduct node delivery testing.Collaboration: Work closely with SRE, Networking, and Storage teams from initial triage to root cause analysis (RCA) delivery.Global Teamwork: Adhere to global team...Full timeTemporary work$124.36k - $146.3k
...level Reliability Engineer specializing in observability, this role partners closely with product owners, application engineering teams, SRE teams, and business stakeholders to translate customer journeys and business outcomes into measurable reliability objectives. The...Full timeWork experience placementLocal area3 days per week$190k - $225k
...Development Experience with agentic solutions including agentic platforms, orchestrating agents and sub-agents to run projects, AI in SRE/RCA/observability, AI in CI/CD, and other agentic solutions within the cloud native solutions space, etc), Internal Developer...Work experience placementLive inWork at officeLocal areaFlexible hours- ...lead the design and implementation of Infrastructure Automation, CI/CD for infrastructure operations, Site Reliability Engineering (SRE), DevSecOps, and Software Development Lifecycle (SDLC) processes. Provide technical leadership in Software-Defined Networking (SDN),...Full timeLocal area
$175k - $215k
...security requirements (FISMA, FedRAMP, NIST), implements zero‑trust principles, and maintains robust security posture. Partner with SRE managers to architect DevSecOps pipelines, CI/CD automation, infrastructure‑as‑code patterns, and observability solutions that enable...Work experience placementH1bWork at officeLocal area$110.11k - $204.49k
...observability initiatives that enhance the overall customer experience. As a Software Engineering Manager - Site Reliability Engineering (SRE), you will lead a team responsible for ensuring the reliability, scalability, and operational excellence of mission-critical...Permanent employmentFull timeTemporary workPart timeWork experience placementWork at officeAfternoon shift$150k - $180k
...in Computer Science or related technical field, or equivalent practical experience. 12+ years of experience in software engineering, SRE, DevOps, platform engineering, or related disciplines. 5+ years in technical leadership roles, including managing senior engineers...Temporary workWork at office$183.71k - $229.64k
...whatever hat the moment calls for: prototyping a new retrieval technique one week, hardening a backend service the next, then doing SRE or QA work when the team needs it. At this level, you’re expected to be a trusted expert beyond your own team — defining technical direction...Full timeWork at officeImmediate startRemote workShift work$190k - $220k
...vulnerability management, and access governance across cloud and datacenter environments. Cross-Functional Collaboration Collaborate with SRE to ensure infrastructure platforms support reliability goals and service level objectives. Work with Product Engineering leaders to...Contract workTemporary workWork at officeWork from homeFlexible hours$185k - $246k
...., PagerDuty, Opsgenie, ServiceNow, Jira Service Management, Rundeck)Experience using Datadog and/or other observability tools in an SRE or DevOps capacity, particularly owning on-call responsibilities or building internal automationExperience building custom applications...Work at office$185k - $246k
...working as a Solutions Architect or Sales Engineer supporting the AI spaceExperience using Datadog and/or other observability tools in an SRE or DevOps capacityExperience with LLM application evaluation and testing frameworksDatadog values people from all walks of life. We...Work at office$38 - $54 per hour
...ago Staff Security Operations Engineer, Incident Response Lead Software Engineering Specialist - Human Data Site Reliability Engineer (SRE, Remote US) Denver, CO $120,000.00-$160,000.00 2 months ago Staff Software Engineer (Online Storage) Denver, CO $198,000.00-$230,000....Weekly payContract workRemote workMonday to Friday$100.8k - $170k
...processes, and operational workflows. What You’ll Need 5+ years of experience as a Software Engineer or related technical role (e.g. SRE, DevOps Engineer, Cloud Engineer). Have played a major role in large software development projects in the cloud (preferably AWS)....Temporary workLocal area$127k - $249k
The TeamPlatform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions that support the broader engineering organization. Among these are our multi-cloud-provider Kubernetes infrastructure, networking...Work at officeLocal areaRemote workWorldwideFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to SRE. Be the first to apply!

