SRE
Saxon Global Inc
Responsibilities:
*5-7 years of experience
• Gathers and analyzes metrics from monitoring platforms to assist in performance tuning
and fault tolerance.
• Participates in system design, platform management and capacity planning.
• Balances feature development speed and reliability with service-level objectives.
• Works closely with the incident response team and restoring service to normal operation.
• Understands debugging and applying troubleshooting skills.
• Investigates, blocks and rate-limits unwanted traffic.
• Utilizes monitoring systems and dashboards for proactive changes and alerting.
• Establishes continuous process improvement cycles where the process, performance,
and supporting technologies are reviewed and enhanced where applicable.
• Partners with development teams to improve services through testing and release
procedures.
KNOWLEDGE, SKILLS, ABILITIES
• Understanding of Kubernetes, containers, clusters and elastic scalability.
• Expertise in SRE principles.
• Mindset of continually finding ways to drive scalability, stability and performance.
• Cloud Services experience with Google Cloud Platform (GCP).
• Experience with API, service-based or microservice-based architecture.
• Proficiency in infrastructure, network, database, operating systems or security
troubleshooting and remediation.
• Architecture-level knowledge of Windows and Linux and Infrastructure systems.
• Experience with production deployment, monitoring and operational support for enterprise-class applications (Dynatrace a plus).
• Experience working with Continuous Integration/ Continuous Deployment tools.
• Experience in performance diagnostics, capacity planning, performance architecture
design, performance tuning and performance monitoring.
• Experience with Azure DevOps (ADO), Dynatrace, Prometheus, Terraform and Grafana
*5-7 years of experience
• Gathers and analyzes metrics from monitoring platforms to assist in performance tuning
and fault tolerance.
• Participates in system design, platform management and capacity planning.
• Balances feature development speed and reliability with service-level objectives.
• Works closely with the incident response team and restoring service to normal operation.
• Understands debugging and applying troubleshooting skills.
• Investigates, blocks and rate-limits unwanted traffic.
• Utilizes monitoring systems and dashboards for proactive changes and alerting.
• Establishes continuous process improvement cycles where the process, performance,
and supporting technologies are reviewed and enhanced where applicable.
• Partners with development teams to improve services through testing and release
procedures.
KNOWLEDGE, SKILLS, ABILITIES
• Understanding of Kubernetes, containers, clusters and elastic scalability.
• Expertise in SRE principles.
• Mindset of continually finding ways to drive scalability, stability and performance.
• Cloud Services experience with Google Cloud Platform (GCP).
• Experience with API, service-based or microservice-based architecture.
• Proficiency in infrastructure, network, database, operating systems or security
troubleshooting and remediation.
• Architecture-level knowledge of Windows and Linux and Infrastructure systems.
• Experience with production deployment, monitoring and operational support for enterprise-class applications (Dynatrace a plus).
• Experience working with Continuous Integration/ Continuous Deployment tools.
• Experience in performance diagnostics, capacity planning, performance architecture
design, performance tuning and performance monitoring.
• Experience with Azure DevOps (ADO), Dynatrace, Prometheus, Terraform and Grafana
Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the SRE in Vestavia Hills, AL vacancy
$55k - $152.38k
Position Overview At PNC, our people are our greatest differentiator and competitive advantage in the markets we serve. We are all united in delivering the best experience for our customers. We work together each day to foster an inclusive workplace culture where all of...SuggestedPermanent employmentFull timeTemporary workPart timeWork experience placementWork at officeShift workNight shift- ...Strong communication and collaboration skills across engineering, security, compliance, and business teams PREFERRED QUALIFICATIONS: CSSLP, GSSP, Security+, SSCP Experience with Kubernetes and container security Familiarity with SRE and resilience practices...SuggestedWork experience placementImmediate startFlexible hours
- ...JOB SUMMARY: This is a full-time position that spans two complementary technical areas within the Hypersonics department at Kratos SRE. Approximately half of the candidate's effort (~20 hours per week) will be devoted to supporting the design, development, and commissioning...SuggestedFull timeWork at office
- ...position serves as a technical team lead and developing subject matter expert within the Hypersonic Structures department at Kratos SRE, one of the nation's leading organizations in the evaluation of thermal protection systems and refractory materials for ballistic and...SuggestedFull timeWork at office
$140k - $150k
...Senior Azure Cloud Engineer / Azure SRE Become part of an experienced cloud engineering organization leading large-scale Azure modernization initiatives across the enterprise. This senior individual contributor position is centered on optimizing and expanding an established...SuggestedWork at officeLocal area- ...position serves as a technical team lead and developing subject matter expert within the Hypersonic Structures department at Kratos SRE, one of the nation's leading organizations in the evaluation of thermal protection systems and refractory materials for ballistic and...Temporary work
- ...configuration inconsistencies across environments- Provide technical leadership and mentoring to junior team members- Partner with engineering, SRE, and platform teams to improve system design, observability, and performance- Establish and promote standards, best practices, and...Full timeTemporary workPart timeWork experience placementWork at office
$122k - $150k
...contributor role, you'll leverage your expertise to strengthen resilient infrastructure, refine IaC and CI/CD artifacts, and elevate SRE practices—building on our solid foundation for even greater reliability. Reporting to the Senior Director of Cloud Infrastructure, you...Full timeTemporary work$42k - $172.25k
...and secure data stores and data movement patterns across MySQL and Oracle platforms, partnering closely with application engineering, SRE/operations, and security teams.You will be accountable for database architecture and implementation decisions that directly impact...Full timeTemporary workPart timeWork experience placementWork at office$100.1k - $204.49k
...efficient workload placement • Lead root cause analysis and prevention strategies for recurring incidents • Partner with Architecture, SRE, DevOps, and Application teams to influence design and scalability decisions • Build and lead a high-performing team, ensuring...Full timeTemporary workPart timeWork experience placementWork at office$75k - $150k
...Similar enterprise monitoring solutions Working knowledge of Python and/or Java. Experience supporting Site Reliability Engineering (SRE), Operations, or Application Support organizations. Familiarity with retail banking or financial services environments....Full timeTemporary workPart timeWork experience placementWork at office- ...reliability and resilience. This role focuses on building automation to reduce manual effort and prevent service-impacting incidents. The SRE combines software and systems engineering to build and support large-scale, distributed, fault-tolerant systems. This role ensures...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to SRE. Be the first to apply!

