Site Reliability Engineering (SRE) Architect
ETHEREUM TECHNOLOGIES LLC
Site Reliability Engineering (SRE) Architect
Location:
Atlanta, GA
Duration:
12Months+ Extension
Hourly Rate:
Depending on Experience (DOE)
Work Authorization:
As an SRE Architect, you will be a pivotal technical leader responsible for designing, building, and evolving the foundational systems and practices that ensure the reliability, scalability, performance, and efficiency of our critical services. Moving beyond day-to-day operations, you will focus on the strategic architectural direction of SRE function, defining standards, blueprints, and frameworks that enable development teams and fellow SRE operations team to build and operate highly resilient systems. Leverage deep expertise in software engineering, distributed systems, cloud infrastructure, and SRE principles to influence technology choices, establish best practices, and foster a proactive culture of reliability across the organization and much beyond observability pillar.
Key Responsibilities:
- Reliability Strategy & Design:
- Architect and design highly available, scalable, secure, and cost-effective infrastructure and application patterns on AWS
- Define and evangelize SRE best practices, standards, and blueprints for service design, deployment, monitoring, and operational readiness across the engineering organization
- Review current observability implementation to identify gaps and define steps to reach next level maturity of observability setup to provide deep insights into system health and behaviour
- With overall maturity lead the definition and implementation strategy for Service Level Indicators (SLIs), Service Level Objectives (SLOs), and Error Budgets for critical services
- Design solutions to systematically reduce operational toil through automation and improved system design
- Evaluate current SRE tools and automation frameworks (e.g., CI/CD pipelines, Infrastructure as Code modules, automated incident remediation, chaos engineering platforms) and suggest enhancement that will help overall enhancement of capability
- Evaluate, prototype, and recommend new technologies, tools, and methodologies to enhance system reliability, developer productivity, and operational efficiency
- Technical Leadership & Consultation:
- Act as a senior technical advisor and subject matter expert on reliability, scalability, and performance for development and platform teams
- Provide architectural guidance during the design phase of new services and features to ensure reliability principles are embedded early (shift-left)
- Mentor and coach other SREs and engineers, fostering technical excellence and adherence to SRE principles
- Lead architectural reviews and production readiness assessments for critical systems
- Resilience:
- Lead blameless postmortems for significant incidents, ensuring root causes are identified and systemic architectural improvements are prioritized and implemented
- Architect and advocate for resilience patterns (e.g., circuit breaking, rate limiting, graceful degradation, chaos engineering) within applications and infrastructure
Required Qualifications:
- Proven experience in an architectural role, designing solutions for reliability, scalability, and performance
- Deep understanding and practical application of SRE principles (SLIs/SLOs, error budgets, toil reduction, automation, incident management, postmortems)
- Expertise in cloud computing platforms (e.g., AWS) including infrastructure, networking, and security services
- Strong experience with containerization and orchestration technologies (Kubernetes, Docker, serverless computing)
- Solid experience designing and implementing observability solutions (e.g., Dynatrace, Prometheus, Grafana, ELK/EFK Stack, Jaeger, OpenTelemetry)
- Strong programming/scripting skills (e.g., Python, Go, Bash) for automation and tool development
- Excellent analytical, problem-solving, and strategic thinking skills.
- Strong communication, collaboration, and leadership skills with the ability to influence technical direction across teams
Preferred Qualifications:
- Experience designing and implementing chaos engineering practices and platforms
ETHEREUM TECHNOLOGIES LLC is an equal opportunity employer inclusive of female, minority, disability and veterans, (M/F/D/V). Hiring, promotion, transfer, compensation, benefits, discipline, termination and all other employment decisions are made without regard to race, color, religion, sex, sexual orientation, gender identity, age, disability, national origin, citizenship/immigration status, veteran status or any other protected status. ETHEREUM TECHNOLOGIES LLC will not make any posting or employment decision that does not comply with applicable laws relating to labor and employment, equal opportunity, employment eligibility requirements or related matters. Nor will ETHEREUM TECHNOLOGIES LLC require in a posting or otherwise U.S. citizenship or lawful permanent residency in the U.S. as a condition of employment except as necessary to comply with law, regulation, executive order, or federal, state, or local government contract
#J-18808-Ljbffr- ...We are seeking an experienced Site Reliability Engineer (SRE) to support and maintain production systems hosted on AWS. The role focuses on production support, incident management, monitoring, observability, troubleshooting, and improving system reliability and availability...SuggestedContract work
$120k - $175k
...of sports fandom. Ready to reimagine the DFS industry together? We are seeking a highly skilled and experienced Senior Site Reliability Engineer to join our team. We are passionate about delivering cutting-edge solutions and pushing the boundaries of what's possible...SuggestedFull timeRemote workWork visaFlexible hours- ...SRE Engineer Location: Atlanta, GA (Hybrid) Qualifications: Strong experience supporting production systems hosted on AWS, including EC2, VPC, ALB/NLB, RDS, Lambda, and EKS. Hands-on experience with incident management and 24/7 production support models...Suggested
- ...us as their trusted partner. If you're ready to make an impact, you're in the right place. Job Details: Job Title: SRE Engineer Work Location: tlanta, GA (Hybrid) Duration: 12 months Contract Job Description: - SRE Engineer...SuggestedContract work
- ...development, cloud infrastructure, DevOps, SRE, and platform engineering. You will test AI-generated commands,... ...workflows for accuracy and reliability. Work with AWS, Azure, GCP, Kubernetes... ...DevOps Cloud Infrastructure Site Reliability Engineering (SRE) Platform...SuggestedRemote jobFor contractors
- As the Senior Site Reliability Engineer, you will serve as a trusted technical resource responsible for deploying, validating, and operationalizing... ..., Platform Engineering, Site Reliability Engineering (SRE), Systems Administration, or related technical roles.Experience...Work at officeImmediate startWorldwideShift work
- ...Lead Engineer, Site Reliability Engineering Team As a lead engineer with Retail, Site Reliability Engineering team, you will be at the forefront... ...for all infrastructure and services within the scope of SRE. Preserve operational visibility and response capabilities...
- ...the following job description: The Site Reliability Engineer role focuses on enhancing the reliability... ...observability practices, mentoring SRE team members, and contributing to enterprise... ...Engineering & Automation Architect and deliver automation solutions that...Permanent employmentFull timePart timeH1bWork at officeLocal areaImmediate startWork visaMonday to FridayShift workDay shift
- ...This position is 60 % SRE and 40% SDE. Also open for candidates... ...metrics so the support engineers can proactively and timely validate... ...with enterprise security architects to design and implement data... ...components. • 1+ Years in Site Reliability Engineering organization...Work experience placement
- ...I have an opportunity for "SRE " _ (Atlanta, GA - ONSITE)" and I am looking for a candidate who can join Immediately if you are... ...Configuration/Continuous Integration/Continuous Delivery/Release Engineering related tasks in JavaEE/C++ Environments. • Experience in...Immediate start
$60 - $68 per hour
...Site Reliability Engineer Immediate need for a talented Site Reliability Engineer. This is a 12+ months contract opportunity with long-term potential... ...the automation AWS Pipeline & Infrastructure, DevOps (SRE) activities, Monitoring & Alerting Our client is a...Contract workLocal areaImmediate start- ...Site Reliability Engineer We're looking for an experienced Site Reliability Engineer (SRE) to help us scale our platform with reliability, observability, and operational excellence at the core. You'll partner with engineers and data scientists to build, automate, and...
$168k - $200k
...Looking For We're looking for a Senior Site Reliability Engineer to join our Data & ML Platform team.... ...for someone who combines a strong SRE mindset with deep cloud... ...complex, hybrid cloud environment and can architect systems that balance velocity, safety,...Remote work- ...Senior Site Reliability Engineer Inspire Brands is hiring two Senior Site Reliability Engineers to help build and scale reliable, resilient, and... ...candidate has hands-on experience applying and implementing SRE principles — not just supporting production systems, but...Worldwide
- ...Technical Support Specialist In Site Reliability Engineering (Sre) Mandatory skills: Scripting and programming languages like Python, Java, Ruby. Cloud and infrastructure management – AWS, Google cloud and Azure is a plus- CI/CD Automation, Database Management. The...
- ...professionalism. We are seeking an experienced AWS solution design engineer/architect to join our infrastructure cloud team. The infrastructure... ...and confidently them into production. As Senior SRE, you will be responsible for providing leadership, design and...
$75.7k - $136.3k
...and building systems that scale? Join our highly skilled Site Reliability Engineering team! Our team designs, develops, and manages applications... ...that support Akamai Cloud's products and services. Our SRE teams solve reliability, security, and usability at scale for...Work experience placementWork at office- ...can create the conditions for educators to teach, students to thrive, and districts to shape the future of education. Site Reliability Engineer (SRE) Overview: We are looking for a Site Reliability Engineer (SRE) to join our Engineering team. This is a build-it-...Full timeLive inWork at office
- ...deliver faster, smarter, more reliable insights to insurance... ...place. The Role As a Site Reliability Engineer, you'll be responsible for... ...Looking For ~4+ years in SRE, DevOps, or cloud operations... ...certifications (Solutions Architect, DevOps Engineer) ~ Experience...Full timeFlexible hours
$100k - $120k
...OverviewThe Site Reliability Engineer is a key force behind improving Origami’s time to resolution and advancing overall site reliability and scalability... ...on our platform.Partners with the larger Cloud Operations, SRE, Engineering teams, and the business-at-large to advance our...Full timeTemporary workWork experience placementFlexible hours- ...At Intercontinental Exchange (NYSE:ICE), we engineer technology, exchanges and clearing houses that... ...people to join our team. We are seeking a Site Reliability Engineer to bring 3+ years of hands-on experience to our SRE team, operating with significant autonomy to...
- ...Join to apply for the Site Reliability Engineer role at Motion Recruitment Join to apply for... ...a point of contact for their complex SRE issues that will span across on-prem and... ...reports Work with enterprise security architects to design and implement data security...Contract workWorldwide
$130k - $145k
...Back Site Reliability Engineer Cloud/Infrastructure Atlanta , GA Sep 2, 2026 Site Reliability Engineer Atlanta, GA / Hybrid Blu Omega... ...experience in cloud infrastructure, systems engineering, DevOps, or SRE roles ~2+ years of hands‑on Microsoft Azure experience...Temporary work$130k - $150k
...to learn more. Base pay range $130,000.00/yr - $150,000.00/yr Overview: We are seeking a highly skilled Site Reliability Engineer (SRE) to join our team and help build and maintain scalable, reliable, and efficient systems. The ideal candidate will have...Full timeRemote work$169.3k - $304.7k
...Join our Network Infrastructure SRE team! Our team designs,... ...fast, efficient, scalable, and reliable routing software and... ...platform. As a Principal Site Reliability Engineer - Network, you will be responsible for: Architecting, building and supporting solutions...Work experience placementWork at office- ...an experienced Observability Architect to design, implement, and mature... ...visibility, performance, and reliability.Key... ...infrastructure, application, and SRE teams to ensure high availability... ...collaborating with DevOps, SRE, cloud engineering, and application teams in large...
$178.13k - $205.4k
...Bachelor's degree or foreign degree equivalent in Computer Engineering, Computer Science, Engineering, or related field plus five (5)... ...websites that are not Workday Careers. Please be aware of sites that may ask for you to input your data in connection with a job...Work at officeRemote workFlexible hours$81.1k - $187k
...Job Description We are looking for a Site Reliability Engineer 3 to support mission-critical cloud services and production operations. The... ...Ingestion and Management: -Takes proactive steps to design and architect infrastructure and/or service according to terms for...Temporary workImmediate startFlexible hoursShift work$123.4k - $222.53k
...! Ready grow your career as part of the Uncarrier journey at T-Mobile? Our team is searching for our next Sr. Site Reliability Engineer to strengthen the reliability and resilience of the systems powering T-Mobile's payment platforms, enabling faster, safer...Full timeTemporary workPart timeWork experience placementLocal areaFlexible hours$71.6k - $119.4k
...implement automation, troubleshoot issues, and work closely with senior engineers to learn and apply best practices. You'll gain exposure to a... .... Requirements: ~1–3 years of experience in DevOps, SRE, cloud engineering, or related IT roles (internships and...Temporary workInternshipLocal areaRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineering (SRE) Architect. Be the first to apply!
- site reliability engineer Atlanta, GA
- site reliability engineer remote Atlanta, GA
- site reliability engineer sre Atlanta, GA
- site safety Atlanta, GA
- website coordinator Atlanta, GA
- on-site clinical research associate (traveling/remote) Atlanta, GA
- site services specialist Atlanta, GA
- on site coordinator Atlanta, GA
- construction site safety Atlanta, GA
- junior website developer Atlanta, GA




