Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior Site Reliability Engineer

DataArt Enterprises Inc

Senior Site Reliability Engineer

Our client is developing a reliable, scalable, and user-friendly ticketing and streaming platform for high school sports. Their goal is to create a solution that allows parents, students, and fans to purchase tickets and stream live events effortlessly, ensuring accessibility from any device, anywhere, at any time.

Working Schedule: 8:00–17:00 Eastern Standard Time (EST). A 4-hour overlap with EST is required for effective collaboration.

The project focuses on building and maintaining a highly available, scalable cloud platform that supports business-critical services. Engineering teams continuously improve system reliability, performance, and operational excellence through automation, observability, and modern software delivery practices.

As a Senior Site Reliability Engineer, you will work at the intersection of software engineering and operations, partnering with application, DevOps, and QA teams. You'll enhance observability, automate operational processes, improve CI/CD workflows, and drive reliability initiatives that enable teams to deliver resilient, high-performing software at scale.

Responsibilities:

  • Enhance platform observability by designing and maintaining metrics, alerts, dashboards, and monitoring capabilities that improve system visibility and reduce incident resolution time.
  • Build and maintain automation, operational tooling, and monitoring solutions that increase service reliability and uptime.
  • Work closely with software development and QA teams to embed reliability best practices into software delivery, release processes, and testing strategies.
  • Promote operational excellence by driving preventive measures, facilitating blameless post-incident reviews, and supporting capacity and scalability planning.
  • Take part in an on-call rotation, ensuring timely investigation and resolution of production incidents affecting critical services.

Requirements:

  • Strong hands-on experience with Python, particularly for scripting, automation, and operational tooling.
  • Proficiency in at least one of the following programming languages: Java, C++, or Go.
  • Solid knowledge of Linux environments, cloud platforms (AWS, GCP, or Azure), and containerized infrastructure using technologies such as Docker, Kubernetes, and Terraform.
  • Experience designing and maintaining CI/CD pipelines, working with version control systems, and implementing automated testing practices.
  • Practical experience with observability and monitoring platforms (such as Prometheus, Grafana, ELK, Datadog, or similar), including troubleshooting through log and metric analysis.
  • Experience identifying and documenting Critical User Journeys and translating them into measurable SLA/SLO objectives that support automation and operational excellence.
  • Strong collaboration and communication skills, with the ability to work effectively across multidisciplinary engineering teams, especially during critical production events.
  • A reliability-first mindset with the belief that system stability is a shared responsibility across engineering teams.
  • Familiarity with AI-assisted engineering tools (such as Claude and Codex) and their use within modern software development workflows.

Nice to have:

  • Experience developing or maintaining end-to-end and integration tests for distributed or microservices-based systems.
  • Knowledge of performance optimization, capacity management, or chaos engineering practices.
  • Experience contributing to internal developer platforms, automation tools, or reliability engineering initiatives.
  • Understanding of production security, compliance requirements, or change management processes.
  • Relevant industry certifications.

What We Offer:

  • Vacation days: Up to 26 business days per year.
  • 10 illness/special days off per year (fully paid, no medical papers needed) for all contract types.
  • Health and life insurance (Luxmed).
  • MyBenefit platform with Multisport option.
  • Internal psychological support service.
  • English language classes from the first working day.
  • Access to external learning platforms: O'Reilly, LinkedIn Learning, Udemy, and a wide catalog of diverse internal training.
  • Flexible workplace: work from the office, from home, or choose a hybrid option.
  • Tech Skills Mentoring Program.
  • Opportunities to develop as a public speaker, mentor, or technical interviewer.
  • Fully paid idle (bench) when not involved in a project.
  • Certification reimbursement (AWS, GCP, Microsoft, etc.)
Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Senior Site Reliability Engineer in United States vacancy
  •  ...Evaluate applications, platforms, and vendors to assess resiliency, reliability, and operational risk.Design and implement processes that...  ...and reliability tooling.Actively participate in reliability engineering and resilience communities of practice, contributing to... 
    Senior
    Full time

    Vanguard

    Wayne, PA
    3 days ago
  • $170k - $220k

    Who We're Looking ForWe’re looking for a hands-on, high-agency Site Reliability Engineer to help shape and scale the reliability layer of our stack. You'll own the release pipeline end-to-end — managing daily releases, weekly deploys, and hotfixes — while also automating... 
    Senior

    Supio

    Seattle, WA
    9 hours ago
  • About the RoleWe’re looking for an experienced Site Reliability Engineer (SRE) to help us scale our platform with reliability, observability, and operational excellence at the core. You’ll partner with engineers and data scientists to build, automate, and maintain the infrastructure... 
    Senior

    Alembic

    San Francisco, CA
    9 hours ago
  • $65 - $75 per hour

    DescriptionKforce has a client seeking a remote Senior Site Reliability Engineer to be a l be a leading member of the team working with a diverse range of technologies. You will enjoy working in a friendly environment and benefit from our investment in staff. The role... 
    Senior
    Remote work

    KForce

    Boca Raton, FL
    4 days ago
  • IXL Learning, developer of personalized learning products used by millions of people globally, is seeking a Senior Site Reliability Engineer to join our team, and help maintain the reliability and optimal performance of our products. We are seeking engineers with a passion... 
    Senior
    Work at office
    Immediate start

    IXL Learning

    Raleigh, NC
    2 days ago
  •  ...TechMContact: Meghana GorusuCompany: SRI Tech SolutionsJob Title: Senior Site Reliability EngineerLocation: Plano , TX (remote)Years of Experience: 8...  ...are seeking a highly skilled Senior Site Reliability Engineer (SRE) to join our dynamic team. The ideal candidate will... 
    Senior
    Remote work

    SRI Tech

    Plano, TX
    9 hours ago
  •  ...work from home day is currently Tuesday.Engineering at Lambda is responsible for building and...  ...and networking teams to improve service reliability and deployment workflowsDeploy and...  ...rotationYouHave 5+ years of experience in Site Reliability Engineering, Production Engineering... 
    Senior
    Work at office
    Local area
    Work from home
    Flexible hours

    Lambda Labs

    San Jose, CA
    2 days ago
  • Senior Site Reliability Engineer ILocationSan Jose, Costa Rica - RemoteSummary of roleOwn availability, the most important product feature, by continually striving for sustained operational excellence of Sumo’s planet-scale observability and security products. Work with... 
    Senior
    Flexible hours

    Sumo Logic

    San Jose, CA
    3 days ago
  • $168k - $270.25k

    NVIDIA is looking for a Senior Site Reliability Engineer (SRE) to join its GeForce Now (GFN) team. SRE at NVIDIA ensures that our internal and external-facing GPU cloud gaming services have reliability and uptime as promised to the users and at the same time enables developers... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  • $152.6k - $191.5k

     ...responsible for partnering with leaders across engineering and technology to define objective reliability goals for services. Key responsibilities include...  ...and continuous improvement.Position Summary:The Senior GCP Site Reliability Engineer acts as an advanced senior... 
    Senior
    Full time
    Work at office
    Day shift

    Bank of America

    Plano, TX
    3 days ago
  • $210k - $230k

    GovCIO is currently hiring for a Senior Site Reliability Engineer (SRE) to design, implement, and maintain highly available, scalable, and resilient infrastructure systems. The ideal candidate will bridge the gap between development and operations, focusing on automation... 
    Senior
    Currently hiring
    Remote work

    Govcio

    Arlington, VA
    1 day ago
  • $104.9k - $174.7k

     ...Data Management. You can learn more about LexisNexis Risk at the link below, About the Role:We are hiring a hands-on Senior Site Reliability Engineer (SRE) to actively build, operate, and improve the reliability of our production systems. This is not a purely advisory... 
    Senior
    Full time
    Work at office
    Local area
    Remote work
    Work from home

    RELX Group

    Alpharetta, GA
    4 days ago
  • $152.5k - $205k

     ...flexible work environment where new ideas are encouraged and everyone is a stakeholder.What you’ll be responsible for:As a Senior Site Reliability Engineer on Circle’s platform team, you’ll design, build, and operate the secure, scalable platform infrastructure behind... 
    Senior
    Flexible hours

    Circle

    San Francisco, CA
    4 days ago
  • $168k - $270.25k

     ...phenomenal people like you to help us accelerate the next wave of artificial intelligence.Join our team at NVIDIA as a Senior Site reliability engineer focused on HPC storage and play a crucial role in designing, implementing, and optimizing on-prem High-Performance... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • Job Description:About the Role: We are looking for a Senior SRE to join our Platform Engineering team as the operations owner of our observability platforms. You’ll be responsible for the reliability, scalability, and continued evolution of the tools that give our engineering... 
    Senior
    Full time

    Dimensional Fund Advisors

    Austin, TX
    9 hours ago
  •  ...professionalism. We are seeking an experienced AWS solution design engineer/architect to join our infrastructure cloud team. The...  ...product features efficiently and confidently them into production.As Senior SRE, you will be responsible for providing leadership, design and... 
    Senior

    Black Knight Financial Services

    Jacksonville, FL
    4 days ago
  • $15k

     ...benefits packages, technology talks by our experts, a beautiful modern office, daily catered lunches, and more.As a Senior Cluster Site Reliability Engineer (SRE), you will help scale our research compute cluster to meet our growing needs, and you will leverage... 
    Senior
    Work at office
    Local area
    Remote work

    The Voleon Group

    Berkeley, CA
    3 days ago
  •  ...Lambda’s designated work from home day is currently Tuesday.Engineering at Lambda is responsible for building and scaling our cloud offering...  ...and SLIs for Kubernetes services, workloads, and platform reliability.You6+ years of experience in a SRE, operations engineer, or... 
    Senior
    Work at office
    Local area
    Work from home
    Flexible hours

    Lambda Labs

    San Jose, CA
    3 days ago
  • LeanData helps the world’s fastest-growing companies automate, simplify, and accelerate revenue.We are looking for a Senior Site Reliability Engineer to lead the strategic evolution of our cloud infrastructure. Reporting directly to the SVP of Engineering, this role is... 
    Senior
    Full time
    Work at office
    2 days per week

    LeanData

    Santa Clara, CA
    9 hours ago
  •  ...candidates that are particularly strong in a few areas, and have some interest and capabilities in others.About the Role:As a Site Reliability Engineer, you’ll join the global Platform SRE team responsible for building, operating, and scaling Kong’s multi-region SaaS... 
    Senior
    Temporary work

    Kong

    Washington DC
    4 days ago
  • $119.8k - $234.7k

     ...yearEmployment type: Full-TimeWork site: 3 days / week in-officeRole type: Individual...  ...EngineeringDiscipline: Site Reliability EngineeringCompany: MicrosoftOverviewMicrosoft...  ...’s most demanding workloads. As a Senior Site Reliability Engineer, you will lead reliability... 
    Senior
    Ongoing contract
    Local area
    3 days per week

    Microsoft

    Redmond, WA
    1 day ago
  • $80k - $140k

    Job DescriptionRBC Wealth Management Technology is seeking a Senior Site Reliability Engineer to join its Wealth Management SRE Team. This team is responsible for ensuring the performance, availability, resilience, and operational excellence of critical applications and... 
    Senior
    Full time
    Flexible hours
    Shift work

    Royal Bank of Canada

    Minneapolis, MN
    1 day ago
  • Job Description:Note: Fidelity will not provide immigration sponsorship for this positionThe RoleOur Site Reliability Engineering group within Enterprise Infrastructure combines Operations Excellence with the Development Experience to deliver services at high scale, high... 
    Senior
    Full time

    Fidelity Investments

    Durham, NC
    4 days ago
  • $104.9k - $174.7k

    Are you passionate about improving reliability, scalability, and resilience in complex database...  ....Own prioritization of reliability engineering tasks within team backlogs.Lead incident...  ...a Service (IaaS).Background in DevOps, site reliability engineering practices, or related... 
    Senior
    Full time
    Local area

    LexisNexis Risk Solutions Group

    Texas
    9 hours ago
  •  ...and foster a dynamic work environment where new ideas thrive. Are you ready to join our team and make an impact?As a Senior Site Reliability Engineer at TeamViewer, you’ll be a key player in ensuring the reliability, scalability, and performance of our Azure-based SaaS... 
    Senior
    Temporary work
    Casual work
    Worldwide

    TeamViewer

    Austin, TX
    9 hours ago
  •  ...the team takes that seriously!The RoleThe Senior SRE at 2K is a hands-on technical leader...  ...regions while partnering with network engineers, systems architects, and game studio developers...  ...technical direction, influencing reliability from architecture review through production... 
    Senior

    2K Games

    Austin, TX
    2 days ago
  • $267k - $356k

     ...day is currently Tuesday.Lambda's Storage Engineering team is the backbone behind our world-...  ...workloads in the industry, which means reliability and performance aren't just goals—they're...  ...defined storage across new and existing sites using tools such as Ansible, Jenkins etc... 
    Senior
    Work experience placement
    Work at office
    Local area
    Work from home
    Flexible hours

    Lambda Labs

    San Jose, CA
    2 days ago
  • $101k - $161k

     ...several prestigious awards, such as Best Engineering Team, Best Company for Diversity,...  ...DescriptionWho You'll Work WithWe’re looking for Site Reliability Engineers to join our growing Arista’s...  ...: EngineeringExperience level: Mid-Senior LevelIndustry: Computer Networking
    Senior

    Arista Networks

    Santa Clara, CA
    3 days ago
  •  ...Georgia, and serves customers in more than 35 countries worldwide.Position OverviewWe are seeking a highly experienced Senior Site Reliability Engineer (Unified Observability) to lead the design, implementation, and operational maturity of the F1 Next Generation... 
    Senior
    Full time
    Worldwide
    Flexible hours

    NCR

    Atlanta, GA
    4 days ago
  • $98k - $176k

     ...joy of everyday life. We bring that vision to life through our values and culture. Learn more about Target here. As a Senior Site Reliability Engineer within Digital Enablement, you specialize in building and supporting the platforms and tools that enable teams to deliver... 
    Senior
    Full time
    Temporary work
    Work experience placement
    Flexible hours

    Target

    Brooklyn Park, MN
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!