Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Site Reliability Engineer II

$95k - $171k

Akamai

Are you passionate about cutting-edge AI infrastructure?

Do you want to build your SRE career on one of the most exciting platforms in cloud computing?

Join the Akamai Inference Cloud Team

The Akamai Inference Cloud team is part of Akamai's Cloud Technology Group. We design, implement, deploy and operate AI platforms that enable customers to run inference models and developers to create AI applications.

Partner with the best

In this role, responsibilities will include automation, monitoring, incident response, and working collaboratively with skilled team members. Candidates should possess expertise in Linux systems, automation, and SRE practices. Daily activities involve coding, improving dashboards, enhancing alerts, and minimizing repetitive tasks. Opportunities exist to focus on GPU infrastructure, Kubernetes, and ensuring reliability for AI workloads within Akamai's serverless inference platform.

As an Site Reliability Engineer II, you will be responsible for:

  • Building and maintaining dashboards, alerts, and monitoring for inference workloads using Akamai's existing observability platform

  • Writing automation and tooling in Python or Go to reduce operational toil and improve system reliability

  • Building and improving runbooks for inference-specific operational procedures, integrating into Akamai's existing incident management processes

  • Contributing to SLO tracking and reporting, identifying trends and areas for improvement

  • Supporting CI/CD pipeline maintenance, deployment safety checks, and rollback procedures

  • Collaborating with product engineering teams to troubleshoot complex problems across the stack

  • Participating in on-call rotations, responding to production incidents, and conducting blameless post-mortems

Do what you love

To be successful in this role you will:

  • Have 2+ years of experience in Site Reliability Engineering and a Bachelor's Degree or its equivalent experience

  • Demonstrate coding ability in at least one programming language (Python or Go) with experience writing automation

  • Have experience with Linux systems administration and the ability to troubleshoot complex infrastructure issues

  • Show familiarity with Kubernetes and containerization concepts

  • Have experience with monitoring and observability tools such as Prometheus, Grafana, or similar

  • Have exposure to CI/CD pipelines and infrastructure-as-code tools (Terraform, SaltStack, or equivalent)

  • Show a willingness to learn and grow, with genuine curiosity about AI infrastructure and distributed systems

Work in a way that works for you

FlexBase, Akamai's Global Flexible Working Program, is based on the principles that are helping us create the best workplace in the world. When our colleagues said that flexible working was important to them, we listened. We also know flexible working is important to many of the incredible people considering joining Akamai. FlexBase, gives 95% of employees the choice to work from their home, their office, or both (in the country advertised). This permanent workplace flexibility program is consistent and fair globally, to help us find incredible talent, virtually anywhere. We are happy to discuss working options for this role and encourage you to speak with your recruiter in more detail when you apply.

Learn ( what makes Akamai a great place to work

Connect with us on social and see what life at Akamai is like!

We power and protect life online, by solving the toughest challenges, together.

At Akamai, we're curious, innovative, collaborative and tenacious. We celebrate diversity of thought and we hold an unwavering belief that we can make a meaningful difference. Our teams use their global perspectives to put customers at the forefront of everything they do, so if you are people-centric, you'll thrive here.

Working for you

At Akamai, we will provide you with opportunities to grow, flourish, and achieve great things. Our benefit options are designed to meet your individual needs for today and in the future. We provide benefits surrounding all aspects of your life:

  • Your health

  • Your finances

  • Your family

  • Your time at work

  • Your time pursuing other endeavors

Our benefit plan options are designed to meet your individual needs and budget, both today and in the future.

About us

Akamai powers and protects life online. Leading companies worldwide choose Akamai to build, deliver, and secure their digital experiences helping billions of people live, work, and play every day. With the world's most distributed compute platform from cloud to edge we make it easy for customers to develop and run applications, while we keep experiences closer to users and threats farther away.

Join us

Are you seeking an opportunity to make a real difference in a company with a global reach and exciting services and clients? Come join us and grow with a team of people who will energize and inspire you!

#LI-Remote

Compensation

Akamai is committed to fair and equitable compensation practices. For US based candidates only - the base salary for this position ranges from $95,000 - $171,000/year; a candidate's salary is determined by various factors including, but not limited to, relevant work experience, skills, certifications and location. Compensation for candidates outside the US will vary. The compensation package may also include incentive compensation opportunities in the form of annual bonus or incentives, equity awards and an Employee Stock Purchase Plan (ESPP). Akamai provides industry-leading benefits including healthcare, 401K savings plan, company holidays, vacation (in the form of PTO), sick time, family friendly benefits including parental leave and an employee assistance program including a focus on mental and financial wellness; Eligibility requirements apply.

Equal Employment Opportunity Rights

Akamai Technologies is an Affirmative Action, Equal Opportunity Employer that values the strength that diversity brings to the workplace. All qualified applicants will receive consideration for employment and will not be discriminated against on the basis of gender, gender identity, sexual orientation, race/ethnicity, protected veteran status, disability, or other protected group status.

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Site Reliability Engineer II in Washington DC vacancy
  • $103.5k - $150k

     ...Bring your whole self. The Role and Team The Site Reliability Engineering organization at Medallia brings together the infrastructure...  ...power a highly reliable global SaaS platform. As an SRE II, you will help operate and improve the reliability, scalability... 
    Suggested
    Temporary work
    Work experience placement
    Local area
    3 days per week

    Medallia

    McLean, VA
    1 day ago
  •  ...Site Reliability Engineer II Join the leader in providing smarter solutions for a safer world. The property technology space is growing rapidly, and Kastle Systems is leading the way. Kastle Systems is the leader in managed security, with a track record of introducing... 
    Suggested
    Remote work

    Kastle Systems

    Falls Church, VA
    8 hours ago
  •  ...Senior Software Engineer – Platform Systems II Confidential Client – Purple Squirrel Enterprises About the Opportunity Our client is...  ...mission-critical and regulated use cases where scalability, reliability, and real-time responsiveness are essential. Purple... 
    Suggested
    Full time
    Flexible hours

    Purple Squirrel Enterprises

    Washington DC
    20 hours ago
  • $165k - $230k

     ...with the ultimate goal of enabling human life on Mars. SR. SITE RELIABILITY ENGINEER (STARSHIELD) Starshield leverages SpaceX’s Starlink...  ...regulations, applicant must be a (i) U.S. citizen or national, (ii) U.S. lawful, permanent resident (aka green card holder), (iii... 
    Suggested
    Permanent employment
    Temporary work
    Immediate start
    Weekend work

    SpaceX

    Washington DC
    3 days ago
  •  ...The Software Engineer II/III develops, tests, and maintains software capabilities in a NAVSEA Program Office Support role. This role supports design implementation, defect resolution, and integration with broader system architectures. This position is contingent upon... 
    Suggested
    Full time
    Work at office

    Warrant Technologies

    Washington DC
    20 hours ago
  • $149.4k - $202k

     ...Senior Software Engineer- Site Reliability Engineering (SRE) DC, MD, VA, CA The Site Reliability Engineering discipline at Noctua Technology,...  ...Security+ certification or an equivalent DoD 8140/8570 IAT Level II baseline certification. Salary Range : $149,400 - $202... 
    Remote work

    Noctua Technology

    Washington DC
    4 days ago
  •  ...SSG)  is seeking a talented Software Developer in Test (SDET) - II Location: Remote Department: Veterans Affairs (VA)...  ...metrics and reporting Occasionally perform other IT systems engineering activities such as requirements, design, installation, operation... 
    Full time
    Work experience placement
    Remote work

    Ssg

    Upper Marlboro, MD
    20 hours ago
  •  ...Building Engineer II The Building Engineer II ensures sanitary, safe and comfortable facility for students, staff, and public; performs...  ...regulations. • Inspects school facilities to ensure that the site is suitable for safe operations and maintained in an attractive... 

    Alexandria City Public Schools

    Alexandria, VA
    1 day ago
  •  ...the US required; US Citizenship preferred*Remote Position TBD*Applicants residing in the DMV area preferred as the client is located in Silver Spring, MD*Three references required POSITION DESCRIPTION:The Systems Engineer II will provide advanced engineering and sys...... 
    Remote work

    Think Tank Inc

    Silver Spring, MD
    1 day ago
  • $200k - $235k

     ...System Engineer II BTS Software Solutions is seeking a System Engineer II with an active TS/SCI w/ POLY to join our team in Laurel...  ...quantitative analysis in non-functional system performance areas like Reliability, Maintainability, Vulnerability, Survivability, Producability... 
    Contract work
    Work experience placement
    Local area

    BTS Software Solutions

    Laurel, MD
    7 hours ago
  •  ...Job Title: Systems Engineer II Position Number: SYS E II 001 Location: Washington D.C. Worksite: Washington Naval Yard Status: Full-Time Contingent Upon Award Clearance: Secret Date Added: December 1, 2025 Job Summary: Applies systems... 
    Full time

    R3 Strategic Support Grp

    Washington DC
    1 day ago
  •  ...Statistical Programmer II/III (Permanent Role) United States Are you interested in working directly for a single sponsor while having the security and additional career opportunities that working for a global CRO can bring? Our team says it's the best of both worlds... 
    Permanent employment

    ClinChoice

    Washington DC
    1 day ago
  •  ...System High Corporation, located in Washington, D.C., is seeking a Program Security Representative II to provide advanced security support for Special Access Programs (SAP). The successful candidate must hold a current Top-Secret Clearance with SCI Eligibility and have... 

    System High Corp

    Washington DC
    4 days ago
  •  ...support Pension plan Paid maternity leave 401(k) Get notified when a new job is posted. Sign in to set job alerts for “Senior Site Reliability Engineer” roles. Bellevue, WA $204,000.00-$259,000.00 1 day ago Seattle, WA $115,000.00-$175,000.00 5 months ago Senior ServiceNow... 
    Contract work
    Remote work

    Signature IT World Inc

    Washington DC
    2 days ago
  • $175k - $250k

     ...Senior Cloud Infrastructure Engineer Location: San Francisco, CA. Remote unavailable. Modality: On‑Site only. Must live within commuting distance of San Francisco or...  ...while ensuring scalability, performance, and reliability across environments. What You’ll Do Design, build... 
    Full time
    Remote work
    Relocation
    Relocation package

    The Recruiting Guy

    Washington DC
    2 days ago
  • Computer Programmer II Location(s): Hyattsville, MD (on‑site) Position Details: Full‑time position No calls, nights, weekends, or holidays Full benefit...  ...developers and technical leads to ensure system reliability, data accuracy, and compliance with federal privacy and... 
    Full time
    Contract work
    Night shift
    Weekend work

    CICONIX

    Hyattsville, MD
    2 days ago
  • $131k - $227.13k

     ...Description: The 1LMX MES COE is seeking an engineer who will own infrastructure‑as‑code, cloud platform, and reliability for the Apriso environment on AWS. This role blends full‑stack development, DevOps, and Site Reliability Engineering (SRE) practices to deliver... 
    Full time
    Temporary work
    Work experience placement
    Work at office
    Remote work
    Relocation
    Flexible hours
    Shift work
    3 days per week

    Lockheed Martin Corporation

    Bethesda, MD
    3 days ago
  • $121.4k - $218.6k

     ...solve complex challenges? Do you have a passion for automation and building systems that scale? Join our highly skilled Site Reliability Engineering team! Our team designs, develops, and manages applications and infrastructure that support Akamai Cloud's products and... 
    Work experience placement
    Work at office

    Akamai

    Washington DC
    1 day ago
  • $153k - $185k

     ...Senior Site Reliability Engineer El Segundo, California, United States About Varda Low Earth orbit is open for business. Varda is accelerating the development of commercial space infrastructure, from in-orbit pharmaceutical processing to reliable and economical... 
    Permanent employment
    Full time
    Immediate start
    Relocation package
    Flexible hours
    Weekend work

    Varda Space Industries

    Washington DC
    4 days ago
  • $135k - $150k

    Senior Site Reliability Engineer Job number: 884 This is a remote position. Ad Hoc is a technology company that empowers organizations to deliver scalable, impactful digital services. Using modern, agile methods, our team creates products that meet people's... 
    Remote work
    Flexible hours

    Ad Hoc LLC

    Silver Spring, MD
    1 day ago
  • $112k - $179k

     ...system, network, software, and security solutions. About The Role Peraton is seeking a self-driven and resourceful Site Reliability Engineer to join our dynamic of Network and UC engineers in Washington, DC. This position combines software engineering and systems... 
    Contract work
    Worldwide
    Shift work

    Peraton

    Washington DC
    3 days ago
  •  ...Qualifications: ~10+ years of overall experience in IT including, with hands-on Development and Systems engineering background ~3-5 years of experience in a Site Reliability Engineering role ~ Experience with Enterprise Cloud transformation efforts ~ Experience with... 
    Temporary work
    Immediate start

    Samprasoft

    Washington DC
    3 days ago
  • $166k - $220k

     ...Senior Site Reliability Engineer Anduril Industries is a defense technology company with a mission to transform U.S. and allied military capabilities with advanced technology. By bringing the expertise, technology, and business model of the 21st century's most innovative... 
    Full time
    Work experience placement
    Immediate start

    anduril

    Washington DC
    1 day ago
  • $106.3k - $221.1k

     ...more. Join us to drive positive, lasting change that moves missions and the government forward! Job Description The Site Reliability Engineer will ensure the reliability, performance, and scalability of the Client System. The engineer will define and track Key... 
    Live in
    Work at office
    Local area

    Accenture

    Arlington, VA
    3 days ago
  • $81.1k - $187k

     ...Job Description We are looking for a Site Reliability Engineer 3 to support mission-critical cloud services and production operations. The role focuses on improving service reliability, reducing operational risk, automating repetitive tasks, and driving faster detection... 
    Temporary work
    Immediate start
    Flexible hours
    Shift work

    Oracle

    Washington DC
    2 days ago
  • $98.16k - $159.27k

     ...Solutions Job Description: The Oracle Application DBA Engineer II develops and maintains technical solutions that adhere to engineering...  .... Provides technical expertise with a focus on efficiency, reliability, scalability, and security; includes planning, evaluating,... 
    Work at office
    Local area
    Work from home
    Flexible hours

    TD Bank

    Laurel, MD
    2 days ago
  • $51.9 per hour

     ...OVERVIEW: This job is responsible for the reliability, availability, and performance of...  ...operational efficiency. This role blends software engineering, clinical engineering, and security...  .... Works cross-functionally with AHN site leaders and teams to navigate and to monitor... 
    For contractors
    Local area

    Highmark Health

    Washington DC
    1 day ago
  • $84.9k - $209.5k

     ...spirit that promotes an upbeat and creative environment. We are unencumbered and will need your contribution to make it a special engineering center with the focus on excellence. Health Data Intelligence Platform has a rare opportunity to play a critical role in how... 
    Temporary work
    Immediate start
    Flexible hours

    Oracle

    Washington DC
    3 days ago
  • $131k - $164k

     ...Staff Site Reliability Engineer New York, New York, United States Position Overview We are seeking a highly skilled Staff Site Reliability Engineer with deep technical expertise across VMware, Linux, and automation frameworks, to join our global Infrastructure... 
    Work at office
    Local area
    Flexible hours

    Diligent

    Washington DC
    4 days ago
  • $220k - $250k

     ...Staff Site Reliability Engineer Yugabyte is the company behind YugabyteDB, the AI-ready, multi-modal, distributed PostgreSQL database for cloud-native apps. Trusted by industry leaders including Shopify, Paramount+, GM, Kroger, Fiserv, and NPCI, YugabyteDB has been... 
    H1b
    Local area
    Worldwide
    Visa sponsorship

    YugaByte

    Washington DC
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Site Reliability Engineer II. Be the first to apply!