Site Reliability Engineer II
$95k - $171kAkamai Technologies
Are you passionate about cutting-edge AI infrastructure?
Do you want to build your SRE career on one of the most exciting platforms in cloud computing?
Join the Akamai Inference Cloud Team
The Akamai Inference Cloud team is part of Akamai's Cloud Technology Group. We design, implement, deploy and operate AI platforms that enable customers to run inference models and developers to create AI applications.
Partner with the best
In this role, responsibilities will include automation, monitoring, incident response, and working collaboratively with skilled team members. Candidates should possess expertise in Linux systems, automation, and SRE practices. Daily activities involve coding, improving dashboards, enhancing alerts, and minimizing repetitive tasks. Opportunities exist to focus on GPU infrastructure, Kubernetes, and ensuring reliability for AI workloads within Akamai's serverless inference platform.
As an Site Reliability Engineer II, you will be responsible for:
Building and maintaining dashboards, alerts, and monitoring for inference workloads using Akamai's existing observability platform
Writing automation and tooling in Python or Go to reduce operational toil and improve system reliability
Building and improving runbooks for inference-specific operational procedures, integrating into Akamai's existing incident management processes
Contributing to SLO tracking and reporting, identifying trends and areas for improvement
Supporting CI/CD pipeline maintenance, deployment safety checks, and rollback procedures
Collaborating with product engineering teams to troubleshoot complex problems across the stack
Participating in on-call rotations, responding to production incidents, and conducting blameless post-mortems
Do what you love
To be successful in this role you will:
Have 2+ years of experience in Site Reliability Engineering and a Bachelor's Degree or its equivalent experience
Demonstrate coding ability in at least one programming language (Python or Go) with experience writing automation
Have experience with Linux systems administration and the ability to troubleshoot complex infrastructure issues
Show familiarity with Kubernetes and containerization concepts
Have experience with monitoring and observability tools such as Prometheus, Grafana, or similar
Have exposure to CI/CD pipelines and infrastructure-as-code tools (Terraform, SaltStack, or equivalent)
Show a willingness to learn and grow, with genuine curiosity about AI infrastructure and distributed systems
Work in a way that works for you
FlexBase, Akamai's Global Flexible Working Program, is based on the principles that are helping us create the best workplace in the world. When our colleagues said that flexible working was important to them, we listened. We also know flexible working is important to many of the incredible people considering joining Akamai. FlexBase, gives 95% of employees the choice to work from their home, their office, or both (in the country advertised). This permanent workplace flexibility program is consistent and fair globally, to help us find incredible talent, virtually anywhere. We are happy to discuss working options for this role and encourage you to speak with your recruiter in more detail when you apply.
Learn ( what makes Akamai a great place to work
Connect with us on social and see what life at Akamai is like!
We power and protect life online, by solving the toughest challenges, together.
At Akamai, we're curious, innovative, collaborative and tenacious. We celebrate diversity of thought and we hold an unwavering belief that we can make a meaningful difference. Our teams use their global perspectives to put customers at the forefront of everything they do, so if you are people-centric, you'll thrive here.
Working for you
At Akamai, we will provide you with opportunities to grow, flourish, and achieve great things. Our benefit options are designed to meet your individual needs for today and in the future. We provide benefits surrounding all aspects of your life:
Your health
Your finances
Your family
Your time at work
Your time pursuing other endeavors
Our benefit plan options are designed to meet your individual needs and budget, both today and in the future.
About us
Akamai powers and protects life online. Leading companies worldwide choose Akamai to build, deliver, and secure their digital experiences helping billions of people live, work, and play every day. With the world's most distributed compute platform from cloud to edge we make it easy for customers to develop and run applications, while we keep experiences closer to users and threats farther away.
Join us
Are you seeking an opportunity to make a real difference in a company with a global reach and exciting services and clients? Come join us and grow with a team of people who will energize and inspire you!
#LI-Remote
Compensation
Akamai is committed to fair and equitable compensation practices. For US based candidates only - the base salary for this position ranges from $95,000 - $171,000/year; a candidate's salary is determined by various factors including, but not limited to, relevant work experience, skills, certifications and location. Compensation for candidates outside the US will vary. The compensation package may also include incentive compensation opportunities in the form of annual bonus or incentives, equity awards and an Employee Stock Purchase Plan (ESPP). Akamai provides industry-leading benefits including healthcare, 401K savings plan, company holidays, vacation (in the form of PTO), sick time, family friendly benefits including parental leave and an employee assistance program including a focus on mental and financial wellness; Eligibility requirements apply.
Equal Employment Opportunity Rights
Akamai Technologies is an Affirmative Action, Equal Opportunity Employer that values the strength that diversity brings to the workplace. All qualified applicants will receive consideration for employment and will not be discriminated against on the basis of gender, gender identity, sexual orientation, race/ethnicity, protected veteran status, disability, or other protected group status.
$98.58k - $138.02k
...Site Reliability Engineer II Restaurant365 is a SaaS company disrupting the restaurant industry! Our cloud-based platform provides a unique, centralized solution for accounting and back-office operations for restaurants. Restaurant365's culture is focused on empowering...SuggestedWork at office$123.71k - $173.2k
...Advanced Concepts and Enterprise Engineering (ACE), supporting Blue Origin’... .... As a Software Engineer II, you will join a team of passionate... .... This role will propel the reliable operation of the company and... ...4,961.00 - $188,945.40 Other site ranges may differ Culture...SuggestedPermanent employmentFull timeTemporary workLocal area$114.5k - $188k
...engagement together!How You Will Make an ImpactAs a Software Engineer II on the Developer Platform Team, you will help improve the Scala... ...to change, the build graph easier to reason about, and CI more reliable under real production-scale usage.This is a hands-on...SuggestedContract workLocal areaImmediate startRemote workWorldwideHome office$78k - $170k
...into global manufacturing capacity.Xometry is seeking a Software Engineer II to join our Xometry Partner Experience Technology Organization... ...expertise and technical leadership to help us build fast, reliable, and intuitive solutions that empower our partners - enabling...Suggested$68.9k - $131.1k
...than 100 years of experience and renowned engineering expertise to meet the needs of today's... ...exciting opportunity for a Systems Engineer II - Modeling, Simulation & Analysis... ...support mission success. The work is on-site in Aurora, CO.What You Will DoDevelop high...SuggestedTemporary workWork experience placementWork at officeRemote workRelocationFlexible hours$68.9k - $131.1k
...than 100 years of experience and renowned engineering expertise to meet the needs of today's... ...exciting opportunity for a Systems Engineer II - EO/IR Processing supporting one of our... ...and mission effectiveness. The work is on-site in Aurora, CO.What You Will DoApply...Temporary workWork experience placementWork at officeRemote workRelocationFlexible hours$197.4k - $232k
...Location Type: Remote Department Engineering Compensation: $197.4K - $232K -... ...the Role: Senior Software Engineers II at Confluent take ownership of critical... ...architecture and technical decisions that balance reliability, scalability, performance, and...Full timeRemote work$160k - $195k
...Senior Software Engineer II Guild is seeking a Senior Software Engineer II to join our Member Experiences pillar, focused on building... ...; to identify and implement enhancements in performance, reliability, and developer velocity. Take ownership of resolving issues and...$70k - $109k
...Location: Denver, CO CONMED is seeking a Software Engineer II to join the Advanced Surgical R&D team based in Denver, CO. This role will partner with cross-functional teams to design, develop, verify, and maintain secure software components and subsystems for medical...Temporary workWork experience placementImmediate startWorldwide$90k - $155k
...ambitious results.It’s the people. Our team is our competitive advantage and we are better together.YOUR MISSIONAs a Flight Software Engineer II at True Anomaly, you will gain experience at every phase of the spacecraft program including design, analysis, manufacturing,...Permanent employment$62k - $141k
Site Reliability EngineerThe Opportunity: Engineering to make a system more resilient and efficient frees up time and money to build more capabilities. Whether you come from a background in network engineering, systems administration, or software development, if you have...Full timeContract workPart timeWork at officeLocal areaRemote work$105.6k - $145.2k
Architect the Future as our Site Reliability Engineer!Are you ready to take your skills to the next level as a self-motivated and enthusiastic Site Reliability Engineer with hands-on experience supporting multiple connected Cloud-based products? Trimble is a global technology...Ongoing contractFull timeWork at officeLocal areaWorldwide$104.43k - $156.65k
...Comcast. (In most cases, Comcast prefers to have employees on-site collaborating unless the team has been designated as virtual... ..., Fox, Disney, NBC, Paramount+, and many others.Our Site Reliability Engineering (SRE) team is at the heart of our mission to deliver seamless...Permanent employmentFull timeWork at officeRemote workWorldwideFlexible hours$100k - $115k
...Internal Developer Platform (IDP) as a product, treating engineering teams as customers and optimizing for reliability, usability, and delivery velocity.Define and... ....4+ years of experience in Platform Engineering, Site Reliability Engineering, DevOps, or Systems Engineering...Temporary work$115k - $170k
SENIOR SOFTWARE ENGINEER I/II - DIGITAL ENGINEERING Are you passionate about advancing aerospace technology through cutting‑edge software... ...join our Digital Engineering team based out of our Littleton, CO site. You will play a pivotal role in developing next‑generation...Permanent employmentWork at officeLocal area$125k - $160k
...technology through cutting‑edge software solutions? We're looking for a driven Senior Software Engineer I/II to join our Digital Engineering team, based out of our Littleton, CO site. You will play a pivotal role in developing next-generation simulation and modeling tools...Permanent employmentLocal area$120k - $160k
...ambitious space missions. SENIOR SOFTWARE ENGINEER I/II – SIMULATION ENGINEERING We are seeking... ...team, based out of our Littleton, CO site. In this role, you will focus on... ...simulation results to ensure accuracy and reliability. Collaborate with domain experts to define...Permanent employmentLocal area$129.2k - $174.8k
...and Agentic AI services.We are seeking an Systems Development Engineer II who can think big and simplify solutions to complex problems,... ...)- 1+ years of designing or architecting (design patterns, reliability and scaling) of new and existing systems experience- Current,...InternshipFlexible hours- ...Senior Site Reliability Engineer (Enterprise Platform) Location: Remote - US - Open to Europe if happy to overlap with EST Compensation: Competitive We are a high-growth software company supporting the development of a premier open-source, EVM-compatible public ledger...Contract workCurrently hiringRemote work
$127k - $249k
The TeamPlatform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions... ...fleet, alongside the critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager, and Gatekeeper)....Work at officeLocal areaRemote workWorldwideFlexible hours$87.4k - $123.4k
...the U.S. We are unable to sponsor or take over sponsorship of an employment visa at this time, including CPT/OPT.*** The Site Reliability Engineer will help ensure the reliability, scalability, and performance of Empower’s financial services platform. This person will...16 hoursContract workTemporary workWork experience placementCasual workWork at officeLocal areaRemote workWork from homeWork visaFlexible hours$121.4k - $218.6k
...will be responsible for ensuring best-in-class uptime and reliability of our AI hardware infrastructure offerings. Partner with... ...and defend them when they are breached. As a Senior Site Reliability Engineer, you will be responsible for: Developing and scaling robust...Work experience placementWork at office$160k - $190k
...Site Reliability Engineer (Classified Deployments) Location: Southern California or Washington, D.C. Clearance: Active Secret required; TS/SCI strongly preferred Work Mode: Hybrid/On-site with government customers Citizenship: U.S. Citizen Compensation:...$94.85k - $135.5k
...powered business communications. This is where you and your skills come in. We're currently looking for: An experienced Site Reliability Engineer (SRE) to join the RingCentral Collaboration team. As a SRE, you will be responsible for maintaining and improving uptime...Full timeLocal areaFlexible hours$118.7k - $218.6k
Position Summary Lead AI and Data Science Engineer II Drive the design and delivery of advanced analytics, artificial intelligence (AI), and generative artificial intelligence (GenAI) solutions that inform Talent strategy and workforce decisions. In this role, you...- ...Dedicated Cloud (ADC) Security Systems Engineering (AS2E) team is responsible for building,... ...the lifeAs a System Development Engineer II, you will collaborate with internal service... ...or architecting (design patterns, reliability and scaling) of new and existing systems...Internship
$100k - $130k
...driven and highly collaborative team that is passionate about creating transformative change in healthcare. We are seeking a Site Reliability Engineer to play a key role in designing, optimizing, and securing the underlying cloud infrastructure that powers our organization...Flexible hours$120k - $225k
...in AI as a core operating capability, not a side experiment, and this role exists to help make that investment real.As a Platform Engineer on the AI team, you will build the infrastructure, tooling, and integrations that put AI directly into the hands and workflows of...Permanent employmentWork at officeShift work3 days per week$110k - $145k
...content reflecting our world. NBCU's Distribution engineering is responsible for the automation and reliability of NBCU's Live sources. Reasonable for the... ...Distribution Engineering is looking to add a talented Site Reliability Engineer to be part of our Video Streaming...Work experience placementWork at officeLocal area- ...focusing on private cloud systems supporting 5G wireless systems. This position will focus on platform monitoring, logging, and reliability aspects supporting the Mobile Core team. A critical goal is to gather metrics of the platform during stress and load events to ensure...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer II. Be the first to apply!

