Site Reliability Engineer II
$95k - $171kAkamai
Are you passionate about cutting-edge AI infrastructure?
Do you want to build your SRE career on one of the most exciting platforms in cloud computing?
Join the Akamai Inference Cloud Team
The Akamai Inference Cloud team is part of Akamai's Cloud Technology Group. We design, implement, deploy and operate AI platforms that enable customers to run inference models and developers to create AI applications.
Partner with the best
In this role, responsibilities will include automation, monitoring, incident response, and working collaboratively with skilled team members. Candidates should possess expertise in Linux systems, automation, and SRE practices. Daily activities involve coding, improving dashboards, enhancing alerts, and minimizing repetitive tasks. Opportunities exist to focus on GPU infrastructure, Kubernetes, and ensuring reliability for AI workloads within Akamai's serverless inference platform.
As an Site Reliability Engineer II, you will be responsible for:
Building and maintaining dashboards, alerts, and monitoring for inference workloads using Akamai's existing observability platform
Writing automation and tooling in Python or Go to reduce operational toil and improve system reliability
Building and improving runbooks for inference-specific operational procedures, integrating into Akamai's existing incident management processes
Contributing to SLO tracking and reporting, identifying trends and areas for improvement
Supporting CI/CD pipeline maintenance, deployment safety checks, and rollback procedures
Collaborating with product engineering teams to troubleshoot complex problems across the stack
Participating in on-call rotations, responding to production incidents, and conducting blameless post-mortems
Do what you love
To be successful in this role you will:
Have 2+ years of experience in Site Reliability Engineering and a Bachelor's Degree or its equivalent experience
Demonstrate coding ability in at least one programming language (Python or Go) with experience writing automation
Have experience with Linux systems administration and the ability to troubleshoot complex infrastructure issues
Show familiarity with Kubernetes and containerization concepts
Have experience with monitoring and observability tools such as Prometheus, Grafana, or similar
Have exposure to CI/CD pipelines and infrastructure-as-code tools (Terraform, SaltStack, or equivalent)
Show a willingness to learn and grow, with genuine curiosity about AI infrastructure and distributed systems
Work in a way that works for you
FlexBase, Akamai's Global Flexible Working Program, is based on the principles that are helping us create the best workplace in the world. When our colleagues said that flexible working was important to them, we listened. We also know flexible working is important to many of the incredible people considering joining Akamai. FlexBase, gives 95% of employees the choice to work from their home, their office, or both (in the country advertised). This permanent workplace flexibility program is consistent and fair globally, to help us find incredible talent, virtually anywhere. We are happy to discuss working options for this role and encourage you to speak with your recruiter in more detail when you apply.
Learn ( what makes Akamai a great place to work
Connect with us on social and see what life at Akamai is like!
We power and protect life online, by solving the toughest challenges, together.
At Akamai, we're curious, innovative, collaborative and tenacious. We celebrate diversity of thought and we hold an unwavering belief that we can make a meaningful difference. Our teams use their global perspectives to put customers at the forefront of everything they do, so if you are people-centric, you'll thrive here.
Working for you
At Akamai, we will provide you with opportunities to grow, flourish, and achieve great things. Our benefit options are designed to meet your individual needs for today and in the future. We provide benefits surrounding all aspects of your life:
Your health
Your finances
Your family
Your time at work
Your time pursuing other endeavors
Our benefit plan options are designed to meet your individual needs and budget, both today and in the future.
About us
Akamai powers and protects life online. Leading companies worldwide choose Akamai to build, deliver, and secure their digital experiences helping billions of people live, work, and play every day. With the world's most distributed compute platform from cloud to edge we make it easy for customers to develop and run applications, while we keep experiences closer to users and threats farther away.
Join us
Are you seeking an opportunity to make a real difference in a company with a global reach and exciting services and clients? Come join us and grow with a team of people who will energize and inspire you!
#LI-Remote
Compensation
Akamai is committed to fair and equitable compensation practices. For US based candidates only - the base salary for this position ranges from $95,000 - $171,000/year; a candidate's salary is determined by various factors including, but not limited to, relevant work experience, skills, certifications and location. Compensation for candidates outside the US will vary. The compensation package may also include incentive compensation opportunities in the form of annual bonus or incentives, equity awards and an Employee Stock Purchase Plan (ESPP). Akamai provides industry-leading benefits including healthcare, 401K savings plan, company holidays, vacation (in the form of PTO), sick time, family friendly benefits including parental leave and an employee assistance program including a focus on mental and financial wellness; Eligibility requirements apply.
Equal Employment Opportunity Rights
Akamai Technologies is an Affirmative Action, Equal Opportunity Employer that values the strength that diversity brings to the workplace. All qualified applicants will receive consideration for employment and will not be discriminated against on the basis of gender, gender identity, sexual orientation, race/ethnicity, protected veteran status, disability, or other protected group status.
- ...Systems Engineer II The Department of Administration's (Admin) Office of Technology & Information... ...lifecycle management policy for groups, sites, and collaboration policies. Perform... ...approved measures to increase security, reliability, and confidentiality of implemented...SuggestedFull timeWork at officeRemote workShift workNight shiftAfternoon shift
$168k - $200k
...is passionate about creating transformative change in healthcare. What We're Looking For We're looking for a Senior Site Reliability Engineer to join our Data & ML Platform team. You'll be at the forefront of building and operating a resilient, observable, and...Suggested$110k - $125k
...The Site Reliability Engineer maintains and improves the availability, performance, resilience, and operational recoverability of enterprise identity, credential, and access-management services. This role supports cloud identity and directory platforms, including Microsoft...SuggestedContract workWork at office$75.7k - $136.3k
...solve complex challenges? Do you have a passion for automation and building systems that scale? Join our highly skilled Site Reliability Engineering team! Our team designs, develops, and manages applications and infrastructure that support Akamai Cloud's products and...SuggestedWork experience placementWork at office$169.3k - $304.7k
...in building and maintaining fast, efficient, scalable, and reliable routing software and infrastructure that is responsible... ...growth and stability of our global platform. As a Principal Site Reliability Engineer - Network, you will be responsible for: Architecting,...SuggestedWork experience placementWork at office$55k - $187k
...Internal Firm Services - Other Management Level Senior Associate Job Description & Summary The Opportunity As a Site Reliability Engineer - Senior Associate, you will play a pivotal role in enhancing the reliability, scalability, and performance of our...Full timeH1b- ...Job Description Job Description Job Title: Senior AWS Site Reliability Engineer (SRE) Location: Birmingham, Alabama Type: Contract To Hire Work Model: Onsite – onsite Hours: 40.0 Security Clearance: Overview Responsibilities Implement and improve...Contract workLocal area
$86.4k - $129.6k
Join our Electrical Systems Design team as an Engineer II and bring your technical expertise to projects that power critical infrastructure... ..., PLC transfers and power monitoring. This position works on-site in Columbia, SC.What will you do?Design and optimize electrical...Full timeTemporary workWorldwideFlexible hours- ...Enforcement Division (SLED) is seeking a highly skilled Senior Systems Engineer to provide advanced-level design, implementation, and support... ..., and backup platforms to ensure maximum performance and reliability. Serve as the technical lead for infrastructure modernization...Local areaNight shift
- ...Multimodality Technologist II- Full-Time Multimodality Technologist II- Full-Time Location: Columbia, South Carolina Employer: Healthcare Support Department: Clinical & Research Support Services Employment Type: Full Time Organization: Hospital Authority...Full timeWork experience placementFlexible hours
- ...Job Description Job Description Job Description Summary The System Engineer II, Network reports to the Manager, Network Infrastructure and Security Engineering in support of MUSC’s academic, research and healthcare missions. Under general supervision, this position...Work experience placementShift workRotating shift
$94.1k - $129.4k
...eligible for an incentive bonus.Position Overview:The Process Engineer II role provides technical leadership to operate manufacturing assets... ...processes.Assist in the development and implementation of the site's quality management system.Work closely with internal and...Temporary workFlexible hours$190k - $205k
...Full-time Description ALEX is seeking two cleared Software Engineers (Level 2) to support software development and full life cycle... ...program at Columbia, MD / Annapolis Junction, MD. This is an on-site position in a Sensitive Compartmented Information Facility (...Full time$94.1k - $152.8k
...Position Overview The ServiceNow Engineer provides advanced engineering, configuration, and integration support for the enterprise ServiceNow... ...with other enterprise tools. Mapped to the Cloud Engineer II standard job, the position also applies modern engineering...Contract workWork at office- ...software development skills to join the Web Application Development team in continuing the delivery of the new Economic Services Re-Engineering projects and the Refugee Management System. Daily Duties / Responsibilities Responsibilities include code development,...Work experience placement
- ...equipment, process efficiency, and plant wide safetyProvide technical assistance in the scope of engineering, preventative and predictive maintenance, and other facets of reliability engineeringUpdate existing plant and maintenance equipment records and processesMonitor and...Work at officeLocal areaImmediate start
- ...application tier utilizes MicroFocus Visual COBOL running under IIS on MS Windows 2022 servers. • This assignment will support development... ..., and reports involving all SCDMV business areas. • This engineer will work closely and will be involved with both the UI tiers which...For contractors
- Engineer, administer, and secure enterprise Windows and RHEL systems Maintain and enforce baseline configurations Administer and optimize... ...certifications qualified as Information Assurance Technician (IAT) Level II, e.g. CompTIA Security+ 10+ years of experience with C4I systems...
- ...providers, applying principles and techniques of computer science, engineering, and mathematical analysis. - Plan and conduct technical... ...-related experience is required. - Security+ or DoD 8570 IAT-II certification required. - Ability to travel 10% - 20%....Minimum wageContract workTemporary workWork experience placement
- ...About the Role: As a CBRE Reliability Engineer, you will support assetreliability and performance through Predictive Maintenance (PdM), conditionmonitoring... ...the network. This role requires up to 75% travelto support site reliability and training needs. What You’ll Do: Serve...Hourly payWork at officeShift work
- ...Reliability Engineer The Reliability Engineer is responsible for overall evaluation and improvement of product reliability, through collecting and analyzing failure data, recommending design and or process improvement, support MTBF modeling activities, interfacing...
$101.2k - $126.5k
...thinking.Woolpert is an award-winning, global leader in architecture, engineering, and geospatial services. We blend design excellence with... ...for career growth.Position OverviewWoolpert is hiring a Site Civil Engineer (PE) to join our dynamic Land Development Engineering...Local areaFlexible hoursNight shift- ...deliver with MORE Paschen . Position Overview: Project Engineer will be responsible for assisting the project manager and... ...completion. Assigned Responsibilities*: Assist with on-site management to ensure project success. Ensure project plan is...Work at officeFlexible hours
- ...019 MS SQL Server 2016/2019 & Reporting Services TFS (Team Foundation Server) Angular, AngularJS, PrimeNG Windows Server / IIS / Active Directory PowerShell scripting Soft Skills Strong communication skills (technical & non-technical stakeholders) Ability...
- ...Description Raba Kistner, Inc. is a premier Engineering Consulting and Program Management firm.... ..., dependable Engineer-In-Training I, II, or III to join our Infrastructure team... ...heavy equipment on construction or roadway sites; potential exposure to hazardous dangerous...Work at officeImmediate start
- ...with modern data processing technologies such as Python, Java, Airflow, etc. using data mesh & data lake. Contribute to the data engineering community and advocate for agency data practices. Engage in hands-on coding and design to implement production solutions. Optimize...Contract workWork at officeRemote workRelocation
- Companion Data Services, a BlueCross BlueShield of South Carolina subsidiary, seeks an experienced systems programmer to maintain and enhance our enterprise systems infrastructure in a fast-paced, multi-platform environment. You will write and debug programs for operating...
$102.3k - $209.5k
...software, firmware, and hardware layers, working in DevOps and incident response paradigms to maintain reliability and availability. Work closely with platform engineering, operations, firmware development, silicon/board vendors, and data center teams to support and...Temporary workImmediate startFlexible hoursShift work- ...providers, applying principles and techniques of computer science, engineering, and mathematical analysis. - Plan and conduct technical tasks... ...-related experience is required. - Security+ or DoD 8570 IAT-II certification required. - Ability to travel 10% - 20%....Minimum wageContract workTemporary workWork experience placement
- ...Multi-Modality Technologist II Garners Ferry FSED-1 MUSC Health Emergency and Urgent Care, a part of MUSC Health Columbia Medical Center Downtown As the health care system of the Medical University of South Carolina, MUSC Health is dedicated to delivering the highest...Hourly payWork experience placementReliefFlexible hoursShift work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer II. Be the first to apply!
- IT site lead Columbia, SC
- site safety Columbia, SC
- site leader Columbia, SC
- on-site clinical research associate (traveling/remote) Columbia, SC
- junior website developer Columbia, SC
- historic site Columbia, SC
- on site coordinator Columbia, SC
- site recruiter Columbia, SC
- construction site safety Columbia, SC
- official site Columbia, SC


