Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Site Reliability Engineer

$150.4k - $277.6k

Apple Oakbrook

Technical Operations & Site Reliability Engineer, Customer SystemsAt Apple, Customer Experience is at the forefront of everything we do. The Customer Systems Operations team is looking for a highly skilled and motivated TechOps Engineer (Technical Operations & Site Reliability) to join us. The team is responsible for maintaining the reliability, availability, and performance of business-critical, globally distributed systems. If you have the desire and motivation to design and develop automation solutions to streamline system sustenance, monitoring, and operational workflows, while collaborating closely with support, engineering and business operations teams, this profile is for you. Ideal candidates will combine a passion for operational excellence with strong software engineering skills, and thrive in a fast-paced, change-driven environment focused on continuous improvement and flawless delivery.Manage large-scale production outages, leading incident response and improving efficiency. Design, build, and maintain automation solutions to streamline the monitoring, sustenance, and management of large-scale distributed systems. Develop tools and software (using Java/JEE, REST, Swift/Objective C, Python, Go, or Bash) to automate repetitive operational tasks, reduce manual intervention, and improve system reliability. Utilize AI & LLM models to achieve Operational Excellence in application support. Plan and execute actionable system health monitoring, incident response, and communication across critical global applications. Drive operational metrics and KPI identification and alignment. Partner with multi-functional teams to improve reliability, efficiency, stability, and processes. Be a self-directed problem-solver exhibiting deftness to handle multiple simultaneous competing priorities and deliver solutions in a timely manner. Create and maintain accurate, up-to-date documentation reflecting architecture, infra configuration, and procedures. Write status and incident reports. Write training material and train users in complex topics. Partner with a team of highly skilled engineers across the globe and guide their work towards operational excellence, gaining efficiency. Build a culture where the regional members are responsible for cultivating strong in-region relationships and getting results for our business partners ensuring they remain informed about significant incidents and problems.Minimum QualificationsExperience in interpreting operational data from systems like Hubble, ExtraHop, Splunk or other monitoring tools along with hands-on experience of production monitoring systems, log analysis, troubleshooting, and support dashboards.B.S in Computer Science, Computer Engineering or similar or equivalent related work experience.Understanding of standard networking protocols and components such as: DNS, TCP/IP, ICMP, the OSI Model, Subnetting and Load Balancing.Experience in using AI and Large Language Models (LLMs) to enhance operational efficiency through tasks such as model training, optimization (including areas like Model Context Protocol or similar methods), and designing effective model utilities.Experience in scripting languages and automation tools such as Java, JEE, REST, Swift/Objective C, database schema design and data access technologies.Preferred QualificationsExperience in strategizing and achieving operational excellence in global distributed systems.Fundamental understanding of distributed systems including: Micro services, Messaging Brokers, and Versioning.Experience in driving operations teams for large-scale mission-critical applications working in a 24x7 environment across multiple locations.Understanding of the Linux Operating System, including Kernel, Memory, Process, Threads, Static / Shared Libraries, IPC, and Signals.Excellent organizational and documentation skills.Excellent interpersonal skills. Proactive, with a strong sense of personal ownership.Pay & BenefitsAt Apple, base pay is one part of our total compensation package and is determined within a range. This provides the opportunity to progress as you grow and develop within a role. The base pay range for this role is between $150,400 and $277,600, and your base pay will depend on your skills, qualifications, experience, and location. Apple employees also have the opportunity to become an Apple shareholder through participation in Apple's discretionary employee stock programs. Apple employees are eligible for discretionary restricted stock unit awards, and can purchase Apple stock at a discount if voluntarily participating in Apple's Employee Stock Purchase Plan. You'll also receive benefits including: Comprehensive medical and dental coverage, retirement benefits, a range of discounted products and free services, and for formal education related to advancing your career at Apple, reimbursement for certain educational expenses — including tuition. Additionally, this role might be eligible for discretionary bonuses or commission payments as well as relocation. Learn more about Apple Benefits Note: Apple benefit, compensation and employee stock programs are subject to eligibility requirements and other terms of the applicable plan or program. Apple is an equal opportunity employer that is committed to inclusion and diversity. We seek to promote equal opportunity for all applicants without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, Veteran status, or other legally protected characteristics. Learn more about your EEO rights as an applicant At Apple, we believe accessibility is a fundamental human right. You'll find that idea reflected in everything here — in our culture, our benefits and our digital tools. By welcoming as many perspectives as possible, we help you build a career where you feel like you belong. Learn about accessibility in Apple's workplace Learn about reasonable accommodations for job applicants Apple accepts applications to this posting on an ongoing basis.

Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Site Reliability Engineer in Sunnyvale, CA vacancy
  • $168k - $270.25k

    NVIDIA is looking for a Senior Site Reliability Engineer (SRE) to join its GeForce Now (GFN) team. SRE at NVIDIA ensures that our internal and external-facing GPU cloud gaming services have reliability and uptime as promised to the users and at the same time enables developers... 
    Suggested
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  • $128.6k - $184.9k

     ...global cloud platform. As a team of six engineers distributed across the US, Canada, and the...  ...with a strong focus on automation, reliability, and operational excellence. We are one...  ...Qualifications7+ years of experience in Site Reliability Engineering, DevOps, Infrastructure... 
    Suggested
    Permanent employment
    Full time
    Temporary work
    Local area
    Worldwide
    Flexible hours

    CISCO Systems

    Santa Clara, CA
    19 hours ago
  • $230k - $250k

     ...network. It's the foundation for autonomous networking, giving engineers and AI agents the ability to know the impact of every change...  ...how things have always been done.Forward is looking for a Site Reliability EngineerAbout the Role This is not a "keep the lights on"... 
    Suggested
    Night shift

    Forward Networks

    Santa Clara, CA
    2 days ago
  • $170k - $200k

    We are seeking a talented and motivated Site Reliability Engineer to join our engineering team. You will be responsible for building, maintaining, and troubleshooting cloud service/cluster, infrastructure, and monitoring systems to ensure high availability, performance,... 
    Suggested
    Full time
    Worldwide

    Fortinet

    Sunnyvale, CA
    4 days ago
  • LeanData helps the world’s fastest-growing companies automate, simplify, and accelerate revenue.We are looking for a Senior Site Reliability Engineer to lead the strategic evolution of our cloud infrastructure. Reporting directly to the SVP of Engineering, this role is... 
    Suggested
    Full time
    Work at office
    2 days per week

    LeanData

    Santa Clara, CA
    2 days ago
  • $148k - $235.75k

     ...see how you can make a lasting impact on the world.Join our team of innovative engineers who are building an AI Data Center AIOps platform that turns raw, high-volume telemetry into reliable, job-centric insights and automation for GPU fleets. We’re hiring a DevOps Engineer... 
    Full time

    Nvidia

    Santa Clara, CA
    4 days ago
  • $167.7k - $245.2k

     ...requiring approximately 2 days per week on-site at Cisco offices in either San Francisco...  ...AI agents behave as intended, improving reliability and reducing risks. This unified...  ...and control.As a Senior Site Reliability Engineer (SRE), you will build, operate, and continuously... 
    Full time
    Temporary work
    Local area
    Flexible hours
    2 days per week

    CISCO Systems

    Sunnyvale, CA
    1 day ago
  • $90k - $180k

     ...nutritionals and branded generic medicines. Our 115,000 colleagues serve people in more than 160 countries.About the RoleThis Senior Site Reliability Engineer position works on-site out of our Sylmar, CA or Sunnyvale, CA location in the Cardiac Rhythm Management Division.We are... 
    Remote work

    Abbott

    Sunnyvale, CA
    3 days ago
  • $152k - $241.5k

     ...infrastructure platforms for automated host lifecycle management, fleet reliability/auto-healing, E2E observability or data-driven operations (...  ...languages such as Python, Go, Perl, or Ruby.Mentored other engineers and influenced technical direction through design reviews,... 
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $101k - $161k

     ...excellence has earned us several prestigious awards, such as Best Engineering Team, Best Company for Diversity, Compensation, and Work-...  ...we do.Job DescriptionWho You'll Work WithWe’re looking for Site Reliability Engineers to join our growing Arista’s CloudVision-as-a-... 

    Arista Networks

    Santa Clara, CA
    5 days ago
  • $168k - $270.25k

     ...phenomenal people like you to help us accelerate the next wave of artificial intelligence.Join our team at NVIDIA as a Senior Site reliability engineer focused on HPC storage and play a crucial role in designing, implementing, and optimizing on-prem High-Performance... 
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  •  ...Overview Title: Site Reliability Engineer SRE – ML platform Location: Austin, TX or Sunnyvale, CA Employment type: Full-time • Seniority: Mid-Senior level • ONLY W2 Responsibilities Continuous Deployment using GitHub Actions, Flux, Kustomize Design and implement cloud... 
    Full time

    Saransh

    Sunnyvale, CA
    2 days ago
  •  ...Google is seeking a Software Engineering Manager II in Site Reliability Engineering, based in Sunnyvale, California. This onsite role leads a team to ensure reliability and performance of critical systems, partnering with product and engineering teams to deliver scalable... 

    Epic Games

    Sunnyvale, CA
    2 days ago
  • Remote DevOp/SRE With AI-First MindsetInsight Global is looking for a remote, DevOp/SRE with an AI-first mindset coming from a start up background to join one of our cyber security customers in the Bay Area. This role can pay 140-160k based on years of experience and skillset...
    Remote work

    Insight Global

    Santa Clara, CA
    3 days ago
  •  ...design by customizing MES tool per business needs Education Requirements, Ideal Experience: Associate’s degree in Industrial Engineering or IT related field Minimum of 0-3 years’ relevant experience Experience in C#, Delphi desired Knowledge of the... 
    Work at office

    Foxconn Industrial Internet - FII

    Sunnyvale, CA
    a month ago
  •  ...of Huobi globe spanning infrastructure. •       Work with engineering teams to make sure new features and changes are deployed quickly...  .... •       Constantly improve our system performance and reliability through better tools, process and monitoring system. •... 
    Worldwide

    Cryptoware Technologies Inc

    Santa Clara, CA
    12 days ago
  • $145k - $165k

     ...A technology solutions firm in Sunnyvale, CA is looking for a highly experienced Site Reliability Engineer (SRE). This role involves maintaining uptime and performance across systems. Exceptional Linux expertise and automation skills in Bash and Python are crucial. Key... 

    Bolt Graphics, Inc.

    Sunnyvale, CA
    2 days ago
  •  ...A leading technology firm is in search of a Senior Wireless Network Site Reliability Engineer to manage and enhance their wireless network infrastructure. The ideal candidate has over 8 years of experience in wireless network operations and a strong background in wireless... 

    TechDigital Group

    Santa Clara, CA
    2 days ago
  •  ...Job Description Job Description About the Role Senior Site Reliability Engineer (Payments Infrastructure) Kody is seeking a Senior Site Reliability Engineer to ensure the reliability, availability, scalability, and operational excellence of our global payment platform... 

    Kody

    Sunnyvale, CA
    a month ago
  •  ...Job Description Job Description Site Reliability Engineer Onsite- Bay Area, CA Skills Relevant Skills and Experience What You’ll Do (Day-to-Day) Own and manage our cloud infrastructure (GCP or AWS, on-prem). Build, maintain, and optimize Kubernetes... 

    Amiri Recruiting

    Mountain View, CA
    more than 2 months ago
  • $150k - $195k

     ...customers worldwide. Our team is growing, and we are looking for engineers with passion for automation. You will help support the...  ...alongside engineering/operations teams to improve the scalability and reliability of internal processes. Participate in an on‑call rotation.... 
    Full time
    Worldwide

    Fortinet

    Sunnyvale, CA
    1 day ago
  •  ...Site Reliability Engineer, Data Platform - USDS Responsibilities Engage in and improve the whole lifecycle of service, from inception and design, through to deployment, operation and refinement. Ensure reliable, fault-tolerant, efficiently scalable and cost-effective data... 

    Tik Tok

    Mountain View, CA
    2 days ago
  • $186k - $279k

     ...and instrumented end to end? At Everpure, identity is the front door to everything we build, and we're looking for an IAM Site Reliability Engineer to keep that door running smoothly — and to make it better every day. The Global Information Security Office (GISO) at... 
    Full time
    Work at office
    Flexible hours

    Everpure

    Santa Clara, CA
    3 days ago
  • $135.6k - $180k

     ...operations team. This role involves overseeing 24/7 operational stability, enhancing processes and systems, and mentoring a diverse engineering team. The ideal candidate will have over 8 years of technical operations experience, proficiency in infrastructure automation... 

    Ruckus

    Sunnyvale, CA
    2 days ago
  •  ..., and the challenges of building in a high-growth startup, we’d love to talk. This is more than a job—it’s a journey. Site Reliability Engineers (SREs) are responsible for the overall performance and reliability of ASAPP's infrastructure and products. The team owns... 
    Remote work

    ASAPP

    Mountain View, CA
    a month ago
  • $120k - $180k

     ...SRE & DevOps EngineerAt CrowdStrike, our engineering organization depends on shared infrastructure platforms that power critical product...  ...platforms require dedicated engineering ownership to operate reliably, scale safely, harden for security, and mature into self-service... 
    Work experience placement
    Work at office
    Local area

    CrowdStrike

    Sunnyvale, CA
    3 days ago
  • $101k - $161k

     ...Requirements: We require a BS or MS in Computer Science, or equivalent relevant experience. We look for 5+ years of software engineering experience. We need experience building or operating distributed database systems or scale-out applications in a SaaS... 
    Full time

    Arastra, Inc.

    Santa Clara, CA
    2 days ago
  • $186.9k - $267.7k

     ...requiring approximately 2 days per week on-site at Cisco offices in either San Francisco...  ...AI agents behave as intended, improving reliability and reducing risks. This unified...  ...and control.As a Staff Site Reliability Engineer (SRE), you will provide technical leadership... 
    Full time
    Temporary work
    Local area
    Flexible hours
    2 days per week

    CISCO Systems

    Sunnyvale, CA
    2 days ago
  • $184k - $287.5k

    At NVIDIA, Site Reliability Engineering provides a rare chance to define, develop, and support large-scale production systems with high efficiency and availability. This demanding position merges software and systems engineering efforts to guarantee flawless service operation... 
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $210.6k - $305.1k

     ...Minimum Qualifications:  You have led a distributed team of 5+ engineers, can demonstrate strong technical vision for your team, and ensure...  ..., and basic life insurance. Please see the Cisco careers site to discover more benefits and perks. Employees may be eligible... 
    Full time
    Temporary work
    Local area
    Flexible hours

    CISCO Systems

    Los Altos, CA
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!