Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Staff Software Engineer, Reliability

$156k - $255k

LinkedIn

Company Description LinkedIn is the world's largest professional network, built to create economic opportunity for every member of the global workforce. Our products help people make powerful connections, discover exciting opportunities, build necessary skills, and gain valuable insights every day. We're also committed to providing transformational opportunities for our own employees by investing in their growth. We aspire to create a culture that's built on trust, care, inclusion, and fun - where everyone can succeed. Join us to transform the way the world works. Job Description This role will be based in Mountain View, CA. At LinkedIn, our approach to flexible work is centered on trust and optimized for culture, connection, clarity, and the evolving needs of our business. The work location of this role is hybrid, meaning it will be performed both from home and from a LinkedIn office on select days, as determined by the business needs of the team. Site Health Platform sits at the core of LinkedIn's Reliability Infrastructure organization, with a primary focus on the end-to-end incident management ecosystem. Our mission is for every member and customer to experience LinkedIn as "always on", every engineer to benefit from a more insightful and proactive site-wide reliability ecosystem, and every business and product owner to be well-informed about service disruptions as they occur. We own the full incident lifecycle across thousands of services and multiple regions, from incident response and mitigation, through problem management and post-incident learning. The platforms we build are the backbone of how LinkedIn detects issues, coordinates incident response, captures context, and turns outages and near misses into structured, actionable insights. By transforming incidents into data and learnings, we enable teams to systematically improve reliability over time. Our work informs engineering priorities, infrastructure investments, capacity planning, and executive decision-making, ensuring the network is dependable when it matters most. You will be exposed to many different technologies, architectures, and systems hosted in state-of-the-art data centers across the globe. Responsibilities: Designing and evolving the core incident management platforms that power LinkedIn's full incident lifecycle, from detection and response to problem management and prevention, across thousands of services and teams. Serving in a critical on-call rotation, providing expert incident triage and coordination during high-severity outages. Partnering closely with service owners and product teams to diagnose issues quickly, mitigate member impact, and drive timely resolution under pressure. Transforming raw, unstructured incident data into clear, actionable intelligence using AI and LLM-based systems, including automated summarization, classification, root cause signals, and mitigation recommendations. Building analytics and insights that surface systemic reliability risks, recurring failure patterns, and cross-service dependencies, enabling org-level prioritization rather than isolated, service-by-service fixes. Building platforms and tools that enable realistic, fleet-wide stress testing of data center and regional capacity, validating incident readiness across dependencies, traffic patterns, and growth scenarios before they impact a significant production outage. Driving consistency, clarity, and quality in how incidents are declared, managed, reviewed, and learned from, raising the reliability bar across a large, fast-moving engineering organization. Influencing service architecture, SLOs, and reliability standards through platforms, data, and technical leadership, ensuring improvements are durable, measurable, and adopted at scale. Qualifications Basic Qualifications: Bachelor's degree in Computer Science, Engineering, or related technical field or equivalent practical experience. Many postings also prefer or require an advanced degree (MS/PhD) for Staff-level roles. 6+ years of professional experience in software development, distributed systems, or reliability engineering. Experience leading technical projects/providing architectural leadership Experience building products and operating large-scale distributed systems. Experience with two or more backend languages such as Go, Python or Java with a track record of owning complex production systems. Full-stack engineering experience, including building user-facing web applications and operational dashboards using modern frontend frameworks such as React.js, along with backend APIs and data pipelines. Understanding of web development fundamentals including API design, performance, accessibility and building intuitive interfaces for engineers and operational users. Understanding of reliability engineering principles, incident management, observability and operating systems under failure conditions. Demonstrated ability to lead technical design across teams, influence architecture beyond direct ownership and drive adoption through well-designed platforms. Debugging and root cause analysis skills, with the ability to communicate complex technical findings clearly to engineers, partners and leadership. Preferred Qualifications: Experience applying AI or LLM-based techniques to operational or incident data, including automated summarization, classification, root cause hypothesis generation or reliability recommendations. Familiarity with vector databases and retrieval-based systems used to power context-aware analytics, search or agentic workflows. Frontend engineering experience beyond basic UI, including building data-dense, high-signal interfaces for engineers using React.js, modern state management and visualization libraries. Experience designing end-to-end full-stack systems where frontend, backend, data and reliability concerns are considered holistically. Background in building internal developer platforms, observability tools, or incident response systems used at scale. A demonstrated ability to simplify complex workflows, reduce operational toil and replace manual processes with well-designed automation. Suggested Skills: High Severity Incident Response Production Troubleshooting & Root Cause Analysis Distributed Systems and Linux fundamentals Software Development (Go / Python / Java) & Architecture Observability & Incident Detection Additional Information LinkedIn is committed to fair and equitable compensation practices. The pay range for this role is $156,000 to $255,000. Actual compensation packages are based on several factors that are unique to each candidate, including but not limited to skill set, depth of experience, certifications, and specific work location. This may be different in other locations due to differences in the cost of labor. The total compensation package for this position may also include annual performance bonus, stock, benefits and/or other applicable incentive compensation plans. For more information, visit Equal Opportunity Statement We seek candidates with a wide range of perspectives and backgrounds and we are proud to be an equal opportunity employer. LinkedIn considers qualified applicants without regard to race, color, religion, creed, gender, national origin, age, disability, veteran status, marital status, pregnancy, sex, gender expression or identity, sexual orientation, citizenship, or any other legally protected class. LinkedIn is committed to offering an inclusive and accessible experience for all job seekers, including individuals with disabilities. Our goal is to foster an inclusive and accessible workplace where everyone has the opportunity to be successful. If you need a reasonable accommodation to search for a job opening, apply for a position, or participate in the interview process, connect with us at View email address on click.appcast.io and describe the specific accommodation requested for a disability-related limitation. Reasonable accommodations are modifications or adjustments to the application or hiring process that would enable you to fully participate in that process. Examples of reasonable accommodations include but are not limited to: Documents in alternate formats or read aloud to you Having interviews in an accessible location Being accompanied by a service dog Having a sign language interpreter present for the interview A request for an accommodation will be responded to within three business days. However, non-disability related requests, such as following up on an application, will not receive a response. LinkedIn will not discharge or in any other manner discriminate against employees or applicants because they have inquired about, discussed, or disclosed their own pay or the pay of another employee or applicant. However, employees who have access to the compensation information of other employees or applicants as a part of their essential job functions cannot disclose the pay of other employees or applicants to individuals who do not otherwise have access to compensation information, unless the disclosure is (a) in response to a formal complaint or charge, (b) in furtherance of an investigation, proceeding, hearing, or action, including an investigation conducted by LinkedIn, or (c) consistent with LinkedIn's legal duty to furnish information. Pay Transparency Policy Statement As a federal contractor, LinkedIn follows the Pay Transparency and non-discrimination provisions described at this link: Global Data Privacy Notice for Job Candidates Please follow this link to access the document that provides transparency around the way in which LinkedIn handles personal data of employees and job applicants: Equal Opportunity Statement We seek candidates with a wide range of perspectives and backgrounds and we are proud to be an equal opportunity employer. LinkedIn considers qualified applicants without regard to race, color, religion, creed, gender, national origin, age, disability, veteran status, marital status, pregnancy, sex, gender expression or identity, sexual orientation, citizenship, or any other legally protected class. LinkedIn is committed to offering an inclusive and accessible experience for all job seekers, including individuals with disabilities. Our goal is to foster an inclusive and accessible workplace where everyone has the opportunity to be successful. If you need a Reasonable Accommodation to search for a job opening, apply for a position, or participate in the interview process, connect with us and describe the specific Accommodation requested for a disability-related limitation. Fill out an Accommodation request here: Reasonable accommodations are modifications or adjustments to the application or hiring process that would enable you to fully participate in that process. Examples of reasonable accommodations include but are not limited to: Documents in alternate formats or read aloud to you Having interviews in an accessible location Being accompanied by a service dog Having a sign language interpreter present for the interview A request for an accommodation will be responded to within three business days. However, non-disability related requests, such as following up on an application, will not receive a response. LinkedIn will not discharge or in any other manner discriminate against employees or applicants because they have inquired about, discussed, or disclosed their own pay or the pay of another employee or applicant. However, employees who have access to the compensation information of other employees or applicants as a part of their essential job functions cannot disclose the pay of other employees or applicants to individuals who do not otherwise have access to compensation information, unless the disclosure is (a) in response to a formal complaint or charge, (b) in furtherance of an investigation, proceeding, hearing, or action, including an investigation conducted by LinkedIn, or (c) consistent with LinkedIn's legal duty to furnish information. San Francisco Fair Chance Ordinance Pursuant to the San Francisco Fair Chance Ordinance, LinkedIn will consider for employment qualified applicants with arrest and conviction records. Pay Transparency Policy Statement As a federal contractor, LinkedIn follows the Pay Transparency and non-discrimination provisions described at this link: Global Data Privacy Notice and Compliance Posters for Job Candidates Please use this link to access documents that provide information about how LinkedIn handles the personal data of employees and job applicants, as well as the E-Verify Participation Notice and the Department of Justice Immigrant and Employee Rights Section Right to Work posters: LinkedIn

Vacancy posted 17 hours ago
Similar jobs that could be interesting for youBased on the Staff Software Engineer, Reliability in Mountain View, CA vacancy
  • $251k - $310k

     ...in simulation across 15+ U.S. states. The Planner/Perception Reliability team’s goal is to build out architectures, tools, and workflows to prevent, identify, and guide fixes of reliability and software integrity issues.  We focus the organization on reliability and... 
    Suggested
    Full time
    Remote work

    Waymo

    Mountain View, CA
    a month ago
  • $218.3k - $327.5k

     ...About Team & About Role The Site Reliability Engineering (SRE) team at Rubrik ensures the absolute...  .... We operate at the intersection of software development and systems engineering,...  ...architectures, and structural resiliency. As a Staff Site Reliability Engineer, you will... 
    Suggested
    Full time
    Local area
    Shift work

    Rubrik

    Palo Alto, CA
    more than 2 months ago
  • $262k - $364k

     ...as system design consulting, developing software platforms and frameworks, capacity...  ...systems by pushing for changes that improve reliability and velocity.Practice sustainable...  ...Master's degree in Computer Science or Engineering.Site Reliability Engineering (SRE) combines... 
    Suggested

    Google

    Sunnyvale, CA
    1 day ago
  • $185k - $275k

     ...cloud: ensuring workloads run seamlessly, reliably, and efficiently across massive GPU...  ...possible with AI. What You'll Do As a Staff Engineer, you will be a technical leader shaping...  .... Who You Are ~8+ years of software engineering experience. ~ Proven track... 
    Suggested
    Permanent employment
    Full time
    Temporary work
    Casual work
    Work at office
    Flexible hours

    CoreWeave

    Sunnyvale, CA
    24 days ago
  •  ...It all started when engineer Fred Luddy wrote code that automated a tedious task for...  ...products build on — bias toward generality, reliability, clean abstractions Drive measurable...  ...technical depth What We Look For at Senior Staff Level You're a multiplier. You trace... 
    Suggested
    Full time
    Work at office
    Immediate start
    Remote work
    Flexible hours

    ServiceNow

    Mountain View, CA
    9 days ago
  • $220k - $275k

     ...Description About the Opportunity Our client is seeking a Staff Software Engineer, Embedded UI to join a highly innovative team focused on...  ...concept through implementation while ensuring quality, reliability, and maintainability. Own the software development... 
    Local area

    LHH US

    Mountain View, CA
    13 days ago
  •  ...We Are Synopsys is the leader in engineering solutions from silicon to systems, enabling...  ...usability, security, and operational reliability. Build, deploy, and operate containerized...  ...6+ years of professional experience in software engineering or a similar domain. You... 
    Immediate start
    Shift work

    Synopsys Inc

    Sunnyvale, CA
    a month ago
  • $210k - $275k

     ...clinical and administrative software for care providers. Its platform...  ...-cycle products. This is a staff-level, full-stack role with an...  ...mature production systems reliable for enterprise customers....  ...observability practices across multiple engineering teams. • Lead multi-month,... 
    Full time
    H1b

    Raydar

    Mountain View, CA
    3 days ago
  • $237k - $288k

     ...across new GPU and CPU server platforms, we're investing deeply in the firmware that underpins fleet reliability, security, and operability — and we're hiring a founding engineer to lead our BMC firmware work. You'll set the technical direction for BMC firmware across... 
    Temporary work

    Crusoe

    Sunnyvale, CA
    a month ago
  • $152k

     ...Staff Software Engineer The Resource Fabric Engineering team builds and operates the foundational infrastructure and developer tools that...  ...lifecycle across Coupang. Our mission is to provide scalable, reliable, and intelligent platforms that enable teams to efficiently... 
    Temporary work
    Flexible hours

    Coupang

    Mountain View, CA
    1 day ago
  • $210k - $240k

     ...Staff Software Engineer Palo Alto, CA About Typeface We help the world's biggest brands move from brief to fully personalized campaigns...  ...related field ~12+ years of experience in developing scalable, reliable, performant, and secure full-stack applications. ~... 
    Work at office
    Flexible hours
    3 days per week

    Typeface

    Palo Alto, CA
    4 days ago
  •  ...Staff Software Engineer Interface.ai is building the infrastructure layer for AI-powered financial services. Our agentic AI platform enables...  ..., data models, deployment patterns, observability, and reliability targets Serve as the technical bridge between product vision... 
    Full time

    Interface AI

    Palo Alto, CA
    3 days ago
  • $206k - $258k

     ...protect it for future generations. Role Summary As a Software Engineer specializing in safety-critical self-driving middleware, you...  ...contributing to the successful implementation of robust and reliable self-driving solutions. Responsibilities Design,... 
    Full time
    Contract work
    Local area

    Rivian

    Palo Alto, CA
    4 days ago
  • $150k - $226k

     ...Harness is the AI Software Delivery Platform company, led by technologist and entrepreneur...  ..., deployments, application security, reliability, compliance, and cost optimization....  ..., and real-time risk detection to help engineering teams ensure software integrity, prevent... 
    Full time
    Local area
    Immediate start
    Flexible hours
    Shift work

    Harness

    Mountain View, CA
    1 day ago
  • $195k - $343k

     ...Saturn, and other digital services. We’re looking for a Staff Software Engineer to join Snap Inc on our Feature Store team! This tech...  ...— while ensuring they meet correctness, performance, and reliability requirements. A central part of the role is enforcing engineering... 
    Full time
    Live in
    Work at office
    Local area

    Snap Inc.

    Palo Alto, CA
    2 days ago
  • $171.15k - $288.5k

     ...your manager). The Role: General Motors is seeking a Staff Software Engineer to shape and scale the engineering platforms that build,...  ...templates, developer self-service, engineering standards, or reliability objectives. Company Vehicle: Upon successful... 
    Full time
    Local area
    Immediate start
    Remote work
    Work from home
    Flexible hours

    General Motors

    Mountain View, CA
    3 days ago
  • $195k - $343k

     ...We are looking for an L6 Staff Technical Lead to define the...  ...across AWS and GCP; and improve reliability, performance, cost...  ...Infrastructure and Product Engineering, mentor engineers, and turn...  ...experience. ~9+ years of software development experience; or a... 
    Full time
    Live in
    Work at office
    Local area

    Snap Inc.

    Palo Alto, CA
    2 days ago
  • $193.93k - $352.29k

     ...other leading investors. About the Role Our software team is growing, and we are looking for talented engineers to join us and be instrumental to one of the...  ...onboard system team's software engineers provide a reliable and high-performance platform that allows our... 
    Immediate start
    Flexible hours

    Nuro

    Mountain View, CA
    24 days ago
  • $300 per month

     ...Crusoe’s Data Center Infrastructure Engineering (DCIE) team is fundamental to our mission...  ...seeking a highly skilled and motivated Software Engineer to join Crusoe’s Data Center Infrastructure...  .... Expertise in distributed systems, reliability, and cloud platforms (Kubernetes, IaC,... 
    Temporary work

    Crusoe

    Sunnyvale, CA
    a month ago
  • $160.2k - $330.4k

     ...systems to intuitive design, intelligent software, and next-generation safety and...  ...and efficiently. As one of the founding engineers, you'll set technical direction from a blank...  ...management services that keep the platform reliable in production. The architecture you choose... 
    Full time
    Local area
    Work from home
    Relocation package
    Flexible hours

    General Motors

    Mountain View, CA
    2 days ago
  • $206.5k - $258.1k

     ...Role Summary In this position, you will be a Lead Staff Engineer developing embedded software for Rivian's next-generation autonomy driving platform...  ...to validate software functionality, safety, and reliability in compliance with automotive standards. Keep up with... 
    Full time
    Contract work
    Temporary work
    Part time
    Work experience placement
    Local area
    Shift work

    Rivian

    Palo Alto, CA
    17 hours ago
  • $235.7k - $277k

     ...Employment Type: FullTime Location Type: Remote Department Engineering Compensation: $235.7K - $277K - Offers Equity At Confluent,...  ...This is production infrastructure serving live inference, so reliability isn't an afterthought. Make the engineers around you better... 
    Full time
    Live in
    Remote work

    Confluent

    Mountain View, CA
    17 hours ago
  • Staff Software Engineer We're looking for a Staff Software Engineer to build and scale Nectar's core product experiences and platform. You'll...  ...the infrastructure that powers our AI workloads Improve reliability, performance, and cost efficiency across our backend and data... 
    Work at office
    Remote work
    Flexible hours

    XRC Ventures

    Palo Alto, CA
    17 hours ago
  • Staff Software Engineer GM is working toward a future defined by Zero Crashes, Zero Emissions and Zero Congestion. Achieving that goal requires...  ...cloud solutions such as in-plant monitoring systems, where reliable software connects cameras, plant infrastructure, machine-... 

    General Motors

    Mountain View, CA
    17 hours ago
  • $179.5k - $260k

     ...scalable, trustworthy systems. We're looking for an Applied AI Engineer with strong backend and AI experience who can architect, build...  ...non-technical stakeholders about trade-offs, performance, and reliability. 3. REQUIRED QUALIFICATIONS 3.1 Experience Proven track... 
    Full time
    Worldwide
    Flexible hours
    Night shift

    Fortinet

    Sunnyvale, CA
    17 hours ago
  • $230k - $275k

     ...candidate is a deeply technical growth engineer who combines hands-on execution with system...  ...bidding optimization engine, building a reliable creative generation pipeline, or...  ...experience building and shipping high-impact software systems. Proven experience in growth engineering... 
    Local area

    Quince

    Palo Alto, CA
    17 hours ago
  • $209.7k - $283.8k

     ...To support this growth, we need strong technical ownership to ensure our ML pipelines remain reliable, scalable, and architecturally sound. We are seeking a staff ML engineer to design and evolve the large-scale offline platform. This role focuses on building reliable... 
    Work at office
    Worldwide
    Relocation package

    Unity

    Mountain View, CA
    5 days ago
  • $190k - $230k

     ...leveraging its commercial self-driving software to develop, test and deploy autonomous capabilities...  ...for an experienced Controls Software Engineer who is passionate about safety-critical...  ...techniques Architect, develop, and test reliable, redundant, and safety-critical software... 
    Temporary work
    Work at office
    Visa sponsorship
    Flexible hours

    Kodiak

    Mountain View, CA
    17 hours ago
  • $279k - $341k

    Senior Staff Software Engineer Mountain View, US About EarnIn As one of the first pioneers of earned wage access, our passion at EarnIn is building...  ...for security, privacy, correctness, and operational reliability. The base salary range for this full-time position is $279... 
    Full time
    2 days per week

    Earnin

    Mountain View, CA
    17 hours ago
  • $143k - $286k

     ...Position Summary... What you'll do... We are seeking a Staff Software Engineer to lead the design and development of highly scalable, distributed...  .... You will drive architectural decisions, ensure system reliability and performance at scale, and build solutions that handle... 
    Full time
    Temporary work
    Part time
    Local area

    Walmart

    Sunnyvale, CA
    17 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Staff Software Engineer, Reliability. Be the first to apply!