Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Principal Staff Software Engineer, Systems Infrastructure

$226k - $369k

Linkedin

Company DescriptionLinkedIn is the world’s largest professional network, built to create economic opportunity for every member of the global workforce. Our products help people make powerful connections, discover exciting opportunities, build necessary skills, and gain valuable insights every day. We’re also committed to providing transformational opportunities for our own employees by investing in their growth. We aspire to create a culture that’s built on trust, care, inclusion, and fun – where everyone can succeed.Join us to transform the way the world works.Job DescriptionAt LinkedIn, our approach to flexible work is centered on trust and optimized for culture, connection, clarity, and the evolving needs of our business. This role may be remote or hybrid. At LinkedIn, hybrid roles are performed both from home and from a LinkedIn office on select days, as determined by the business needs of the team. Remote roles are performed from the designated home work location upon time of hire, and any changes to this home work location requires a review of remote status and approval. LinkedIn’s Reliability Infrastructure team is responsible for defining and driving the reliability strategy, standards, and practices that keep LinkedIn’s most critical systems stable, resilient, and available at massive scale.As a Principal Staff Software Engineer, Reliability Infrastructure, you will serve as a senior technical authority for reliability across LinkedIn Engineering. You will help define how critical services are designed, built, operated, and measured, partnering broadly across infrastructure and product engineering teams to improve resiliency, reduce incidents, and raise the reliability bar across the company.A key focus of this role is driving the adoption and evolution of LinkedIn’s service criticality framework, including reliability expectations for the most business-critical systems. You will help classify services based on criticality and blast radius, define appropriate reliability standards, and influence system architecture to ensure the right levels of availability, redundancy, observability, and failure handling are in place.As AI-assisted software development, agent-based automation, and autonomous operational systems become more prevalent, this role will also help define how LinkedIn safely builds and operates reliable AI-enabled systems. You will shape standards for evaluating, deploying, monitoring, and governing AI-generated code and agentic workflows, ensuring that automation introduced into critical environments is observable, explainable, auditable, and designed with appropriate safeguards, rollback mechanisms, and human oversight.This is not a traditional SRE role focused on operating a single service or team. It is a company-wide technical leadership role for someone with deep distributed systems expertise, strong reliability judgment, and the ability to influence architecture and engineering practices across large organizations.ResponsibilitiesDefine and drive company-wide reliability strategy, standards, and best practices across LinkedIn EngineeringLead adoption and evolution of service criticality models that set reliability expectations based on business impact and blast radiusServe as a technical authority for architecture decisions related to reliability, resiliency, availability, and failure handlingPartner with infrastructure and product engineering teams to improve system design, reduce incident risk, and strengthen operational readinessIdentify high-risk systems and drive cross-organizational initiatives to improve reliability of critical servicesEstablish and evolve reliability standards including SLOs, SLIs, uptime expectations, redundancy, monitoring, alerting, and failover patternsInfluence engineering culture by promoting reliability-focused design, incident review rigor, and postmortem-driven improvementsProvide architectural guidance and mentorship to senior engineers and technical leaders across teamsBalance technical strategy, hands-on engineering judgment, and cross-functional influence to drive measurable improvements in site stabilityHelp shape how LinkedIn builds and operates resilient systems as the platform continues to scaleDrive the strategy for applying LLMs to alert triage, root cause analysis, and incident summarization at scale, ensuring systems are explainable, auditable, and safe to operate autonomously in Ring0/Ring1 environments.QualificationsBasic QualificationsBA/BS degree in Computer Science or related technical field, or equivalent practical experience10+ years of experience in software engineering, infrastructure engineering, distributed systems, SRE, production engineering, or reliability engineering5+ years of experience in a technical leadership, architect, or principal-level engineering roleExperience designing, building, or operating large-scale distributed systemsExperience defining or driving reliability standards such as SLOs, SLIs, uptime targets, incident reduction, or operational readiness frameworksUnderstanding of high availability, redundancy, fault tolerance, failure modes, and resiliency patternsExperience influencing architecture and engineering practices across multiple teams or organizationsSoftware engineering experience in one or more languages such as Java, Go, C++, Python, or similarPreferred QualificationsMS or PhD in Computer Science or related technical fieldExperience operating at company-wide or large org-wide scope as a reliability, infrastructure, SRE, or production engineering technical leaderExperience with tiered service criticality models, priority-based reliability frameworks, or large-scale reliability governanceDeep expertise in distributed systems reliability, service resilience, and failure isolation at scaleExperience with incident management, postmortems, operational reviews, and driving long-term corrective actions across organizationsExperience with observability, monitoring, alerting, capacity planning, disaster recovery, and multi-region failover strategiesBackground in mature SRE, production engineering, platform reliability, or infrastructure resilience environmentsExperience driving reliability transformations across large engineering organizationsExecutive-level communication skills with the ability to align technical decisions to business impactDemonstrated ability to influence technical direction without direct authority and drive adoption of standards across teamsPrior work on self-healing or auto-remediation systems at companies with large-scale infrastructure (hyperscalers, large internet companies).Experience defining reliability, safety, or governance standards for AI-enabled systems, agentic workflows, or AI-assisted software developmentFamiliarity with LLM and agent evaluation, production monitoring, guardrails, human oversight, and rollback strategies for autonomous systemsSuggested SkillsDistributed systems reliabilitySite Reliability Engineering / Production EngineeringHigh availability and fault toleranceSLO / SLI designIncident management and postmortem practicesObservability and monitoringResiliency engineeringAI agent reliability and safetyLLM and agent evaluationAI observability and monitoringAutonomous remediation and guardrailsLarge-scale infrastructureCross-organizational technical leadershipReliability standards and governanceFamiliarity with emerging standards and frameworks for agentic AI safety and evaluation (evals pipelines, red-teaming autonomous systems, policy guardrails for production agents).LinkedIn is committed to fair and equitable compensation practices. The pay range for this role is $226,000 to $369,000. Actual compensation packages are based on several factors that are unique to each candidate, including but not limited to skill set, depth of experience, certifications, and specific work location. This may be different in other locations due to differences in the cost of labor. The total compensation package for this position may also include annual performance bonus, stock, benefits and/or other applicable incentive compensation plans. For more information, visit . Additional InformationEqual Opportunity Statement We seek candidates with a wide range of perspectives and backgrounds and we are proud to be an equal opportunity employer. LinkedIn considers qualified applicants without regard to race, color, religion, creed, gender, national origin, age, disability, veteran status, marital status, pregnancy, sex, gender expression or identity, sexual orientation, citizenship, or any other legally protected class.LinkedIn is committed to offering an inclusive and accessible experience for all job seekers, including individuals with disabilities. Our goal is to foster an inclusive and accessible workplace where everyone has the opportunity to be successful.If you need a Reasonable Accommodation to search for a job opening, apply for a position, or participate in the interview process, connect with us and describe the specific Accommodation requested for a disability-related limitation. Fill out an Accommodation request here: accommodations are modifications or adjustments to the application or hiring process that would enable you to fully participate in that process. Examples of reasonable accommodations include but are not limited to:Documents in alternate formats or read aloud to youHaving interviews in an accessible locationBeing accompanied by a service dogHaving a sign language interpreter present for the interviewA request for an accommodation will be responded to within three business days. However, non-disability related requests, such as following up on an application, will not receive a response.LinkedIn will not discharge or in any other manner discriminate against employees or applicants because they have inquired about, discussed, or disclosed their own pay or the pay of another employee or applicant. However, employees who have access to the compensation information of other employees or applicants as a part of their essential job functions cannot disclose the pay of other employees or applicants to individuals who do not otherwise have access to compensation information, unless the disclosure is (a) in response to a formal complaint or charge, (b) in furtherance of an investigation, proceeding, hearing, or action, including an investigation conducted by LinkedIn, or (c) consistent with LinkedIn's legal duty to furnish information.San Francisco Fair Chance Ordinance ​Pursuant to the San Francisco Fair Chance Ordinance, LinkedIn will consider for employment qualified applicants with arrest and conviction records.Pay Transparency Policy Statement ​As a federal contractor, LinkedIn follows the Pay Transparency and non-discrimination provisions described at this link: Data Privacy Notice and Compliance Posters for Job Candidates Please use this link to access documents that provide information about how LinkedIn handles the personal data of employees and job applicants, as well as the E-Verify Participation Notice and the Department of Justice Immigrant and Employee Rights Section Right to Work posters: Full-timeFunction: SalesExperience level: DirectorIndustry: Internet

Vacancy posted 16 hours ago
Similar jobs that could be interesting for youBased on the Principal Staff Software Engineer, Systems Infrastructure in Mountain View, CA vacancy
  • $226k - $369k

     ...status and approval. We are seeking a Principal Staff Software Engineer to join our organization. The team...  ...and driving next-generation infrastructure that powers AI-first unified developer...  ...Architect large-scale infrastructure systems for end-to-end software lifecycle platformsDesign... 
    Suggested
    For contractors
    Work at office
    Remote work
    Work from home
    Flexible hours

    Linkedin

    Mountain View, CA
    2 days ago
  • $226k - $369k

     ...As part of our world-class software engineering team, you will take the...  ...building the next-generation infrastructure and platforms for LinkedIn...  ..., API design and systems design, and your passion for...  ...impact within our company.As a Principal Staff Software Engineer, you will... 
    Suggested
    For contractors
    Work at office
    Flexible hours

    Linkedin

    Sunnyvale, CA
    1 day ago
  • $231k - $378k

     ...determined by the business needs of the team.As a Principal Staff Software Engineer of the Compute Infrastructure team at LinkedIn, you will play a crucial role in...  ...developing large-scale distributed systems and using agentic software engineeringPreferred Qualifications... 
    Suggested
    For contractors
    Work at office
    Flexible hours

    Linkedin

    Mountain View, CA
    16 hours ago
  • $231k - $378k

     ...the business needs of the team.We are seeking a Principal Staff Software Engineer to join LinkedIn’s Physical Infrastructure organization. As the technical partner to the...  ...organization builds and operates the systems that translate business and product demand into... 
    Suggested
    For contractors
    Work at office
    Immediate start
    Flexible hours

    Linkedin

    Mountain View, CA
    4 days ago
  • $207k - $340k

     ...and Machine Learning Engineers are both data/research scientists and software engineers, who develop...  ...implementation. As a Principal Staff Software Engineer you...  ..., models, and systems that power our LinkedIn...  ...strengthen LinkedIn’s AI infrastructure.QualificationsBasic QualificationsBS... 
    Suggested
    For contractors
    Work experience placement
    Work at office
    Flexible hours

    Linkedin

    Mountain View, CA
    1 day ago
  • $198k - $326k

     ...the business needs of the team. LinkedIn’s AI Infrastructure organization is responsible for building the...  ...serving frameworks.We are looking for a Senior Staff Software Engineer with deep expertise at the intersection of systems, machine learning, GPU infrastructure, and... 
    For contractors
    Work at office
    Flexible hours

    Linkedin

    Mountain View, CA
    2 days ago
  • $142.8k - $274.8k

     ...ContributorTravel: Less than 25%Profession: Software EngineeringDiscipline: Software...  ...within Microsoft’s Azure Hardware Systems and Infrastructure (AHSI) organization, the team...  ...and Xbox Live. We are seeking a Principal Software Engineer - Networking & Storage Systems to... 
    Ongoing contract
    Local area
    Worldwide
    3 days per week

    Microsoft

    Mountain View, CA
    2 days ago
  • $280k - $350k

     ...top researchers and engineers, building the world’s...  ...ambitious and capable Staff/Principal Backend Engineer to join...  .... You'll own core systems for multi-provider failover...  ...and exciting Infrastructural projects: platformization...  ...As a Staff/Principal Software Engineer, you would... 
    Full time
    Work at office
    Relocation

    Inworld AI

    Mountain View, CA
    2 days ago
  • $142.8k - $274.8k

     ...ContributorTravel: Less than 25%Profession: Software EngineeringDiscipline: Software...  ...: MicrosoftOverviewThe AI Infrastructure Engineering Systems team at Microsoft builds the cloud platform...  ...collaborative and inclusive culture.As a Principal Software Engineer - AI... 
    Ongoing contract
    Local area
    3 days per week

    Microsoft

    Mountain View, CA
    2 days ago
  • $207k - $300k

     ...coach a distributed team of engineers. Facilitate alignment and clarity...  ..., and enhance large scale software solutions.Minimum...  ...working with embedded operating systems.5 years of experience testing...  ...architecture built by the Technical Infrastructure team to keep it running.... 

    Google

    Sunnyvale, CA
    2 days ago
  • $250k - $350k

     ...persistent, multimodal AI system that predicts, simulates,...  ...AIDO Foundry, our AI-native engineering framework for automating...  ...development, we are building the software infrastructure for the next generation of...  ...seeking an experienced Staff to Principal-level Software Engineer to... 
    Full time

    GenBio AI

    Palo Alto, CA
    3 days ago
  •  ...A leading software delivery company in Mountain View, CA, is searching for an experienced Software Engineer to help build a cutting-edge developer productivity platform. The ideal...  ...Responsibilities include developing scalable systems and mentoring other engineers. The role... 

    Menlo Ventures

    Mountain View, CA
    6 hours ago
  • $272k - $431.25k

     ...the team and see how you can make a lasting impact on the world.At NVIDIA, as a Principal Rack Scale Systems Infrastructure Engineer, you will build and guide the development of software systems. These systems support our upcoming rack-scale infrastructure products and... 
    Full time
    Remote work
    Shift work

    Nvidia

    Santa Clara, CA
    2 days ago
  • $276k - $414k

     ...other digital services.Snap Engineering teams build fun and...  ...forefront.We’re looking for a Principal Software Engineer to join Snap Inc!What...  ...distributed, and multi-cloud infrastructure that powers all Snap's...  ...infrastructure, distributed systems, and operational excellence... 
    Full time
    Temporary work
    Live in
    Work at office
    Local area

    Snap

    Palo Alto, CA
    2 days ago
  • $258k - $387k

     ...leading investors.About the RoleAs a Principal Software Engineer, you will help define and build the...  ...Platform, Performance, and Onboard Systems, requiring deep technical leadership...  ...technical direction of Nuro’s onboard infrastructure. We are looking for a technical... 
    Immediate start
    Flexible hours

    Nuro

    Mountain View, CA
    5 days ago
  • $193.93k - $352.29k

     ...investors.About the RoleOur software team is growing, and we are looking for talented engineers to join us and be instrumental...  ...the following areas: Onboard Systems, Performance, and Devices Platform...  ...and tracing tools and infrastructure (perf, eBPF, Perfetto, pprof,... 
    Immediate start
    Flexible hours

    Nuro

    Mountain View, CA
    5 days ago
  • $240k - $265k

     ...is also leveraging its commercial self-driving software to develop, test and deploy autonomous capabilities...  ...of Defense.We are looking for a Senior or Staff Software Engineer to build infrastructure, tools, and systems that help our Planning team develop, debug, evaluate... 
    Visa sponsorship

    Kodiak Robotics

    Mountain View, CA
    3 days ago
  • $207k - $301k

     ...implement, and analyze computer systems and their interactions with...  ...in Warsaw.Develop junior engineers on the team.Minimum qualifications...  ...testing, and launching software products.5 years of...  ...and developing large-scale infrastructure, distributed systems or networks... 
    Worldwide

    Google

    Sunnyvale, CA
    5 days ago
  • $207k - $300k

    Lead the design and end-to-end software delivery of specialized AI compute platforms...  ....Partner closely with hardware engineering and chip design teams to influence...  ...building and developing large-scale infrastructure, distributed systems or networks, or experience with... 
    Remote work
    Worldwide

    Google

    Sunnyvale, CA
    2 days ago
  • $198k - $326k

     ...learning models at scale. Our AI systems power recommendations,...  ...frameworks that empower engineers and researchers to rigorously...  ...engineers robust, highly scalable infrastructure that delivers continuous,...  ...at scale.As a Sr. Staff Software Engineer, you will help define... 
    For contractors
    Work at office
    Flexible hours
    Shift work

    Linkedin

    Sunnyvale, CA
    16 hours ago
  • $198k - $326k

     ...model training, feature engineering and serving with...  ..., data infra, compute software, and hardware to harness...  ...queries.Model Training Infrastructure: As an engineer on the...  ...inference at scale.As a Sr. Staff Software Engineer, you...  ...of various systems with a focus on improving... 
    For contractors
    Work at office
    Flexible hours

    Linkedin

    Sunnyvale, CA
    16 hours ago
  • $193.93k - $352.29k

     ...leading investors.About the RoleOur software team is growing, and we are looking for talented engineers to join us and be instrumental...  ..., Simulation, and Technical Infrastructure.Data Platform: The Data...  ...as a comprehensive management system for Nuro AI Driver's data, labels... 
    Immediate start
    Flexible hours

    Nuro

    Mountain View, CA
    2 days ago
  • $207k - $301k

     ...team of high-performing engineers who design and model...  ...roadmap for modeling infrastructure, exercising strong technical...  ...high standards for system design, code quality,...  ...testing and launching software products and working...  ...experienced and passionate Staff Software Engineering... 
    Remote work
    Worldwide

    Google

    Sunnyvale, CA
    4 days ago
  • $193.93k - $352.29k

     ...investors.About the RoleThe Autonomy ML Infrastructure team is responsible for building &...  ...model compression. Work with autonomy engineers to optimize, validate, and deploy large...  ...framework, FTL.Write robust, high quality software to increase our confidence in our vehicle... 
    Work experience placement
    Immediate start
    Flexible hours

    Nuro

    Mountain View, CA
    5 days ago
  • $207k - $300k

     ...Influence and coach a distributed team of engineers.Facilitate alignment and clarity...  ...maintain, and enhance large-scale software solutions.Minimum qualifications:...  ...and developing large-scale infrastructure, distributed systems or networks, or experience with compute... 
    Worldwide

    Google

    Sunnyvale, CA
    2 days ago
  • $262k - $364k

     ...coach a distributed team of engineers.Facilitate alignment and clarity...  ..., and enhance large-scale software solutions.Minimum...  ...working with embedded operating systems.5 years of experience with design...  ...software solutions.The AI and Infrastructure team is redefining what’s... 
    Worldwide

    Google

    Sunnyvale, CA
    3 days ago
  • $218.8k - $335.3k

     ...edge robotics, optimization, and machine learning to build systems that are both intelligent and trustworthy. The Secondary Driving...  ...to bring the vehicle to a safe stop. We are looking for a Staff Software Engineer to provide technical leadership for the Secondary Driving... 
    Full time
    Local area
    Remote work
    Work from home
    Flexible hours

    General Motors

    Sunnyvale, CA
    5 days ago
  • $207k - $300k

     ...roadmap for Google’s accelerator software stacks, focusing on distributed systems, Linux OS, networking, power...  ...interfaces.Collaborate with HW engineering to design, test, deploy, and debug...  ...building and developing large-scale infrastructure, distributed systems or networks... 
    Worldwide

    Google

    Sunnyvale, CA
    2 days ago
  • $262k - $364k

     ...and coach a distributed team of engineers.Facilitate alignment and clarity...  ...maintain, and enhance large scale software solutions.Work on automation systems that execute complex network lifecycle...  ...in networking and systems infrastructure.Preferred qualifications:Master’s... 
    Worldwide

    Google

    Sunnyvale, CA
    2 days ago
  • $229k - $343k

     ...are looking for an L6 Staff Tech Lead to join our...  ...the critical Service Infrastructure that powers Snap's entire...  ...workloads. Our systems are widely adopted across...  ...role is for a backend engineer focused on Service & Compute...  ....9+ years of software development experience... 
    Full time
    Live in
    Work at office
    Local area

    Snap

    Palo Alto, CA
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Principal Staff Software Engineer, Systems Infrastructure. Be the first to apply!