Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Director, Site Reliability Engineering

$189.59k - $220k

NBCUniversal Inc

Director, Site Reliability Engineering

NBCUniversal is one of the world’s leading media and entertainment companies. We create world‑class content, distribute across film, television, streaming, and bring to life through global theme parks and consumer products.

Job Description
  • As a member of NBCUniversal’s Production Software Engineering team, responsible for leading and performing custom architectural design, implementation, monitoring, and maintenance for a portfolio of production application environments.
  • Responsible for hands‑on configuration and support as well as managing the work of other architects and engineers.
  • Work closely with our Principal Software Engineer on technical architecture and design based on customer product requirements, translating product requirements to technical designs and implementations.
  • Collaborate with cross‑functional team members such as Scrum Leads, Software Engineers, QA Engineers, UX Designers, Product Managers, other Architects & Site Reliability Engineers (Contractors and/or Staff), and third‑party vendors.
  • Effectively delegate responsibilities to team members, mentoring and providing them with repeatable processes, and verifying the quality of their work.
  • Utilize metrics to measure accomplishments and monitors progress, ensuring milestones and projects are completed on‑time.
  • Communicate progress and the impact of solutions in technical terms to technology partners and in business terms to business partners.
  • Establish a reputation as the subject matter expert for every tech stack used in Production Software Engineering applications and how they all fit together while keeping current with new technologies, developing innovative technical ideas, and generating proposals.
  • Work with product teams to learn business objectives, development teams to plan platform needs, QA to understand test strategy, and SRE on environments and deployments.
  • Participate in Scrums, demos, and other Agile ceremonies and ensure accurate and timely status updates to the team.
  • Serve as primary interface with the NBCU Cyber Security team for all security‑related initiatives, patching, remediations, etc.
  • Hands‑on commissioning, configuration, administration, documentation, and support for all on‑prem & cloud (AWS) environments (Servers, Storage, Databases, Networking, Security, etc.).
  • Technical impact analysis, implementation, and monitoring of all cyber, technology audit, enterprise engineering, & IT (Databases, Monitoring, etc.) activities related to Production Software Engineering applications and platforms.
  • Create and manage CI/CD pipelines using tool likes Cloud Formation, Foreman, Jenkins, Nexus, Rundeck, Ansible, and Puppet.
  • Lead implementation of monitoring and reporting framework using tools like Grafana, Influx, Graylog/Splunk, Selenium, New Relic, and Icinga.
  • Recognize and identify potential technical impacts of enterprise change controls which could affect our applications and customers.
  • Help improve performance, scalability, and reliability.
  • Build and maintain distributed infrastructure and automation.
  • Solve problems quickly and automates processes for the future.
  • Direct management of other engineers and architects (Contractors and/or Staff). 24x7x365 availability for production outages, emergencies, and deployments.
  • 100% telecommuting is permitted for this role.
Qualifications
  • Bachelor’s degree in Computer Science, Information Technology, or related field (or foreign degree equivalent), plus 10 years of experience as a Software Architect or in a related occupation.
  • Hands‑on systems engineering experience on Linux/Unix platforms.
  • Experience with technical leadership and people management.
  • Experience with Continuous Delivery and SDLC practices.
  • DevOps principles, experience with operational tools (Ansible, Puppet, Chef, Terraform) and best practices for infrastructure (on‑prem or cloud) and software deployment.
  • Operational experience with large‑scale applications.
  • Experience with NoSQL data stores (MarkLogic, MongoDB, Cassandra, DynamoDB, Couchbase, PostgreSQL, etc.).
  • Experience with a broad range of enterprise technologies.
  • Experience building real‑time, large‑scale, low‑latency distributed systems.
  • Experience with Agile tools like Jira, GitHub or similar.
  • Additional skills (8 years experience required):
    • Experience using AWS Cloud in a production environment.
    • Experience with AWS IAM, EC2, RDS, S3, Lambda, batch and step functions.
Benefits

This position is eligible for company‑sponsored benefits, including medical, dental and vision insurance, 401(k), paid leave, tuition reimbursement, and a variety of other discounts and perks. Learn more about the benefits offered by NBCUniversal by visiting the Benefits page of the Careers website.

Compensation

Salary range: $189,592 – $220,000 per year

Full‑time: 40 hours/week

Equal Opportunity Statement

NBCUniversal’s policy is to provide equal employment opportunities to all applicants and employees without regard to race, color, religion, creed, gender, gender identity or expression, age, national origin or ancestry, citizenship, disability, sexual orientation, marital status, pregnancy, veteran status, membership in the uniformed services, genetic information, or any other basis protected by applicable law.

If you are a qualified individual with a disability or a disabled veteran, you have the right to request a reasonable accommodation if you are unable or limited in your ability to use or access nbcunicareers.com as a result of your disability. You can request reasonable accommodations by emailing View email address on click.appcast.io.

For LA County and City Residents Only: NBCUniversal will consider for employment qualified applicants with criminal histories, or arrest or conviction records, in a manner consistent with relevant legal requirements, including the City of Los Angeles' Fair Chance Initiative For Hiring Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, where applicable.

#J-18808-Ljbffr
Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the Director, Site Reliability Engineering in New York, NY vacancy
  • $120k - $142k

     ...on making journalism so good that it’s worth paying for. Mission Overview & Responsibilities: At The New York Times, our Site Reliability Engineering (SRE) team is central to how we design, test, and operate the systems that support our most critical customer... 
    Suggested
    Local area
    Flexible hours

    The New York Times

    New York, NY
    1 day ago
  • $151k - $297k

    As a TPM for SRE, you will partner with SRE leaders and engineers to scale the platform that underpins all of MongoDB’s cloud products. You will drive program execution, strengthen production reliability practices, and coordinate cross-functional efforts across US and EMEA... 
    Suggested
    Local area
    Remote work
    Worldwide
    Flexible hours

    MongoDB

    New York, NY
    2 days ago
  • $167.7k - $245.2k

     ...very effective.We’re looking for talented engineers with a software or operations background...  ...development teams to ensure the reliability, performance and security of our infrastructure...  ...insurance. Please see the Cisco careers site to discover more benefits and perks.... 
    Suggested
    Full time
    Temporary work
    Work at office
    Local area
    Flexible hours
    1 day per week

    CISCO Systems

    New York, NY
    4 days ago
  • $141k - $216.6k

     ...—it means helping shape the future of emergency response and building a safer, more connected world.Position OverviewAs a Site Reliability Engineer, you'll own the reliability, observability, and operational excellence of our Unified Call (UC) platform—the mission-critical... 
    Suggested
    Work experience placement
    Work at office

    Axon

    New York, NY
    3 days ago
  • $158.5k - $172k

     ...exceptional value they deserve.About The OpportunityAs a Senior Engineer on the Runtime Automation team, you will design, automate, and...  .... This is a high-impact position driving continuous reliability, deep system optimization, and automation across our entire technology... 
    Suggested
    Full time
    Temporary work
    Work at office
    Flexible hours
    3 days per week

    GrubHub

    New York, NY
    18 hours ago
  • $200k - $250k

    Hudson River Trading (HRT) is seeking a Senior Site Reliability Engineer to join our growing Enterprise SRE team. This team is responsible for developing and maintaining productivity service infrastructure for the entire firm, both on-prem and in the cloud. They ensure... 
    Work at office
    Local area
    Immediate start

    Hudson River Trading

    New York, NY
    13 hours ago
  • $139k - $257.55k

    The ChallengeThe Adobe Creative Community CCM organization is seeking an outstanding Senior Site Reliability Engineer (SRE) to support innovation through machine learning, autonomous AI workflows, and cloud-native infrastructure. Adobe Stock gives designers and businesses... 
    Full time
    Temporary work
    Local area
    Remote work
    Worldwide

    Adobe Systems

    New York, NY
    2 days ago
  •  ...and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems.As a Site Reliability Engineer III at JPMorgan Chase within the Commercial & Investment Bank, Production management team, you will solve complex and broad... 
    Shift work

    JP Morgan Chase

    New York, NY
    1 day ago
  • $120k - $200k

     ...PermContact: Kunal DaveContact Email: ****@*****.*** Reliability Engineer(SRE) ResponsibilitiesGlobal Architecture & Disaster Recovery...  ...practices (e.g., Chaos Engineering, resilience testing, automated recovery)SkillsBilingual Mandarin Site Reliability Engineer(SRE)
    Overseas

    Comrise

    New York, NY
    2 days ago
  • $110k - $120k

     ...largest companies to small and mid-market firms, rely on SS&C for expertise, scale, and technology.Job DescriptionJob Title: Site Reliability Engineer (SRE) / L3 Support EngineerGetting to know us:As a leading financial services and healthcare technology company based on... 
    Ongoing contract
    Full time
    Casual work
    Remote work
    Flexible hours

    SS&C Technologies

    New York, NY
    2 days ago
  •  ...customers. Cohere is a team of researchers, engineers, designers, and more, who are all...  ...building high-performance, scalable and reliable machine learning systems? Do you want to...  ...advanced NLP applications? We are looking for a Site Reliability Engineer to join the Model Serving... 
    Full time
    Work experience placement
    Work at office
    Local area
    Remote work
    Home office

    Cohere

    New York, NY
    18 hours ago
  •  ...globally recognized firm and have a direct and significant effect in a realm tailored for top achievers in site reliability. As a Lead Site Reliability Engineer at JPMorgan Chase within the Commercial & Investment Bank, Production Management team, you hold a leadership... 

    JP Morgan Chase

    New York, NY
    3 days ago
  • $131k - $164k

    Position OverviewWe are seeking a highly skilled Staff Site Reliability Engineer with deep technical expertise across VMware, Linux, and automation frameworks, to join our global Infrastructure & Operations team. This role is a hands-on senior engineering position responsible... 
    Work at office
    Local area
    Visa sponsorship
    Flexible hours

    Diligent

    New York, NY
    18 hours ago
  • $194k - $267k

     ...all in on this mission. If you are too, let's talk.Position Overview:We are seeking a highly technical StaffObservabilitySite Reliability Engineer with a specialty in Splunk to own and evolve our Splunk ecosystem. In this role, you will move beyond simple monitoring to... 
    Permanent employment
    Work at office
    Local area
    Worldwide
    Flexible hours

    Okta

    New York, NY
    4 days ago
  • $150k - $250k

    What We DoAt Goldman Sachs, our Engineers don't just make things - we make things possible. Change the world by connecting people...  ...markets.Within the firm's Global Banking & Markets business, the Site Reliability Engineering (SRE) team ensures the availability, resilience,... 
    Full time
    Temporary work
    Part time

    Goldman Sachs

    New York, NY
    4 days ago
  • $182k - $250.8k

     ...Team at Okta is the backbone of our platform's reliability and operational excellence. We are a forward-thinking group of engineers and leaders who believe that great...  ...for millions of users worldwide. As a Manager, Site Reliability Engineer, you'll lead this team with... 
    Permanent employment
    Local area
    Remote work
    Worldwide
    Flexible hours
    Weekend work
    Weekday work

    Okta

    New York, NY
    13 hours ago
  • $194k - $267k

     ...career-defining work. We're all in on this mission. If you are too, let's talk.The TeamWe are looking for an experienced Staff Site Reliability Engineer to join Okta's Emerging Products Group (EPG). Our mission is to build highly reliable, scalable, and secure cloud services... 
    Local area
    Worldwide
    Flexible hours

    Okta

    New York, NY
    3 days ago
  • $194k - $267k

     ...do something more than once, automate it” and who can rapidly self-educate on new concepts and tools.Position Overview:The Site Reliability Engineer (SRE) will play a key role in building and managing Kubernetes platforms that support cloud-native applications and... 
    Permanent employment
    Work at office
    Local area
    Worldwide
    Flexible hours

    Okta

    New York, NY
    13 hours ago
  • $141.8k - $195k

     ...work, grow fast, and bring their full selves to the herd. Why You'll Love This Role Cribl Inc is seeking a Senior Site Reliability Engineer to join our mission where you will unlock the value of all observability data, as we expand our team in the U.S. Cribl provides... 
    Temporary work
    Remote work

    Jobiqo

    New York, NY
    2 days ago
  • $130k - $250k

    What We DoAt Goldman Sachs, our Engineers don't just make things - we make things possible. Change the world by connecting people...  ...digital possibility? Begin your journey here.Securities Frontline Site Reliability Engineers (SREs) play a critical role in our fast-paced... 
    Full time
    Temporary work
    Part time
    Immediate start

    Goldman Sachs

    New York, NY
    4 days ago
  • $150k - $190k

    Senior Site Reliability Engineer, VPAt Morgan Stanley, we advise, originate, trade, manage and distribute capital for governments, institutions and individuals, and always do so with a standard of excellence. We are a leading global financial services firm that conducts... 
    Temporary work
    Worldwide
    Flexible hours
    Weekend work

    Morgan Stanley

    New York, NY
    18 hours ago
  • $195k - $275k

     ...Management, Capital Markets Application & Data Services, Deployment Planning & Release Management, and the Chief Operating Office.The Reliability Operations (RO) within WMT is responsible for providing swift, courteous, and knowledgeable customer service to end users of the... 
    Temporary work
    Work at office
    Worldwide
    Night shift

    Morgan Stanley

    New York, NY
    3 days ago
  • $150k - $160k

    Front-End & AdTech Site Reliability Engineer (SRE)Haymarket Media, Inc. is seeking a Front-End & AdTech Site Reliability Engineer (SRE) to join the Engineering team. This position is located in our New York, NY office; three (3) days in office depending on business needs... 
    Work at office
    Local area

    Haymarket Media Group

    New York, NY
    18 hours ago
  •  ...Site Reliability Engineer (SRE) Job Title Site Reliability Engineer (SRE) Job Summary We are seeking a skilled Site Reliability Engineer (SRE) to build, automate, and maintain highly available, scalable, and reliable infrastructure and applications. The ideal candidate... 
    Flexible hours

    Ova Technologies

    New York, NY
    1 day ago
  •  ...self-healing, deployment/rollback automation). Establish reliability standards: SLOs/SLIs, error budgets, production readiness reviews...  ..., and release risk controls. Performance and reliability engineering: capacity planning, load/performance analysis, resilience... 

    Bahwan CyberTek

    New York, NY
    2 days ago
  • $150k - $175k

     ...Site Reliability Engineer At ASAPP, our mission is simple: deliver the best AI-powered customer experience—faster than anyone else. To achieve that, we're guided by principles that shape how we think, build, and execute. We value customer obsession, purposeful speed... 
    Remote work

    ASAPP

    New York, NY
    5 days ago
  •  ...DevOps Engineer Responsible for reliability and support of container platform on-prem and external clouds (Azure /AWS /Google) Monitor and troubleshoot container platform environment performance issues, connectivity issues, security issues, etc. Perform deep dives... 

    Samprasoft

    New York, NY
    5 days ago
  •  ...Senior Site Reliability Engineer (SRE) Our client is seeking a Senior Site Reliability Engineer (SRE) with 10–15 years of experience to support front-office trading systems in a production environment. This role focuses on troubleshooting complex trading infrastructure... 

    TheStaffed

    New York, NY
    5 days ago
  •  ...Chariot Engineering Hire Chariot's engineering hire will be responsible for taking the Chariot platform to the next level. You will lead the evolution of our banking product, DAF product and technology strategy, along with building a growing team of software engineers... 
    Work experience placement
    Work at office
    Work from home
    Monday to Friday
    Monday to Thursday

    Chariot

    New York, NY
    2 days ago
  • $150k - $170k

     ...Senior Site Reliability Engineer – Zip Co Join to apply for the Senior Site Reliability Engineer role at Zip Co At Zip, we build cloud‑native software applications that serve millions of customers and process billions of dollars in payments. We’re looking for... 
    Casual work
    Work at office
    Remote work
    Flexible hours

    ZIP

    New York, NY
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Director, Site Reliability Engineering. Be the first to apply!