Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Observability SRE (Site Reliability Engineer)

$115k - $135k
Full-time

Nomura Holdings, inc.

Job Title: Observability SRE (Site Reliability Engineer)

Department: Technology

Location: Philadelphia

Corporate Title: Associate

The pay range for this position at commencement of employment is expected to be between $115,000 - $135,000 per year* (see below footnote for additional compensation and benefits information).

Company overview

Nomura is a financial services group with an integrated global network. By connecting markets East & West, we service the needs of individuals, institutions, corporates and governments through our four business divisions: Wealth Management, Investment Management, Wholesale (Global Markets and Investment Banking) and Banking.

Driven by the insights of some 28,000 people worldwide, we put our clients at the center of everything we do, delivering unparalleled access to, from and within Asia. For further information about Nomura, visit

Aon’s Benefit Index®, Nomura’s benefits rank #1 amongst our competitors

Department Overview:

The Information Technology department at Nomura is at the forefront of innovation, driving technology solutions that empower our business and enhance client experiences. We leverage cutting-edge technologies to develop and maintain robust systems and infrastructure, ensuring the security, reliability, and efficiency of our operations. Join our team and be part of a dynamic and collaborative environment that embraces technological advancements to deliver value and drive our digital transformation journey.

Key Responsibilities:

SRE within the Group Platform Services & Engineering division which provides the Nomura group common services to Development, Infrastructure and Production Services. This is an SRE/support position responsible for administering and supporting Production environment as well as engineering reliability into the products / services we support i.e. monitoring & observability platform. The successful candidate will have a vital role in shaping future monitoring strategy and direction within the Nomura Group.

A fantastic opportunity for somebody with 3+ years IT experience to work with state-of-the-art technologies to deliver industry leading solutions in the Telemetry, Observability and Monitoring space. The successful candidate would join a team of enthusiastic, creative and forward-thinking SRE in the US who are working in tandem with the engineers to radically transform how the Nomura Group manages the operation of its estate. The position is within a global team consisting of 20 team members, across Engineering and SRE, bringing change across the organisation. The candidate will work closely with their peers in other regions as well as other teams to facilitate the strategic objectives of the team. The challenges we strive to solve include availability, scalability and performance related to delivering a platform used by the entire Nomura Group.

The observability platform consists of a combination of platforms and frameworks from in-house, vendors, and open source. These include:

  • Grafana LGTM stack (open source)
  • RightITNow, EverBridge, Sentinel (3rd Party tools)
  • AMBER, Bing, MCM, CMS (homegrown)
  • Cross functional engagement to champion and provide necessary support for the adoption of TOM platform across the Nomura group of companies.
  • Gain understanding of the various tools and frameworks that together provide observability and notification service to the organization and assist development and production support teams with queries / issues related to their usage of our platform.
  • Act as custodian of production environment and engage within the team and outside, if need be, towards building and maintaining robust, scalable, highly available production systems in accordance with our service level objectives
  • Preventing production incidents but when they do occur, performing effective incident and problem management and RCA to minimize downtime as well as possibility of recurrence.
  • Pushing out changes and releases to production environment reliably via effective change and release management
  • Quick and effective response to alerts before they become incidents, with an approach to prevent them from occurring ever again
  • Effectively triaging alerts, requests, emails such that things that needs attention get addressed first and in a timely manner in the order of their priority, the drivers for which should be production stability and user satisfaction
  • Continuous and effective engagement with users, with the required empathy, providing the right guidance so as to provide a good customer experience
  • Collaborate in a global agile team environment using established support practices, participating in sprint planning, reviews, and continuous improvement initiatives
  • Build and maintain scalable, reliable monitoring solutions that support Nomura's global infrastructure
  • Engage with engineers, architect towards contributing to architectural decisions that influence the future direction of Nomura's observability platform
  • Champion observability best practices across the organization, helping teams leverage data-driven insights to improve system reliability and performance
  • Partner with engineers as needed to optimize operational efficiency and enhance system resilience
  • Effectively leveraging AI tools such as Claude, CoPilot etc. with adequate guardrails to bring efficiencies into operational processes in a consistent, repeatable and risk averse manner.
  • Mentor and guide other SREs, sharing your knowledge and expertise across other team members for the benefit of the team.


Required Qualifications:

Mandatory:

  • Minimum 2 years’ experience with Grafana or any other modern observability tools in an administrative capacity for a medium/large scale enterprise.
  • At least 2 years’ exposure to Linux OS with a decent hold on general purpose troubleshooting and day to day commands
  • Exposure to one or more of following – Python / Ansible
  • Production support experience – Request handling, incident management, problem management, change management, release management, on-call handling, user engagement, responding to alerts etc.
  • Good communication and interpersonal skills
  • Strong analytical and trouble-shooting skills, with the ability to exercise mature judgement
  • Basic understanding of cloud platforms
  • Basic understanding of CI/CD tools such as GitLab, Jenkins, Ansible, Nexus etc.
  • Good Team player

Preffered:

  • Understanding of Open Telemetry standards
  • Understanding of containerization technologies such as Kubernetes, EKS, Docker etc.
  • Supporting a medium / large scale production environment
  • Knowledge of ITIL
  • Decent understanding of DB Platforms – Sybase / MySQL / MSSQL – general RDBMS concepts, SQL
  • Collaboration Tools – Confluence / JIRA
  • Basic knowledge of / familiarity with other infrastructure technologies such as Middleware (ActiveMQ / Solace / EMS / Tibco etc.), Web servers, Load balancers, Directory Services etc.
  • Experience working with a globally dispersed team

Nomura Leadership Behaviours

  • Explore Insights & Vision: Identify the underlying causes of problems faced by you or your team and define a clear vision and direction for the future.
  • Making Strategic Decisions : Evaluate all the options for resolving the problems and effectively prioritize actions or recommendations.
  • Inspire Entrepreneurship in People : Inspire team members through effective communication of ideas and motivate them to actively enhance productivity.
  • Elevate Organizational Capability : Engage proactively in professional development and enhance team productivity through the promotion of knowledge sharing.
  • Inclusion : Foster a culture of inclusion and psychological safety in the workplace and cultivate a "Risk Culture" (Challenge, Escalate and Respect).

* base pay offered may vary depending on multiple individualized factors, including market location, corporate and functional title and duties, job-related knowledge and advanced degrees, skills, and experience. The total compensation package for this position may also include other elements, including a sign-on bonus, restricted stock units, and discretionary awards in addition to a full range of medical, financial, and/or other benefits (including 401(k) eligibility and various paid time off benefits, such as vacation, sick time, and parental leave), dependent on the position offered. Details of participation in these benefit plans will be provided if an employee receives an offer of employment.

If hired in the U.S., employee will be in an “at-will position” and the Company reserves the right to modify base salary (as well as any other discretionary payment or compensation program) at any time, including for reasons related to individual performance, Company or individual department/team performance, and market factors”.

Nomura is an Equal Opportunity Employer

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Observability SRE (Site Reliability Engineer) in Philadelphia, PA vacancy
  • $1,000 per month

     ...the AWS environment, deployments, observability, and the SOC 2 and PCI-DSS...  ...building the AI infrastructure our engineers use every day. When this team does...  ...of gaps.As a Senior DevOps / SRE Engineer on this team, you'll own reliability and deployments across our AWS and... 
    Suggested
    Temporary work
    Work at office
    Immediate start
    Remote work
    Flexible hours

    Creditly Corp

    Philadelphia, PA
    1 day ago
  • $130k - $150k

     ...Site Reliability Engineer (SRE) Engineer Reliability into the Systems That Move the Nation’s Food Supply Who We Are US Cold owns and...  ...automation interfaces — and design controls, automation, and observability that reduce incidents over time. Success in this role... 
    Suggested

    USCS

    Camden, NJ
    1 day ago
  •  ...We are seeking a Lead Site Reliability Engineer (SRE) who combines deep technical expertise with strong leadership and client-facing capabilities...  ...system scalability, performance, and resilience Observability & Monitoring Implement and enhance monitoring, alerting... 
    Suggested

    The Judge Group

    Philadelphia, PA
    1 day ago
  •  ...that runs the way a financial exchange does. A successful Site Reliability Engineer at Sporttrade will be an operations-minded self-starter who...  ...teams to keep a live, regulated marketplace fast, observable, and compliant. Duties Own the daily operation of... 
    Suggested
    Temporary work
    Flexible hours

    sporttrade inc

    Camden, NJ
    3 days ago
  • $135.2k

     ...manual intervention and increase platform reliability Identify vulnerabilities and...  ...This role requires three days per week on-site. Basic Qualifications: Minimum 5 years...  ...needs such as for a disability or religious observance, please call us toll free at 1 (877) 889... 
    Suggested
    Hourly pay
    Full time
    Live in
    Work at office
    Local area
    Flexible hours
    3 days per week

    Accenture

    Philadelphia, PA
    1 day ago
  • $145k - $160k

     ...We are seeking a specialized Observability & Infrastructure Engineer to lead the monitoring, telemetry, and platform-as-code initiatives critical to our multi-region disaster recovery roadmap. You will architect and implement robust observability pipelines, ensure deep... 
    Temporary work
    Remote work
    Flexible hours

    EPAM Systems Inc

    Conshohocken, PA
    21 hours ago
  •  ...Description Forhyre is looking for engineers who can bring unique perspectives and...  ...practices while building a culture of reliability and observability Engage in and improve the end to end...  ...Serve as subject matter expert in an SRE mindset, best practices, and cloud-... 

    Forhyre

    Philadelphia, PA
    4 days ago
  • Senior Site Reliability EngineerLocation: Exton or Philadelphia, PA (Hybrid - 3 times a week in-office)Position SummaryAre you ready to start...  ...looking for you!We are looking for a Senior Site Reliability Engineer to take on the responsibility of automating cloud-based... 
    Casual work
    Work at office
    Worldwide

    Bentley Systems

    Philadelphia, PA
    3 days ago
  •  ...relating to any production issues; and guide and mentor junior-level engineers. Position is eligible for 100% remote work.REQUIREMENTS:...  ...everyday life.Please visit the benefits summary on our careers site for more details.Comcast is an equal opportunity workplace. We... 
    Full time
    Remote work

    Comcast

    Philadelphia, PA
    3 days ago
  • $155k - $287k

     ...seeking a Principal Platform Engineer to develop, operate, and...  ...secure, scalable, reliable, and high-performing platform...  ...Infrastructure Engineering, DevOps, Site Reliability Engineering (SRE), or related fields....  .... Experience with observability, monitoring, logging, and... 
    Full time
    Temporary work
    H1b
    Work at office
    Home office
    Flexible hours
    3 days per week

    NASDAQ OMX

    Philadelphia, PA
    1 day ago
  • $55k - $187k

     ...Internal Firm Services - Other Management Level Senior Associate Job Description & Summary The Opportunity As a Site Reliability Engineer - Senior Associate, you will play a pivotal role in enhancing the reliability, scalability, and performance of our... 
    Full time
    H1b

    PwC

    Philadelphia, PA
    1 day ago
  • $125k - $145k

     ...Sr DevOps Engineer US Cold owns and operates one...  ...We continue to enhance reliability and accelerate engineering...  ...by strengthening our SRE and AI practices. This...  ..., logging, and observability solutions to proactively...  ...DevOps, Cloud Engineering, Site Reliability... 
    Permanent employment
    Temporary work
    Work at office
    Flexible hours

    USCS

    Camden, NJ
    1 day ago
  •  ...teams to develop, deploy, and run secure, reliable, and scalable solutions. Ensures...  ...strong DevOps practices, while supporting observability, governance, compliance, and AI oversight...  ..., root cause analysis, and reliability engineering.Implement platform security, DevSecOps... 

    Dale Workforce Solutions

    Philadelphia, PA
    21 hours ago
  • $51 - $61 per hour

     ...demands to travel. Key Responsibilities:As a Release Train Engineer, you will be responsible for facilitating Agile Release Train...  ...have accommodation needs such as for a disability or religious observance, please call us toll free at (***) ***-**** or send us an email... 
    Hourly pay
    Live in
    Work at office
    Local area
    Immediate start
    Flexible hours
    Shift work

    Accenture

    Philadelphia, PA
    2 days ago
  • $145k - $177k

    Slalom's Platform Engineering team is looking for Senior Consultants...  ...pipelines from scratch.* Build observability, security controls, and...  ...engineering teams on DevOps, SRE, cloud-native, and platform...  ...security, networking, monitoring, reliability engineering, and governance... 
    Temporary work
    Local area

    Slalom

    Philadelphia, PA
    21 hours ago
  • This position is on-site in PhiladelphiaAbout ProsciaProscia is revolutionizing pathology...  ...in medicine.About This PositionAs a Reliability Engineer reporting to the VP, Technical...  ...-heavy workloads.Working knowledge of observability practices (logs, metrics, tracing) and... 
    Work at office
    Shift work

    Proscia

    Philadelphia, PA
    4 days ago
  • DevOps/Cloud Platform Engineer Cadent ignites seamless connections between...  ...partner closely with development, SRE, data, and security teams to deliver reliable, repeatable deployments and high-...  ...FTO) policy and company breaks 11 observed Federal holidays Summer Fridays... 
    Temporary work
    Work experience placement
    Summer work
    Immediate start
    Flexible hours

    Cadent

    Philadelphia, PA
    3 days ago
  •  ...supporting mission-critical environments through modern platform engineering and cloud-native infrastructure. This full-time opportunity...  ..., Python, Helm, Docker, Podman, GitOps, and cloud-native observability tools. This role offers the opportunity to build and evolve enterprise... 
    Full time
    Flexible hours

    Motion Recruitment

    Philadelphia, PA
    1 day ago
  • $113.1k - $154.1k

     ...runbooks, and escalation proceduresDevelop and maintain cloud engineering standards, governance processes, and onboarding...  ...hybrid connectivity, DNS, and segmentationExperience implementing observability, monitoring, alerting, backup, resiliency, and disaster recovery... 
    Full time
    Contract work
    Local area
    Flexible hours

    Armanino

    Philadelphia, PA
    2 days ago
  •  ...seeking a Senior Backend DevOps Engineer for a 100% remote position supporting...  ...-grade CI/CD pipelines, implementing observability and SRE practices, and supporting secure, compliant...  ...production issue triage and service reliability improvements. • Guide releases... 
    Full time
    Remote work

    VetsEZ

    Philadelphia, PA
    a month ago
  • $86.6k - $144.4k

    Senior Software Engineer: Clinical Data Platform · Wellsheet at ElsevierAbout Wellsheet...  ...end to end: ingestion pipelines, sync reliability, normalization. Keeping the stack healthy...  ...durable fixes.Keeping correctness and observability strong as the platform takes on new consumers... 
    Full time
    Local area

    Elsevier

    Philadelphia, PA
    1 day ago
  •  ...Development team is a software engineering group within Strategic Global Reliability focused on designing, building,...  ...streamline workflows, improve system observability, and empower teams to...  ..., including software engineers, site reliability engineers, product teams... 

    Susquehanna International Group

    Bala Cynwyd, PA
    4 days ago
  •  ...casual atmosphere. FreedomPay is seeking a Platform Operations Engineer to help the Platform Operations team maintain operational...  ...Required Technical Skills Proficiency with an enterprise APM or observability platform (e.g., Dynatrace, Datadog, New Relic, or comparable)... 
    Full time
    Casual work
    Flexible hours

    FreedomPay

    Philadelphia, PA
    4 days ago
  •  ...about backlog prioritizationPractical knowledge of software engineering best practices, including Agile development methodologies, DevOps...  ...debugging complex distributed systems using cloud-native observability tooling, experience with cloud-native architecture design a plusExperience... 
    Apprenticeship
    Easy work

    McKinsey & Company

    Philadelphia, PA
    21 hours ago
  • $93.85k - $143.7k

     ...is seeking an AWS Serverless Software Engineer to join the Product Engineering Unit as...  ...operational processes to improve system reliability and efficiency.Monitor application health...  ..., and authorization.Experience with observability and monitoring tools such as CloudWatch... 
    Work experience placement
    Work at office
    Local area
    Remote work

    National Board of Medical Examiners

    Philadelphia, PA
    1 day ago
  •  ...efficiency, query governance, and system reliability. Responsibilities include diagnosing...  ...integration risks early, and supports observability standards to improve system...  ...cross-functional collaboration across engineering, data, and product teams to deliver reliable... 
    Contract work
    Work at office

    Morgan Properties

    Conshohocken, PA
    3 days ago
  •  ...driven and highly collaborative, bringing together researchers, engineers, and traders to design and deploy impactful strategies in our...  ...productionImprove production behavior of the platform — observability, debugging tooling, and performance tuning under loadWork across... 

    Susquehanna International Group

    Bala Cynwyd, PA
    4 days ago
  •  ...Health.About the RoleSignant Health is looking for a Software Engineer - AI Accelerated Development to build AI-powered, agentic...  ...outputs in trusted data.Instrument agents for evaluation and observability so their outputs can be verified, measured, and audited.Apply... 

    Signant Health

    Blue Bell, PA
    21 hours ago
  • $293.9k - $406.8k

     ...outcomes, as a Distinguished Engineer. The team delivers secure,...  ..., with a strong emphasis on reliability, interoperability, and long-...  ...across platform services such as observability, data platforms, AI/ML...  ...Please see the Cisco careers site to discover more benefits and... 
    Full time
    Temporary work
    Local area
    Remote work
    Flexible hours

    CISCO Systems

    Philadelphia, PA
    2 days ago
  •  ...medicine.About the RoleAs a Senior Software Engineer, you’ll help build and evolve the...  ...operations teams to deliver highly scalable, reliable, and maintainable solutions.Our...  ...drive best practices for scalability, observability, maintainability, and performanceImprove... 
    Work at office
    Shift work

    Proscia

    Philadelphia, PA
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Observability SRE (Site Reliability Engineer). Be the first to apply!