Principal Site Reliability Engineer, Infrastructure Observability
$159k - $272kT. Rowe Price
At T. Rowe Price, we identify and actively invest in opportunities to help people thrive in an evolving world. As a premier global asset management organization with more than 85 years of experience, we provide investment solutions and a broad range of equity, fixed income, and multi-asset capabilities to individuals, advisors, institutions, and retirement plan sponsors. We take an active, independent approach to investing, offering our dynamic perspective and meaningful partnership so our clients can feel more confident. We believe doing the right thing for our clients and our associates is good business. With a career at the firm, you can expect opportunities to create real impact at work and in your community. You’ll enjoy resources to support your career path, as well as compensation, benefits, and flexibility to enrich your life. Here, you’ll find a collaborative culture that respects and values differences and colleagues who share a spirit of generosity. Join us for the opportunity to grow and make a difference in ways that matter to you. Role SummaryIn this role as Principal Site Reliability Engineer, Infrastructure Observability you will help formulate, develop, and implement a team of Site Reliability Engineers (SREs) focused on the observability, sustainability, scalability, measurability and recoverability of T. Rowe Price’s innovative cloud & on-prem solutions by leveraging automation and best-of-breed tools. The successful candidate will have a strong operations & engineering background, is hands-on when needed, and has expertise in the cloud environments (public, private), infrastructure operations, DevOps practices, CI/CD toolchain and systems, code build and deployment, incident response, and 24x7 monitoring and support.The candidate will also have extensive experience operating within a SRE function within a complex, distributed environment. They will have a demonstrated ability to work horizontally and vertically within an organization with diverse partners and sponsor groups.ResponsibilitiesPossesses extensive knowledge in own area of expertise and extensive in-depth knowledge of the broader portfolio for comprehensive understanding of up/downstream impacts across technology infrastructureResponsibility for the design of technology solutions to prevent or minimize service disruptionsPrevents technology service disruptions through technology solution recommendations and automationsFosters a culture of deep learning through blameless post-mortems to improve the shared goal of reliability across servicesTransform operations teams by facilitating internal change to adopt SRE standard methodologies across the organization and driving strategic growth in this area within Global TechnologyAnalyzes incidents impacting technology availability for high-level trends across the broad portfolioDrive initiatives to reduce or prevent technology failures in a complex, distributed technology environmentPulls together information from disconnected systems into cohesive views of the technology portfolio for identifying trends, redundancies, and riskDemonstrates outstanding awareness of the complexities of the tech and asset management industriesMay lead initiatives of varying degrees of complexity that span multi-functional areas and of varying degrees of complexityContributes to definition of target state architecture and design of the technology environmentQualificationsRequired:Bachelor's degree or the equivalent combination of education and relevant experience AND 10+ years of experience designing and operating cloud infrastructure with senior‑level impact.5+ years building and supporting solutions in Amazon AWS5+ years of experience building and running a DevOps and/or SRE functionExperience with implementation and operation of the chaos model at scaleStrategic and program-level implementation experienceDemonstrable experience implementing new technology, tools, and platformsSystem administration and scripting experienceDemonstrable experience leveraging automation to proactively prevent or quickly remediate incidentsFluent in multiple programming languages (e.g., Python, Java, GO, Node.js, .Net Core, etc)Proficiency with database development (SQL Server, PostgreSQL, MySQL, etc)Proficiency with defining, right-sizing, tracking, and reporting on Service Level Objectives (SLOs), Service Level Indicators (SLIs), system availability, and the progress and outcomes related to reliabilityExperience with implementing and managing Error BudgetsProficiency with understanding and explaining incident situations and their recovery plans to prevent recurrenceKnowledge/experience driving dashboard standardization across the ecosystem for observability, APM and infrastructure monitoring, and application-specific loggingKnowledge/experience with observability tools such as New Relic, SolarWinds DPA, Elastic Stack, Prometheus, Grafana, Splunk, and cloud native toolsKnowledge/experience with cloud management tools such as Ansible, Terraform, Vault, and VagrantWorks independently, with guidance in only the most complex situationsMakes sound decisions with limited facts or resourcesBalances strategic and pragmatic concerns when solving problemsAdjusts communication style and materials to suit a given audienceAble to clearly articulate operational principles, practices, and policiesStays abreast of industry trends and technologiesAccountable for work of self and others; sets standards around which others will operateMaintains a broad internal professional network and knows when to engage/activate itDevelops or mentor’s diverse talent on the teamAbility to be on-call and/or work during off-hoursPreferred:Cloud or SRE‑related certificationsWorking knowledge of AzureApplicants for employment in the US must have work authorization that does not now or in the future require sponsorship of a visa for employment authorization in the United States (e.g., H1-B visa, F-1 visa (OPT), TN visa or any other non-immigrant work status) FINRA Requirements FINRA licenses are not required and will not be supported for this role. Work Flexibility This role is eligible for hybrid work, with up to three days per week from home. Base Salary RangesPlease review the job posting for the location of this specific opportunity.$159,000.00 - $272,000.00 for the location of: Maryland, Colorado, Washington and remote workers$175,000.00 - $299,000.00 for the location of: Washington, D.C.$199,000.00 - $339,000.00 for the location of: New York, CaliforniaPlacement within the range provided above is based on the individual’s relevant experience and skills for the role. Base salary is only one component of our total compensation package. Employees may be eligible for a discretionary bonus, which is determined upon company and individual performance.Commitment to Diversity, Equity, and InclusionAt T. Rowe Price, our associates are our greatest asset. We thrive because our company culture is built on inclusion and because we sustain a work environment where associates can bring their best selves to work every day. The backgrounds, talents, and experiences of our global associates allow us to embrace new ideas and perspectives that move our business priorities forward and enable us to deliver strong client outcomes. Here, you can expect equal opportunity and fair and consistent treatment for all. BenefitsWe value your goals and needs, at work and in life. As an associate, you’ll be supported with resources, benefits, and work-life balance so you can thrive in ways that matter to you. Featured employee benefits to enrich your life: Competitive compensation Annual bonus eligibility A generous retirement plan Hybrid work schedule Health and wellness benefits, including online therapy Paid time off for vacation, illness, medical appointments, and volunteering days Family care resources, including fertility and adoption benefits Learn more about our benefits. T. Rowe Price is an equal opportunity employer and values diversity of thought, gender, and race. We believe our continued success depends upon the equal treatment of all associates and applicants for employment without discrimination on the basis of race, religion, creed, color, national origin, sex, gender, age, mental or physical disability, marital status, sexual orientation, gender identity or expression, citizenship status, military or veteran status, pregnancy, or any other classification protected by country, federal, state, or local law.SummaryLocation: Owings Mills, MDType: Full time
$159k - $272k
...Role SummaryCloud Reliability operates as a... ...and reliability engineering, delivering secure... ...the firm. The Principal Cloud Reliability... ...on reliability, observability, automation, and... ...design authority, site reliability... ...Build and maintain infrastructure as code using Terraform...PrincipalFull timeLocal areaRemote work3 days per week- ...understanding of how reliability, availability, recoverability... ..., roadmaps, and engineering priorities.* Makes... ...and operating cloud infrastructure with senior-level impact... ..., resilience, observability, recovery, incident response... ....* Cloud or site reliability engineering...Principal
$159k - $272k
...SummaryCloud Storage Platform Engineering provides secure, scalable,... ...The Cloud Storage Platform Principal is a senior individual... ...architecture, engineering, and reliability outcomes for enterprise... ...automation while partnering with infrastructure, cloud engineering and...PrincipalFull timeLocal areaRemote work3 days per week$110k - $188k
...Lead Workplace Experience Engineer - Power Platform is responsible... ...platform and AI service reliability, regulatory alignment, and... ...consultants, and represents Infrastructure Operations in architecture,... ...monitoring, telemetry, and observability strategy for Power Platform...SuggestedFull timeLocal areaRemote workWork from home3 days per week$121k - $206k
...you. About the TeamThe AI Engineering and Application Development... ...This team delivers secure, reliable, and forward-looking engineering... ..., build, and implement infrastructure and software solutions for... ...software, tools, and related observability) is sufficiently robust,...SuggestedFull timeWork experience placementLocal areaRemote work3 days per week- ...T. Rowe Price is seeking a Principal Cloud Reliability Engineer to lead enterprise cloud foundations with... ...guiding teams across applications and infrastructure to deliver scalable cloud... ...and CloudFormation, and advancing observability, incident response, and platform...
$121k - $206k
...the opportunity to grow and make a difference in ways that matter to you. Role Summary We are seeking a hands-on Senior Software Engineer with deep expertise in Oracle Cloud technologies, including ERP, EPM, Oracle Integration Cloud (OIC), OTBI, and APEX. This role is...Full timeLocal areaRemote work3 days per week$170k - $220k
Job DescriptionDewberry is expanding its Energy Market Sector and seeking a Physical Electrical Engineer to lead the growth of our medium and high voltage substation design practice. This is a unique opportunity to build and mentor a high-performing team while delivering...Principal$122k - $209k
...automation, and cloud-native security controls. Working across infrastructure, security, and application teams, you will establish the... ...enterprise standards, assess and mitigate risk, guide critical engineering decisions, and mentor technical talent across the...Full timeLocal areaRemote workWork from home3 days per week- ...Computer Science, Mathematics, Engineering or a related field.Masters... ...must be willing to work on-site in Woodlawn, MD 5 days a week... ...for performance and reliability.Demonstrate a strong understanding... ...using existing IBM DataPower infrastructure. Ensure interoperability with...PrincipalTemporary work
- ...Description Apply now: DevOps Engineer, location is Owings Mills,... ...the cloud-native infrastructure that powers quantitative research... .... Implement monitoring, observability, and alerting using Prometheus... ...teams to improve platform reliability, scalability, and developer...Contract workWork at officeImmediate start2 days per week
$110.6k - $178k
...The Job: As a Senior Manager, Customer Identity & Platform Engineering, you’ll be part of our IT - Customer Engagement team... ...best practices for DevSecOps, CI/CD, automated testing, observability, reliability, and platform performance.Partner with Product Management...Full timeH1bLocal areaRemote work$145k - $247k
...leader to shape how AI-enabled software engineering evolves across our mobile organization... ...processes Improve software quality, observability, telemetry, and release confidence through... ...experiences, including performance, reliability, usability, telemetry, and release...Full timeLocal areaRemote work3 days per week- ...Computer Science, Mathematics, Engineering or a related field.Masters... ...must be willing to work on-site in Woodlawn, MD 5 days a week... ...for performance and reliability.Demonstrate a strong understanding... ...using existing IBM DataPower infrastructure. Ensure interoperability with...PrincipalTemporary work
- ...T. Rowe Price is seeking a Cloud Storage Platform Principal to own the architecture, automation, and operational direction for enterprise cloud storage. You will work across NetApp, Nasuni, and AWS storage services to enable CloudNext and data center exit while mentoring...
$115 per hour
...Principal Data Product Manager (PBM) Location: Remote (Candidates must reside in approved states) Type: Contract-to-Hire... ...ecosystem projects. Partner with business leaders, architects, engineers, and analysts to deliver high-value data capabilities. Manage...PrincipalContract workTemporary workRemote work$84.84k - $153.55k
...directed by management and senior staff. This position will provide Software solutions delivery support and mentoring for Software Engineers. Lead the design and implementation of software solutions that meet business requirements and technical specifications....Full timeTemporary workWork at officeVisa sponsorshipWork visaFlexible hours- ...T. Rowe Price Hong Kong is seeking a senior cloud/site reliability engineer to design, build, and operate enterprise-scale cloud platforms. The role emphasizes reliability, security, and performance across AWS with knowledge of Azure and related tooling. Experience...
$105.5k - $168.8k
...of us—from design and engineering to the manufacturing... ...implementing monitoring, observability, systems management,... ...optimal performance, reliability, and maintainability... ...) Experience with Infrastructure as code (e.g.... ...learn more on our career site under "Our Commitment...Hourly pay- ...Our client, a leader in cloud infrastructure and enterprise solutions, is seeking a Cloud Engineer III to join their team. As a Cloud Engineer III, you will be part of the Cloud Operations Department supporting cross-functional teams. The ideal candidate will demonstrate...Weekly payTemporary workFlexible hours
- ...Senior AI Software Engineer at T. Rowe Price will design, build, and scale production-grade AI agents within a Salesforce-centric ecosystem. You will lead technical workstreams, incubate AI products, and partner with tech teams to enable broad adoption at scale. This...
$130k - $140k
...Full-time Description Software Systems Engineer Position Summary Trust Consulting Services, Inc. is seeking a Software... ...engineering changes, and corrective actions that strengthen software reliability, maintainability, performance, and product quality Perform...Full timeTemporary workLocal areaRemote workMonday to Friday$121k - $206k
...ways that matter to you. Role SummaryAs a Senior AI Software Engineer, you will design, build, and scale production-grade AI agents... ...technical debt and drive ongoing improvements in AI platforms and infrastructure.Proactively seek opportunities to apply Agentic AI...Full timeLocal areaRemote work3 days per week$121k - $206k
...investment operations and portfolio of investment applications through key long-term technology initiatives. As a Senior Software Engineer, you will play a critical role in designing and delivering next-generation, cloud-native applications that solve complex business,...Full timeLocal areaRemote work3 days per week$134k - $171k
...the existing IBM DataPower infrastructure. Ensure interoperability... ...Computer Science, Mathematics, Engineering, or a related field [... ...~ Willingness to work on-site in Woodlawn, MD, 5 days a week... ...clusters for performance and reliability. Strong understanding of...PrincipalFull time$143k - $156k
...Shift5 Shift5 is the observability platform for onboard... ..., and other critical infrastructure. Come join us.... ...is seeking a Systems Engineer to join our team. In... ...to support customer site integrations and testing... ...aerospace, defense, or high-reliability electronics sectors....Contract workFor contractorsRemote workFlexible hours$159k - $272k
...that matter to you. Role SummaryThe Principal Desktop Engineer is a hands-on engineering leader responsible... ...macOS platforms, virtual desktop infrastructure (VDI), and modern endpoint... ...OS lifecycle compliance, deployment reliability, automation maturity (including zero...PrincipalFull timeContract workLocal areaRemote workWork from home3 days per week- ...Kinzo Staffing is seeking a Splunk Enterprise Security Engineer who can develop custom detection content (correlation rules) identify... ...access. Design, manage, and maintain enterprise SIEM infrastructure to improve data ingestion processes, including architectural...Remote workNight shift
- Job-ID27127872Reference26-01207Information Technology - Engineer, Software Sr PURPOSE: Performs complex analysis, design, development... ...applications. Works with cross functional teams to develop highly reliable software that runs at scale. Provides recommendations to infuse...Work experience placement
- ...innovation, ERP and CRM counselling, Product Engineering, Business Intelligence, Data Management... ..., SharePoint Consulting and IT Infrastructure. Our other offerings include modified solutions... ...11). We are a project-driven firm that reliably meets the IT needs of our State and...Worldwide
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Principal Site Reliability Engineer, Infrastructure Observability. Be the first to apply!
- chief engineer Owings Mills, MD
- principal developer Owings Mills, MD
- general engineer Owings Mills, MD
- engineering director Owings Mills, MD
- hotel chief engineer Owings Mills, MD
- principal engineer Owings Mills, MD
- data center chief engineer Owings Mills, MD
- infrastructure engineer Owings Mills, MD
- infrastructure developer Owings Mills, MD
- senior principal cloud computing engineer Owings Mills, MD



