Site Reliability Engineer
Morgan Stanley
In the Technology division, we leverage innovation to build the connections and capabilities that power our Firm, enabling our clients and colleagues to redefine markets and shape the future of our communities.
This is a Lead Software Production Management & Reliability Engineering position at Director level which is part of the job family responsible for overseeing the production environment, ensuring the operational reliability of deployed software, and implementing strategies to optimize performance and minimize downtime. Since 1935, Morgan Stanley is known as a global leader in financial services, always evolving and innovating to better serve our clients and our communities in more than 40 countries around the world. Department ProfileServices Technology is a division within Wealth Management Technology that enhances technology solutions to improve business processes and service delivery. The division leverages a range of tools and technologies to automate processes, increase efficiency, and improve the effectiveness of business services. The Services Technology organization delivers platforms that support core client and advisor experiences. Our teams build and maintain solutions for the Contact Center, digital business automation, workflow orchestration, CRM & Salesforce, and client reporting. Our systems are designed to be innovative and resilient, helping to serve clients more efficiently and enabling seamless operations across the business. Job Summary
We are looking for a Site Reliability Engineer with a minimum of 5 years of industry experience, preferably working in the financial IT community. The position in the WM Product Technology team is focused on delivering exceptional services to both BU and Dev partners to minimize/avoid any production outages. The role will focus on production support within the WM Product Technology automating deployments and working with the agile teams to build and support stable and reliable production systems. The ideal candidate will be passionate about automation and skilled in one of the programming language Python/PERL/SHELL, Ruby, JAVA, C# or the like. Candidate should possess a strong understanding of database concepts, job scheduler, MQ, Web services, UNIX/LINUX/Windows OS as well as experience with debugging applications. We are looking for a strong leader with excellent communications skills who is committed to continuously improving and delivering results. Candidate should be organized, disciplined, detail-oriented, self-motivated, and delivery-focused.
Responsibilities:
- Maintain applications once they are live by measuring and monitoring availability, latency and overall system health with a focus on business activities and continuously evaluate cost and TOIL.
- Engage in and improve the whole lifecycle of services from inception and design, through deployment, operation, capacity planning and launch reviews.
- Scale systems sustainably through mechanisms like automation and evolve systems by pushing for changes that improve reliability and velocity; includes automation for other various operational needs.
- Troubleshoot infrastructure issues, reviewing log files, updating documentation, and having knowledge base with resolutions
- Work closely with the application Development team to understand the platform and create tools/utilities to help with production management
- Work with upstream data providers and upstream consumers, and reducing the amount of escalation to development teams
- Develop scripts and assist with code changes along with operational tasks/activities.
- Work closely with Application Development to ensure that the support team has excellent knowledge of the application set, own and maintain support knowledgebase and documents.
- Use analytical skills to find trends in the environment and drive out problems.
- Lead effort to determine improvement areas to stabilize the plant.
- Identify risks and work with a sense of urgency, working within a team or independently.
- Test and tune network, hardware, and software configurations to maximize performance
- Interface with different teams like IT Dev managers, Infrastructure teams and lead as a Subject Matter Expert (SME) for the application(s) supported.
- Understand the overall business flow of supported application systems and its interface with clients
- Take ownership and managing production requests, questions, issues and perform Root Cause Analysis for outages/incidents
- Understand the overall business flow of supported application systems and its interface with clients
- Be flexible to provide weekend on call rotation and available for offshore time lead
- Be accountable for the Production Environments as well as the non-Production Environments and be part of 24/7 production support coverage.
- 5+ years of experience in a production environment with a solid software development background and understanding of performance tuning, end-to-end troubleshooting, networking fundamentals and appropriate attention to detail
- Ability to focus, provide resolutions for production issues in a high demanding and pressured environment
- 5+ years' hands-on experience in designing, developing, and implementing technical solutions, or significant experience in deep technical support
- Strong experience in scripting language (Shell scripting, Python, Perl, etc.) and cloud driven development
- Strong database skills with DB2, Sybase or Oracle
- Hands-on experience with Autosys or other batch scheduling software
- Strong experience in Continuous Integration and Continuous Deployment
- Strong experience in environment on demand for both Virtual Machines and containers
- Knowledge and hands-on experience with monitoring tools like Splunk, IP Soft, Sockeye
- Practical experience in Agile Methodology (e.g. Scrum)
- Knowledge or experience with automating deployments using Jenkins and Train
- Ability to diagnose technical problems, debug, optimize code, and automate routine tasks
- Hands-on experience in application and database troubleshooting/issue resolution in a fast-paced environment
- Excellent communication and ability to think out of the box for process improvements.
- Knowledge of Cloud based deployment, security, networking concepts in Azure and AWS
- Hands on experience leveraging generative AI tools to enhance research, automate and improve productivity
- Knowledge or experience with algorithms, data structures, complexity analysis and software design
- Interest in designing, analyzing and troubleshooting large-scale distributed systems
- Minimum BS degree in Computer Science, Engineering or a related field
At Morgan Stanley, we raise, manage and allocate capital for our clients - helping them reach their goals. We do it in a way that's differentiated - and we've done that for 90 years. Our values - putting clients first, doing the right thing, leading with exceptional ideas, committing to diversity and inclusion, and giving back - aren't just beliefs, they guide the decisions we make every day to do what's best for our clients, communities and more than 80,000 employees in 1,200 offices across 42 countries. At Morgan Stanley, you'll find an opportunity to work alongside the best and the brightest, in an environment where you are supported and empowered. Our teams are relentless collaborators and creative thinkers, fueled by their diverse backgrounds and experiences. We are proud to support our employees and their families at every point along their work-life journey, offering some of the most attractive and comprehensive employee benefits and perks in the industry. There's also ample opportunity to move about the business for those who show passion and grit in their work.
To learn more about our offices across the globe, please copy and paste into your browser. Morgan Stanley is an equal opportunity employer committed to building and maintaining a workforce that is diverse in experience and background. Our recruiting efforts reflect our strong commitment to a culture of inclusion, where individuals are hired, developed, and advanced based on their skills and talents. Our workforce reflects a broad cross-section of the global communities in which we operate, bringing a variety of backgrounds, talents, perspectives, and experiences. For more information, please visit:
Vacancy posted 5 days ago
Similar jobs that could be interesting for youBased on the Site Reliability Engineer in Alpharetta, GA vacancy
- ...communities.This is a Lead Software Production Management & Reliability Engineering position at Director level which is part of the job family responsible... ...across the business.Job SummaryWe are looking for a Site Reliability Engineer with a minimum of 5 years of industry...SuggestedFlexible hoursWeekend work
$128k - $216k
...another millions of times a day - quickly, reliably, and securely. Any time you swipe your... ...make a difference at Fiserv.Job TitleSr. Site Reliability EngineerAbout CloverClover is... ...does a successful Senior Site Reliability Engineer do at Fiserv?As a Senior Site Reliability...SuggestedFull timeWorldwide- ...unwavering security to responsibly propel the global lottery industry ever forward.Position SummaryWe are looking for a skilled Site Reliability Engineer (SRE) to enhance the stability, performance, and reliability of our production systems. The SRE will work closely with...SuggestedPermanent employmentFull timeWork experience placementLocal area
$86.6k - $144.4k
...platforms and using automation to solve complex security and reliability challenges?Do you enjoy shaping the future of security... ...You can learn more about LexisNexis Risk at our TeamOur Site Reliability Engineering (SRE) team plays a critical role in ensuring the...SuggestedFull timeLocal area$129k - $161k
...Job title: Senior Site Reliability Engineer Reports to: Director, Site Reliability Engineering Department: Cloud Platforms Location: Remote Grade: 20 About Priority Commerce: Priority Commerce is a leading financial technology company on a...SuggestedRemote work$125k - $175k
...our Firm, enabling our clients and colleagues to redefine markets and shape the future of our communities. This is a Lead Site Reliability Engineer position at Vice President level, which is part of the job family responsible for overseeing the production environment, ensuring...Temporary work- ...and shape the future of our communities.This is a Software Engineering position at Director level, which is part of the job family... ...our businesses. This role is for an experienced and driven Site Reliability Engineer (SRE) to join our AI Platform team to help support,...
- ...can create the conditions for educators to teach, students to thrive, and districts to shape the future of education. Site Reliability Engineer (SRE) Overview: We are looking for a Site Reliability Engineer (SRE) to join our Engineering team. This is a build-it...Full timeLive inWork at office
$118.3k - $219.8k
Are you excited to lead Site Reliability Engineering teams that keep mission-critical, 24/7 services running reliably and securely?Do you enjoy building automated cloud platforms, hardening security, and driving ongoing cost optimization through strong FinOps practices...Full timeLocal area$167.7k - $245.2k
...Cisco Meraki, we are responsible for building and growing the cloud that supports these customers and their networks. As a Site Reliability Engineer, you will be focused on supporting a specific, highly available, and very secure production environment. You will analyze...Full timeTemporary workLocal areaFlexible hours- ...of a company that values diversity, integrity, and growth.Role OverviewPDI Technologies is looking for a Senior Manager, Site Reliability Engineering to lead the SRE organization supporting Paylo, PDI’s payments, loyalty, and fuel-pricing product suite. This role owns...Full time
$125k - $175k
...overseeing the production environment, ensuring the operational reliability of deployed software, and implementing strategies to optimize... ...in quantitative discipline (Computer Science, Computer Engineering). - 5+ years’ experience in leading a small to medium team of...Full timeTemporary work- ...On-Site role Job Description: ~4 - 10 hour days (Sunday-Wednesday 7am-5pm) ~ Database Site Reliability Engineer (Database Operations) Position Summary: We are seeking an experienced Database Site Reliability Engineer (SRE) to support and operate...Permanent employmentTemporary work
- ...of your work. You are visible, your talents are valued, and you are empowered to shape the future of payments.As a Senior Site Reliability Engineer (SRE) - Azure & GitOps (CI/CD) in Norcross, GA or Omaha, NE, you will join a diverse, passionate team, dedicated to powering...Full timeLocal areaWorldwide
- ...storage tanks, water metering, energy metering, gas monitoring, and asset management. Our founders are hardcore telecommunications engineers with combined 200 + years of experience in designing, optimizing and performance engineering; for several mid - large wireless...Work experience placement
- ...passionate and love what we do! We are at the forefront of future engineering technologies, with solutions that ensure the success of our... ...shaping the future of the world we live in.Job Description: Reliability EngineerJob Location: Hartford City, INTravel: Approximately...Live inWork at office
- ...storage tanks, water metering, energy metering, gas monitoring, and asset management. Our founders are hardcore telecommunications engineers with combined 200 + years of experience in designing, optimizing and performance engineering; for several mid - large wireless...
- ...storage tanks, water metering, energy metering, gas monitoring, and asset management. Our founders are hardcore telecommunications engineers with combined 200 + years of experience in designing, optimizing and performance engineering; for several mid - large wireless...
- ...storage tanks, water metering, energy metering, gas monitoring, and asset management. Our founders are hardcore telecommunications engineers with combined 200 + years of experience in designing, optimizing and performance engineering; for several mid - large wireless...
$101.6k - $152.4k
We are seeking a talented Senior Engineer I, Digital Solutions to join our team and take charge of designing, developing, and deploying... ...alarms, and reports.Travel: Willingness to travel to customer sites as required. Travel is roughly expected to be around 25% but is...Full timeTemporary workImmediate startRemote workWork from homeFlexible hours- ...IT capabilities and seeking an IT Support Engineer who thrives in a communicative,... ...Service Provider (MSP) to ensure smooth and reliable service delivery. This position offers significant... ...access governance, permissions, and site structure Microsoft Teams configuration...
- ...Software Systems Engineer - IVAmerica Networks is a leading sensor and networking solutions partner for companies in any Industrial, Manufacturing, and Waste management space. We design and manufacture sensors for storage tanks, water metering, energy metering, gas monitoring...
$63.22k - $94.55k
...maintain Windows Server environments with a focus on performance, reliability, and scalability Develop and maintain PowerShell scripts... ...excellence and process optimization Mentor junior engineers and contribute to knowledge sharing across the team Required...Remote work$63.22k - $94.55k
...States (US). The role involves designing, implementing, and maintaining Windows Server environments with a focus on performance, reliability, and scalability. The Lead Systems Programmer will also develop and maintain PowerShell scripts for automation of administrative...Remote work$101k - $194k
...together — lifting our communities and building trust in how we show up, everywhere & always. Want in? Join the #VTeamLife.As a Senior Engineer, you will serve as the technical custodian and primary oversight for the Customer Experience Platform (CXP) services. You will...Full timeTemporary workPart timeWork experience placementWork at officeWork from homeShift work3 days per week$102.3k - $147.05k
...so do you.About The Team: UKG is looking for a Senior Software Engineer III - Eng to build and maintain our public/private cloud’s... ...architectural design of new features and systems, ensuring scalability, reliability, and maintainability.Code Review: Diligent about reviewing...- Job PostingInstallation, upgrade and maintenance of z/OS and related productsSound knowledge in SMPEInstallation and maintenance of program productsProficiency in Assembler, Rexx, CLIST is advantageGood knowledge of system initialization and tuningGood Knowledge & Experience...
- ...IT capabilities and seeking an IT Support Engineer who thrives in a communicative,... ...Service Provider (MSP) to ensure smooth and reliable service delivery. This position offers significant... ...access governance, permissions, and site structureMicrosoft Teams configuration, policies...
- The Software Engineer III (Electronics) is responsible for the development of software for embedded microcontrollers. Software development... .... Our full line of global air and water solutions deliver reliable performance, comfort and energy savings for residential and commercial...Contract workWork experience placementRemote workWorldwideMonday to Thursday
$85k - $101k
...and consumer electronics solutions. Our North American R&D and engineering team works closely with Global HQ in South Korea to develop and... ...Post-Sales SupportSupport pre-sales technical assessments and on-site control system evaluations.Develop application guidelines,...Full timeTemporary workFor contractorsLocal areaImmediate start
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!
Related searches
- site reliability engineer Alpharetta, GA
- site reliability engineer sre Alpharetta, GA
- junior website developer Alpharetta, GA
- construction site safety Alpharetta, GA
- site services specialist Alpharetta, GA
- website content developer Alpharetta, GA
- on-site clinical research associate (traveling/remote) Alpharetta, GA
- historic site Alpharetta, GA
- official site Alpharetta, GA
- site leader Alpharetta, GA



