Site Reliability Engineer/L3 Support
$110k - $120kSS&C Technologies
As a leading financial services and healthcare technology company based on revenue, SS&C is headquartered in Windsor, Connecticut, and has 27,000+ employees in 35 countries. Some 20,000 financial services and healthcare organizations, from the world's largest companies to small and mid-market firms, rely on SS&C for expertise, scale, and technology.Job DescriptionJob Title: Site Reliability Engineer (SRE) / L3 Support EngineerGetting to know us:As a leading financial services and healthcare technology company based on revenue, SS&C is headquartered in Windsor, Connecticut, and has 27,000+ employees in 35 countries. Some 20,000 financial services and healthcare organizations, from the world's largest companies to small and mid-market firms, rely on SS&C for expertise, scale, and technology.Kick off your software engineering career on our Quality & Automation team. You will learn modern test engineering practices while contributing real code, automated tests, and quality improvements. We welcome candidates new to financeWhy You Will Love It Here! Flexibility: Hybrid Work Model & a Business Casual Dress Code, including jeansYour Future: 401k Matching Program, Professional Development ReimbursementWork/Life Balance: Flexible Personal/Vacation Time Off, Sick Leave, Paid HolidaysYour Wellbeing: Medical, Dental, Vision, Employee Assistance Program, Parental LeaveWide Ranging Perspectives: Committed to Celebrating the Variety of Backgrounds, Talents and Experiences of Our Employees Training: Hands-On, Team-Customized, including SS&C UniversityExtra Perks: Discounts on fitness clubs, travel and more!What You Will Get To Do:We are looking for a Site Reliability Engineer (SRE) to join our Platform Engineering team and take ownership of the operational health, reliability, and availability of our FedRAMP High cloud platform.This role combines modern Site Reliability Engineering practices with advanced production support responsibilities. You will act as the highest level of operational support (L3), proactively identifying and resolving issues before they impact customers, driving continuous improvement, and working closely with engineering teams to improve the reliability and operability of the platform.This is not a traditional operations role. You will use automation, observability, and engineering best practices to reduce operational toil while helping development teams build resilient, secure services.Due to the nature of the environment, this position requires the successful candidate to be a U.S. Citizen and eligible to work on systems supporting FedRAMP High workloads.What you will get to do:Monitor the health, availability, performance, and security of production services.Proactively identify emerging issues using telemetry, logs, metrics, and distributed tracing.Investigate, troubleshoot, and resolve complex production incidents across application and infrastructure layers.Act as the L3 escalation point for operational issues that cannot be resolved by L1 or L2 support.Participate in an on-call rotation for critical production incidents.Lead incident response activities, including coordination, communication, and post-incident reviews.Perform root cause analysis and ensure corrective actions are implemented to prevent recurrence.Develop and maintain operational runbooks, dashboards, alerts, and standard operating procedures.Improve platform observability by enhancing monitoring, alerting, dashboards, and service-level indicators.Work closely with software engineering teams to improve service reliability, scalability, and resilience.Identify opportunities to automate operational tasks and eliminate repetitive manual work.Support production deployments, infrastructure changes, and maintenance activities.Assist with disaster recovery exercises, resilience testing, and operational readiness reviews.Ensure operational activities comply with FedRAMP High security and compliance requirements.Contribute to continuous improvement initiatives across reliability, performance, and operational excellence.What you will Bring:U.S. Citizenship (required).3–6 years of experience in Site Reliability Engineering, Production Engineering, DevOps, Platform Engineering, or a senior production support role.Experience supporting mission-critical cloud-based production systems.Strong understanding of Linux operating systems and networking fundamentals.Experience troubleshooting distributed applications running in Kubernetes.Experience with public cloud platforms, preferably AWS.Experience with infrastructure as code and configuration management.Strong scripting or programming skills (e.g. Python, Bash, PowerShell, Go, or similar).Experience using monitoring and observability platforms such as Prometheus, Grafana, CloudWatch, Datadog, Splunk, or OpenTelemetry.Experience analysing application logs, metrics, and traces to diagnose production issues.Understanding of incident management, problem management, and root cause analysis.Strong analytical and troubleshooting skills.Excellent written and verbal communication skills.Preferred QualificationsExperience supporting systems operating under FedRAMP High, DoD IL5/IL6, or similar regulated environments.Experience with Kubernetes in production.Experience with AWS services including EKS, RDS, IAM, CloudWatch, Route 53, VPC networking, and AWS Backup.Experience with CI/CD pipelines and deployment automation.Knowledge of service mesh technologies such as Istio.Familiarity with security best practices including IAM, least privilege, vulnerability management, and compliance monitoring.Experience with PagerDuty, Jira Service Management, or similar incident management platforms.AWS certification (Associate or Professional) is desirable.What Success Looks LikeWithin your first year you will:Maintain high platform availability and service reliability.Detect and resolve issues before customers experience impact.Reduce mean time to detect (MTTD) and mean time to recover (MTTR).Improve monitoring coverage and reduce unnecessary alert noise.Increase operational automation and reduce manual support effort.Produce high-quality incident reviews with actionable improvements.Partner effectively with engineering teams to continuously improve platform resilience.Why Join Us?You'll work on a modern cloud-native SaaS platform running in a highly secure FedRAMP High environment, where reliability, automation, and engineering excellence are fundamental. You'll collaborate with software engineers, platform engineers, and security specialists to build and operate systems that customers depend upon every day.Unless explicitly requested or approached by SS&C Technologies, Inc. or any of its affiliated companies, the company will not accept unsolicited resumes from headhunters, recruitment agencies, or fee-based recruitment services.SS&C Technologies offers a comprehensive total rewards package designed to support your wellbeing, growth, and future. Our benefits include medical, dental, and vision coverage; a 401(k) plan with company match; paid time off, holidays, and parental leave; and professional development reimbursement opportunity.Actual base salary will vary based on several factors, including but not limited to relevant skills, prior experience, education, demonstrated performance, and geographic location.New York: The expected base salary for the position is between 110000 USD to 120000 USD.In addition, employees in this role may be eligible for consideration on an annual basis for a discretionary bonus and/or equity awards, such as restricted stock units or stock options, based upon individual and business performance at the company’s discretion.Applications will be accepted on an ongoing basis until the position is filled.SS&C Technologies is an Equal Employment Opportunity employer and does not discriminate against any applicant for employment or employee on the basis of race, color, religious creed, gender, age, marital status, sexual orientation, national origin, disability, veteran status or any other classification protected by applicable discrimination laws.SummaryLocation: Remote - New York, US; Remote - Kansas, US; Remote - Pennsylvania US; Remote - New Hampshire, USType: Full time
$120k - $165k
...Lead Software Production Management & Reliability Engineering position at Director level which is part... ...the design, development, delivery and support of the technical solutions behind the... ...Management Production Management Site Reliability Engineer position is a highly...SuggestedTemporary workWork at office$131k - $164k
Position Overview We are seeking a highly skilled Staff Site Reliability Engineer with deep technical expertise across VMware, Linux, and... ...(RHEL/CentOS/Ubuntu) and Windows Server operating systems supporting enterprise workloads. * Integrate and maintain Active Directory...SuggestedWork at officeLocal areaVisa sponsorshipFlexible hours$197.3k - $313.7k
...duplicating efforts. Job Category Software Engineering Job Details About Salesforce... ...of Salesforce. Job Title: Director, Site Reliability Engineering Location: New York, NY; San... ...response, capacity planning, and operational support. Build automation that reduces...SuggestedFull timeImmediate start$130k - $165k
...with order data across a vast array of markets and asset classes Support systems processing that data for Surveillance and Regulatory Reporting... ...L1 and L2 support for users and collaborate with developers on L3 support Onboard support processes for new development teams and...SuggestedTemporary workFlexible hours$100k - $150k
...getting started. The Role As an Application Support Engineer, you will cover the full support spectrum... ...deep infrastructure‑level diagnosis (L3). You're the person who gets things done... ...Willingness to travel occasionally for on‑site customer visits — you're comfortable getting...SuggestedWork at officeRemote work- ...We are seeking an experienced Aladdin Application Support Engineer to join our Trading Technology team, providing L2/L3 production support for BlackRock's Aladdin platform across front office trading and portfolio construction workflows. The ideal candidate has hands-...Night shift
- ...highly experienced Senior Application Support Engineer with strong expertise in .NET technologies... ...the stability, performance, and reliability of customer-facing and backend business... ...environment. Key Responsibilities Provide L2/L3 production and application support for...
- ...SAP HCM Functional Consultant with Production support (US Citizen only) Direct message the job poster from AppLab Systems, Inc Job Summary... ...an enhancement. Location Remote Key Responsibilities Provide L2/L3 functional support for SAP HCM modules: PA, OM, TM, US Payroll,...Contract workRemote work
$123k - $165k
...Summary:Department/Group OverviewOur engineering fleet is a horizontal set of teams... .... Our specific team provides reliability engineering and operational support to backend service development teams... ...products and brands.We are seeking a Site Reliability Engineer who will...$200k - $250k
Hudson River Trading (HRT) is seeking a Senior Site Reliability Engineer to join our growing Enterprise SRE team. This team is responsible for developing and maintaining productivity service infrastructure for the entire firm, both on-prem and in the cloud. They ensure...Work at officeLocal areaImmediate start- ...the world's most complex and mission-critical systems.As a Site Reliability Engineer III at JPMorgan Chase within the Commercial & Investment... ...designs and gaining consensus from peers where appropriate, supporting adoption of site reliability engineering best practices...Shift work
$165k - $241.4k
...very effective.We’re looking for talented engineers with a software or operations background... ...development teams to ensure the reliability, performance and security of our infrastructure... ...insurance. Please see the Cisco careers site to discover more benefits and perks....Full timeTemporary workWork at officeLocal areaFlexible hours1 day per week$120k - $200k
...PermContact: Kunal DaveContact Email: ****@*****.*** Reliability Engineer(SRE) ResponsibilitiesGlobal Architecture & Disaster Recovery... ...practices (e.g., Chaos Engineering, resilience testing, automated recovery)SkillsBilingual Mandarin Site Reliability Engineer(SRE)Overseas$138.1k - $198.2k
...intuitive with technology that simply works. The SRE Engineering Enablement Team supports our CI Platforms, Developer Environments, tools, automation... ...are all engineers at Cisco. Your Impact As a Site Reliability Engineer, you will be at the epicenter of our engineering...Permanent employmentFull timeTemporary workWork experience placementLocal areaRemote workFlexible hours$158.5k - $172k
...exceptional value they deserve.About The OpportunityAs a Senior Engineer on the Runtime Automation team, you will design, automate, and... .... This is a high-impact position driving continuous reliability, deep system optimization, and automation across our entire technology...Full timeWork at office3 days per week$141k - $216.6k
...building a safer, more connected world.Position OverviewAs a Site Reliability Engineer, you'll own the reliability, observability, and operational... ...the systems, tooling, and operational practices that support our applications today while building the foundation for the...Work experience placementWork at office$140k - $205k
Senior Technology Site Reliability EngineerCooley is seeking a Senior Site Reliability Engineer to join the Infrastructure & Development Operations team.Position summary:... ...automated, resilient, and observable systems that support high availability and operational excellence....Full timeTemporary workWork at officeFlexible hoursWeekend work- ...platform and performing reconciliation reporting. Monitor systems, diagnose issues, implement fixes, and provide mission critical support. Perform application and system administration for platforms that include document class maintenance, security maintenance, document...
$160k - $240k
Senior Software Engineer - Kubernetes as a Service Location New York Business Area... ...Bloomberg’s development teams with a reliable, scalable, and feature-rich Kubernetes ecosystem... ...playbooks.Provide mentorship and support to other team members on Kubernetes best...Temporary workFor contractorsWork experience placement$150k - $250k
What We DoAt Goldman Sachs, our Engineers don't just make things - we make things possible... ...Global Banking & Markets business, the Site Reliability Engineering (SRE) team ensures the... ...load/performance testing, and production support in high-availability, latency-sensitive...Full timeTemporary workPart time- ...Cohere is a team of researchers, engineers, designers, and more, who are... ...-performance, scalable and reliable machine learning systems? Do... ...? We are looking for a Site Reliability Engineer to join... ...custom Kubernetes operators that support language model deployments.Automate...Full timeWork experience placementWork at officeLocal areaRemote workHome office
$194k - $267k
...all in on this mission. If you are too, let's talk.Position Overview:We are seeking a highly technical StaffObservabilitySite Reliability Engineer with a specialty in Splunk to own and evolve our Splunk ecosystem. In this role, you will move beyond simple monitoring to...Permanent employmentWork at officeLocal areaWorldwideFlexible hours$194k - $267k
...and who can rapidly self-educate on new concepts and tools.Position Overview:The Site Reliability Engineer (SRE) will play a key role in building and managing Kubernetes platforms that support cloud-native applications and services. This position focuses on architecting...Permanent employmentWork at officeLocal areaWorldwideFlexible hours- As a Site Reliability Engineering at JPMorgan Chase within the Enterprise technology, liquidity risk team, you are the non-functional requirement... ...environment to improve SRE workflows (e.g., incident investigation support and knowledge capture) with strong validation habits and...
$194k - $267k
...If you are too, let's talk.The TeamThe Site Reliability team is dedicated to architecting and... ...infrastructure tooling and CI/CD platforms that support Okta’s SRE ecosystem. In this... ...maximize platform reliability and engineering velocity.The ideal candidate is someone...Local areaWorldwideFlexible hours- ...firm, driven by pride in ownership.As a Senior Manager of Site Reliability Engineering at JPMorgan Chase within the Corporate Investment Bank,... ...North America leadership for production management teams supporting trading desks across multiple Markets lines of business;...Bank staffShift work
$150k - $190k
Senior Site Reliability Engineer, VPAt Morgan Stanley, we advise, originate, trade, manage and distribute capital for governments, institutions... ...any production outages. The role will focus on production support within the WM Product Technology automating deployments and...Temporary workWorldwideFlexible hoursWeekend work$195k - $275k
...department responsible for the design, development, delivery and support of the technical solutions behind the products and services... ...Planning & Release Management, and the Chief Operating Office.The Reliability Operations (RO) within WMT is responsible for providing swift,...Temporary workWork at officeWorldwideNight shift$150k - $160k
Front-End & AdTech Site Reliability Engineer (SRE)Haymarket Media, Inc. is seeking a Front-End & AdTech Site Reliability Engineer (SRE) to join the Engineering team. This position is located in our New York, NY office; three (3) days in office depending on business needs...Work at officeLocal area$160k - $240k
Senior Software Engineer - Service Mesh Security and Configuration Location New York... ...products and services, responsible for reliably handling billions of dollars of transactions... ...with flexible routing rules and support for enterprise security and monitoring capabilities...Temporary workFor contractorsWork experience placementFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer/L3 Support. Be the first to apply!
- site reliability engineer New York, NY
- site reliability engineer sre New York, NY
- site reliability engineer remote New York, NY
- after school site coordinator New York, NY
- site services specialist New York, NY
- construction site safety New York, NY
- site merchandiser New York, NY
- site leader New York, NY
- official site New York, NY
- website content developer New York, NY

