Platform ULL - Colo - Reliability
$120kSQUAREPOINT CAPITAL
Position: Colo LL Reliability Specialist - Compute Business Area:Infrastructure Job Summary: Squarepoint is looking for a talented and highly motivated Ultra Low Latency Platform Engineer to provide solutions across Squarepoint’s global colocation (COLOs) estate consisting of 400+ servers across 30 global sites. The candidate will be responsible for project delivery, support escalations, monitoring, automation, security, documentation, and capacity management for Squarepoint’s low latency infrastructure. This will involve collaborating with our business partners, application owners, clients, vendors, and internal teams (SRE, Network, Application Support and Application Development, Quants, etc.) to deliver end to end solutions in a timely manner. Manage systems efficiently at scale through standardization, automation, testing, and in-depth monitoring Enforce development standards for source control, testing, and continuous integration for infrastructure, OS, patches, and configuration management Manage a distributed compute environment and multiple petabyte-scale storage systems Install, manage, and monitor the Linux operating system (RHEL based) Troubleshoot complex hardware and software issues throughout the Squarepoint technology stack Create self-healing systems and automated recovery processes Respond to system incidents and participate in on-call rotations Conduct root cause analysis of incidents and outages Reduce operational toil through the development of user-driven automated workflows Work with business owners to regularly re-prioritize the book of work, while delivering both tactical and long-term objectives Required Qualifications: 5+ years of experience working with Linux (RHEL/CentOS/Rocky preferred) in a large complex or niche environment with the following areas of focus: operations, systems engineering and systems performance.Server Management and Support: HP, SuperMicro, Dell, various overclock servers. Experience with Low latency network interfaces and kernel bypass (configuration and optimization): Solarflare with onload, Mellanox with VMA. Experience with build and configuration management tools, specifically Chef or Ansible. Experience with observability tools, specifically Grafana and Prometheus. Highly motivated and a keen eye for scripting and automation in Python, Ruby, and Bash. In depth knowledge of server network stack configuration, tuning and troubleshooting including TCP, UDP(unicast/multicast), NTP, PTP, wireshark/tshark Strong communication: verbal and written. Critical thinking and problem-solving skills to tackle troubleshooting the unknown, glitches and the obscure. Well-organized, proactive, resourceful, able to handle a fast-paced environment, question the status quo, accountable and possesses an ownership mindset.Good understanding of trading venues such as Nasdaq, LSE, Euronext etc. Degree in Engineering, Computer Science or related experience. The minimum base salary for this role is $120,000 if located in New York. This expectation is based on available information at the time of posting. This role may be eligible for discretionary bonuses, which could constitute a significant portion of total compensation. This role may also be eligible for benefits, such as health, dental, and other wellness plans, as well as 401(k) contributions. Successful candidates’ compensation and benefits will be determined in consideration of various factors.
$120k
Position Overview:The role of the Network Reliability Specialist is to ensure the stability... ...technology team to deliver an optimal trading platform. High Level Designs and implementing... ...inter and intra region between colo to colo, colo to DC, or DC to colo.Project...PlatformWork at office- ...A dynamic fintech company in New York is seeking a Product & Platform Monitoring Manager to ensure the reliability of its fintech products. The role focuses on end-to-end monitoring of customer journeys, incident management, and collaboration with various teams. Candidates...Platform
- ...A financial technology company based in New York is seeking a Product & Platform Monitoring Manager to ensure the reliability and health of their fintech products. The role involves monitoring customer journeys, APIs, and incident management, requiring 5+ years of experience...Platform
- ...Brooklyn, NY, seeks a Principal Automation Engineer to provide Site Reliability engineering for automation services and lead DevOps... ...sessions for new services, mentor engineers, and ensure security and scalability across the platform in a 24/7 operation. #J-18808-Ljbffr...PlatformWork at office
- ...Senior Database Reliability Engineer (DBRE)Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted... ...optimize the data persistence layer that powers our large-scale, mission-critical systems. You will work closely with SRE, Platform, and....Platform
- ...Type AnnuallyIndustry Broker DealerSelling Points Drive impactful system reliability initiatives in a dynamic financial environment. Collaborate with experts to enhance trading and settlement platforms. Elevate your career with continuous improvement opportunities.Job...Platform
- ...is seeking a hands-on SRO Engineer to help establish and scale reliability practices across its technology organization in a hybrid cloud... ...reliability, resilience, and supportability of enterprise platforms and services. You will translate reliability principles into clear...Platform
- ...Bank of New York is seeking an experienced Cloud AWS Support Reliability Engineer (SRE) to build and maintain scalable AWS infrastructure... ..., security, and resilience across enterprise cloud platforms. Key responsibilities include deploying Terraform-based infrastructure...Platform
- Basis is seeking a Site Reliability Engineer to ensure the reliability, scalability, and performance of our AI-powered accounting platform. You’ll join a high-leverage infrastructure team at the intersection of product and platform, owning systems that keepBasis fast,...Platform
- Future Secure AI is seeking a VP of Platform Engineering to lead Platform Engineering, DevOps, Cloud, and Data Platforms across the engineering... ...organization. This leader is accountable for delivery, reliability, and operational integrity of our enterprise AI Co-Worker...Platform
$160k - $180k
Socure is seeking a Site Reliability Engineer in New York to enhance our identity trust infrastructure. In this role, you will take full ownership of AWS and Kubernetes platforms, ensuring high reliability and operability. The ideal candidate will possess extensive experience...Platform- ...Infrastructure/SRE engineer to design and automate large-scale infrastructure. You will own provisioning, migrations, and reliable, self-service platforms for other engineers across Kubernetes clusters, databases, and services. You will reduce toil, implement scalable...Platform
- ...leader for the role of Vice President, Global Production Operations & Reliability. This executive will shape strategy, drive execution, and elevate the reliability of our cloud-native SaaS platform worldwide. Reporting to the CTO, you will lead global SRE and DRE teams...PlatformWorldwide
- the company, a fast-growing global crypto company, seeks a Head of Site Reliability Engineering (SRE) to lead the infrastructure reliability strategy at scale. You will join senior engineering leadership, shaping the vision while ensuring uptime and security to meet the...Platform
$150k - $180k
...curious Support Engineer to join the Front Office Technology team in New York. The role involves owning the reliability of Keystone, Capstone's trading platform, and engaging in automation through Python tooling and Grafana dashboards. Responsibilities include monitoring...Platform- ...accounts and product teams to resolve the most challenging issues on our API platform. You will help design and run monitoring, alerting, and incident response processes to ensure reliability at scale. You will work closely with engineering and infrastructure teams, tackling...PlatformRemote work
- ...Overview We are seeking an experienced Site Reliability Engineer (SRE) – Microsoft Hyper-V & Private Cloud to operate highly available private cloud and Virtual Desktop Infrastructure (VDI) platforms based on Microsoft Hyper-V. This role combines deep Hyper-V expertise...Platform
- Zelis is modernizing healthcare data platforms and building a data‑native future. The Director, ZDC Operations will own the run‑state... ...foundations, establish a CoE, and lead a distributed team to deliver reliable data pipelines and dashboards across Snowflake. You will drive...Platform
$200k - $250k
Hudson River Trading (HRT) is seeking a Senior Site Reliability Engineer to join our growing Enterprise SRE team. This team is responsible... ...or AnsibleExperience building and operating services on cloud platforms (e.g. AWS, Azure, GCP)Prior experience working within an IT...PlatformWork at officeLocal areaImmediate start- United States Digital Space LLC is seeking a Senior Database Reliability Engineer to design, build, and scale the company’s database platform used across applications. This role blends database expertise with software and infrastructure engineering to create reliable,...Platform
$120k - $200k
...: PermContact: Kunal DaveContact Email: ****@*****.*** Reliability Engineer(SRE) ResponsibilitiesGlobal Architecture & Disaster Recovery... ...system stability for overseas businesses.Infrastructure Platform Deployment & Operations Manage the deployment, operation, and...PlatformOverseas- ## Site Reliability Engineer (FedRAMP / Security)New York, US · Full-time · Senior#### About The PositionCoralogix is a modern, full-stack observability platform transforming how businesses process and understand their data. Our unique architecture powers in-stream analytics...PlatformFull timeRemote work
- ...We are seeking an experienced Site Reliability Engineer (SRE) – Microsoft Hyper-V & Private Cloud to operate highly available private cloud and Virtual Desktop Infrastructure (VDI) platforms based on Microsoft Hyper-V. This role combines deep Hyper-V expertise with modern...PlatformLocal area
$98.28k - $154.44k
...VibrationDuPont is seeking a highly experienced Maintenance & Reliability Engineer to serve as the Company's global subject matter expert... ...Predictive Maintenance (IPM), machinery health monitoring platforms, rotating equipment diagnostics, dynamic balancing, laser alignment...PlatformContract workFor contractors$151k - $297k
...TPM for SRE, you will partner with SRE leaders and engineers to scale the platform that underpins all of MongoDB’s cloud products. You will drive program execution, strengthen production reliability practices, and coordinate cross-functional efforts across US and EMEA...PlatformLocal areaRemote workWorldwideFlexible hours$182k - $250.8k
...on this mission. If you are too, let's talk.The SRE Leadership TeamThe SRE Leadership Team at Okta is the backbone of our platform's reliability and operational excellence. We are a forward-thinking group of engineers and leaders who believe that great infrastructure...PlatformPermanent employmentLocal areaRemote workWorldwideFlexible hoursWeekend workWeekday work- ...significant effect in a realm tailored for top achievers in site reliability. As a Lead Site Reliability Engineer at JPMorgan Chase within... ...the Production Management team supporting Sales Execution platforms across Rates, Credit, FX, SPG, and Repo, setting direction, priorities...Platform
$120k - $142k
...Overview & Responsibilities: At The New York Times, our Site Reliability Engineering (SRE) team is central to how we design, test, and... ...Product Manager to lead the strategy for reliability programs and platforms that help teams ship resilient systems with confidence. These...PlatformLocal areaFlexible hours$125k - $130k
...Astronomer Customer Reliability Engineering RoleAstronomer empowers data teams to bring mission-critical software, analytics, and AI to... ...the company behind Astro, the industry-leading unified DataOps platform powered by Apache Airflow®. Astro accelerates building reliable...PlatformRemote workWeekend work$200k - $300k
Hudson River Trading (HRT) is seeking a Software Engineer focused on GPU reliability to join our Systems Development team. The Systems Development team builds and maintains the platform that is shared by all Systems teams to provision, monitor, and manage HRT’s server...PlatformWork at officeLocal areaImmediate start
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Platform ULL - Colo - Reliability. Be the first to apply!

