Platform ULL - Colo - Reliability

Squarepoint Capital
London, UK
3 days ago
Apply on find.jobs
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Compensation
£120,000.0
Working hours
Regular working hours
Job source

Tech stack

Systems Engineering Bash Shell CentOS Configuration Management Protocol Stack Continuous Integration Linux Python (Programming Language) Multicasting Network Interface Performance Tuning Red Hat Enterprise Linux
+11 more
Ansible Prometheus Ruby Software Engineering Transmission Control Protocol (TCP) Wireshark Grafana Infrastructure Automation Frameworks Information Technology Low Latency Software Version Control

Job description

Position: Colo LL Reliability Specialist - Compute Business Area:Infrastructure Job Summary: Squarepoint is looking for a talented and highly motivated Ultra Low Latency Platform Engineer to provide solutions across Squarepoint’s global colocation (COLOs) estate consisting of 400+ servers across 30 global sites. The candidate will be responsible for project delivery, support escalations, monitoring, automation, security, documentation, and capacity management for Squarepoint’s low latency infrastructure. This will involve collaborating with our business partners, application owners, clients, vendors, and internal teams (SRE, Network, Application Support and Application Development, Quants, etc.) to deliver end to end solutions in a timely manner. Manage systems efficiently at scale through standardization, automation, testing, and in-depth monitoring Enforce development standards for source control, testing, and continuous integration for infrastructure, OS, patches, and configuration management Manage a distributed compute environment and multiple petabyte-scale storage systems Install, manage, and monitor the Linux operating system (RHEL based) Troubleshoot complex hardware and software issues throughout the Squarepoint technology stack Create self-healing systems and automated recovery processes Respond to system incidents and participate in on-call rotations Conduct root cause analysis of incidents and outages Reduce operational toil through the development of user-driven automated workflows Work with business owners to regularly re-prioritize the book of work, while delivering both tactical and long-term objectives Required Qualifications: 5+ years of experience working with Linux (RHEL/CentOS/Rocky preferred) in a large complex or niche environment with the following areas of focus: operations, systems engineering and systems performance.Server Management and Support: HP, SuperMicro, Dell, various overclock servers.

Requirements

Experience with Low latency network interfaces and kernel bypass (configuration and optimization): Solarflare with onload, Mellanox with VMA. Experience with build and configuration management tools, specifically Chef or Ansible. Experience with observability tools, specifically Grafana and Prometheus. Highly motivated and a keen eye for scripting and automation in Python, Ruby, and Bash. In depth knowledge of server network stack configuration, tuning and troubleshooting including TCP, UDP(unicast/multicast), NTP, PTP, wireshark/tshark Strong communication: verbal and written. Critical thinking and problem-solving skills to tackle troubleshooting the unknown, glitches and the obscure. Well-organized, proactive, resourceful, able to handle a fast-paced environment, question the status quo, accountable and possesses an ownership mindset.Good understanding of trading venues such as Nasdaq, LSE, Euronext etc. Degree in Engineering, Computer Science or related experience. The minimum base

Benefits & conditions

salary for this role is $120,000 if located in New York. This expectation is based on available information at the time of posting. This role may be eligible for discretionary bonuses, which could constitute a significant portion of total compensation. This role may also be eligible for benefits, such as health, dental, and other wellness plans, as well as 401(k) contributions. Successful candidates’ compensation and benefits will be determined in consideration of various factors.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on find.jobs
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

50 sec

Why developer happiness matters in web frameworks

Eileen Uchitelle Eileen Uchitelle +1 · Coffee With Developers

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · World Congress 2025

1:45 min

Building IT systems for global finance markets

Anastasia Troitskaya Anastasia Troitskaya · World Congress 2024

3:30 min

Falling in love with Ruby and creating Basecamp

David Heinemeier Hansson David Heinemeier Hansson +1 · Coffee With Developers

3:55 min

Demonstrating .NET installation on Debian and Azure Linux

Silvano Coriani Silvano Coriani · Europe 2026 Virtual

Videos

See all

Related articles

See all