Platform ULL - Colo - Reliability

Squarepoint Capital · London, United Kingdom · London, Montreal, New York, Singapore · Singapore · Montreal, Canada · United States

Apply on Squarepoint CapitalPosted last yearConfirmed still open

Position:  Colo LL Reliability Specialist  - Compute 

Business Area: Infrastructure  

Job Summary:   

Squarepoint is looking for a talented and highly motivated Ultra Low Latency Platform Engineer to provide solutions across Squarepoint’s global colocation (COLOs) estate consisting of 400+ servers across 30 global sites.  The candidate will be responsible for project delivery, support escalations, monitoring, automation, security, documentation, and capacity management for Squarepoint’s low latency infrastructure. This will involve collaborating with our business partners, application owners, clients, vendors, and internal teams (SRE, Network, Application Support and Application Development, Quants, etc.) to deliver end to end solutions in a timely manner. 

  • Manage systems efficiently at scale through standardization, automation, testing, and in-depth monitoring 
  • Enforce development standards for source control, testing, and continuous integration for infrastructure, OS, patches, and configuration management 
  • Manage a distributed compute environment and multiple petabyte-scale storage systems 
  • Install, manage, and monitor the Linux operating system (RHEL based) 
  • Troubleshoot complex hardware and software issues throughout the Squarepoint technology stack 
  • Create self-healing systems and automated recovery processes 
  • Respond to system incidents and participate in on-call rotations 
  • Conduct root cause analysis of incidents and outages 
  • Reduce operational toil through the development of user-driven automated workflows 
  • Work with business owners to regularly re-prioritize the book of work, while delivering both tactical and long-term objectives

  

Required Qualifications: 

  • 5+ years of experience working with Linux (RHEL/CentOS/Rocky preferred) in a large complex or niche environment with the following areas of focus: operations, systems engineering and systems performance.
  • Server Management and Support:  HP, SuperMicro, Dell, various overclock servers.
  • Experience with Low latency network interfaces and kernel bypass (configuration and optimization):  Solarflare with onload, Mellanox with VMA.
  • Experience with build and configuration management tools, specifically Chef or Ansible.
  • Experience with observability tools, specifically Grafana and Prometheus.
  • Highly motivated and a keen eye for scripting and automation in Python, Ruby, and Bash.
  • In depth knowledge of server network stack configuration, tuning and troubleshooting including TCP, UDP(unicast/multicast), NTP, PTP, wireshark/tshark
  • Strong communication: verbal and written.
  • Critical thinking and problem-solving skills to tackle troubleshooting the unknown, glitches and the obscure.
  • Well-organized, proactive, resourceful, able to handle a fast-paced environment, question the status quo, accountable and possesses an ownership mindset.
  • Good understanding of trading venues such as Nasdaq, LSE, Euronext etc.
  • Degree in Engineering, Computer Science or related experience.

 

The minimum base salary for this role is $120,000 if located in New York.  This expectation is based on available information at the time of posting.  This role may be eligible for discretionary bonuses, which could constitute a significant portion of total compensation.  This role may also be eligible for benefits, such as health, dental, and other wellness plans, as well as 401(k) contributions.  Successful candidates’ compensation and benefits will be determined in consideration of various factors.

Apply on Squarepoint Capital

Listing sourced from the employer's careers page. Applications are handled by Squarepoint Capital.

Similar trading jobs across Europe

Ranked by how much they share with this one, not by how recent they are.

Jane Street · London · London, England

We are looking for an experienced Data Centre Engineer who can help our team keep Jane Street’s data safe and accessible around the clock, involving a mix of hands on work in our data and colocation centres, process thinking and project management.

SeniorEngineering5 months ago

Qube Research & Technologies · London

Qube Research & Technologies (QRT) is a global quantitative and systematic investment manager, operating in all liquid asset classes across the world. We are a technology and data driven group implementing a scientific approach to investing.

SeniorEngineering5 months ago

Qube Research & Technologies · London

Qube Research & Technologies (QRT) is a global quantitative and systematic investment manager, operating in all liquid asset classes across the world. We are a technology and data driven group implementing a scientific approach to investing.

SeniorEngineering5 months ago

Squarepoint Capital · London · Montreal · SG

The role of the Network Reliability Specialist is to ensure the stability and integrity of all aspects that comprise the Squarepoint global trading network, which includes the regional data centers, regional offices, and various cloud service providers.

SeniorEngineeringlast year

Jane Street · London · London, England

We are looking for an experienced Network Engineer to join our Network Security team to help us to run, scale and continuously improve our network security infrastructure through hands on operations, design, and implementation.

SeniorEngineering4 years ago

Gunvor Group · London

The System Engineer Team Lead is part of IT Infrastructure team and reports to the IT Infrastructure Manager. This role is a senior technical leader responsible for overseeing the global systems infrastructure team.

LeadEngineering17 days ago

Get these by email

One email when there is something new on Trading Careers. No account, no CV, no recruiters.

We only use it for this alert, and every email unsubscribes in one click.