8 часов назад
Colo LL Reliability Specialist - Compute (Linux)
120 000$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Colo LL Reliability Specialist - Compute (Linux): Managing and optimizing Squarepoint’s ultra-low-latency colocation infrastructure across 400+ servers and 30 global sites with an accent on Linux systems engineering, hardware support, network performance, and automation. Focus on building self-healing systems, tuning kernel-bypass networking, resolving complex infrastructure incidents, and improving reliability through observability and capacity management.
Location: London, Montreal, New York, or Singapore
Salary: Minimum base salary of $120,000 per year when located in New York, with potential discretionary bonuses and benefits.
Company
operates a global technology and infrastructure platform supporting ultra-low-latency trading systems.
What you will do
- Deliver, support, monitor, secure, document, and manage capacity for a global colocation estate of 400+ servers across 30 sites.
- Operate distributed compute environments and petabyte-scale storage systems.
- Install, manage, tune, and monitor RHEL-based Linux systems.
- Automate infrastructure operations, configuration management, testing, recovery, and user-driven workflows.
- Troubleshoot complex hardware, software, and network issues, respond to incidents, participate in on-call rotations, and perform root cause analysis.
- Collaborate with business partners, application owners, clients, vendors, SRE, network, application support, development, and quantitative teams.
Requirements
- 5+ years of experience with Linux, preferably RHEL, CentOS, or Rocky, in operations, systems engineering, or systems performance.
- Experience managing and supporting HP, SuperMicro, Dell, and overclocked servers.
- Experience configuring and optimizing low-latency network interfaces and kernel bypass, including Solarflare with Onload or Mellanox with VMA.
- Experience with Chef or Ansible, and observability tools such as Grafana and Prometheus.
- Scripting and automation skills in Python, Ruby, and Bash.
- In-depth knowledge of TCP, UDP unicast and multicast, NTP, PTP, Wireshark, and Tshark, plus familiarity with trading venues such as Nasdaq, LSE, and Euronext.
Culture & Benefits
- Work involves a fast-paced environment with ownership, proactive problem-solving, and continuous improvement.
- Benefits may include health, dental, wellness plans, and 401(k) contributions.
- Discretionary bonuses may form a significant part of total compensation.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →