2 месяца назад
Junior SRE (ML/HPC)
150 000$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Junior SRE (ML/HPC): Building and operating production machine learning and high-performance computing infrastructure with an accent on data pipelines, workflow orchestration, performance, scalability, and reliability. Focus on designing resource allocation and data movement systems, improving observability, and supporting quantitative research and trading workloads.
Location: New York, London, or Singapore
Salary: Minimum base salary of $150,000 annually if located in New York; discretionary bonuses and benefits may also apply.
Company
develops and operates technology supporting quantitative research and trading workloads.
What you will do
- Design, implement, and maintain data pipelines for machine learning training and inference.
- Develop workflow orchestration and infrastructure components for scheduling, resource allocation, data movement, and filesystems.
- Analyze and improve the performance and scalability of systems and critical code paths.
- Improve the reliability, observability, and usability of ML and HPC platforms.
- Collaborate with quantitative researchers and developers to translate requirements into robust technical solutions.
- Progressively take ownership of well-defined platform components while learning existing systems.
Requirements
- BSc or MSc in Computer Science, Engineering, Applied Mathematics, or a related field, or equivalent practical experience.
- Strong foundations in algorithms, data structures, operating systems, and distributed systems.
- Software development experience in Python, C++, Rust, or another programming language.
- Experience writing, testing, and maintaining production-quality code in Linux environments.
- Exposure to machine learning workflows, data engineering, or ML platforms.
- Introductory knowledge of HPC, distributed systems, or cloud platforms for CPU and GPU workloads, plus experience with profiling, benchmarking, or performance analysis.
Nice to have
- Familiarity with CUDA or other GPU development frameworks.
Culture & Benefits
- Full-time role with strong mentorship and technical guidance.
- Gradual ownership of meaningful systems and components.
- Exposure to production-scale performance, scalability, and reliability challenges.
- Close collaboration with experienced engineers, quantitative researchers, and developers.
- Potential eligibility for health, dental, wellness, and 401(k) benefits.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
8 дней назад
Production Engineer (Cybersecurity)
102 400 - 128 000$
6 дней назад
Site Reliability Engineer (Python/Go)
200 000 - 225 000$
4 дня назад
Senior Site Reliability Engineer (Fintech)
160 000 - 200 000$
6 дней назад
Site Reliability Engineer (Cloud Infrastructure)
122 574 - 259 200$
3 дня назад
Senior Incident Commander (SRE/Python)
187 000 - 233 500$
4 дня назад
Junior SRE
104 500 - 162 500$