8 дней назад
Sr Staff Software Engineer (Genomics)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Sr Staff Software Engineer (Genomics): Building cloud-native genomic data platforms, statistical genetics pipelines, APIs, and scientific interfaces for analyzing population-scale datasets with an accent on lakehouse architecture, high-performance systems, and production-grade genomics software. Focus on designing million-sample workflows, optimizing concurrent workloads and cloud costs, and translating advanced analytical methods into commercially releasable products.
Location: Singapore - Woodlands - NorthCoast
No salary information provided.
Company
develops genomic technologies that support early disease detection, diagnosis, treatment discovery, and broader health equity.
What you will do
- Transform statistical genetics methods, including genome-wide and phenome-wide association testing, into high-performance, scientifically accurate commercial software.
- Design cloud-native lakehouse analysis pipelines that scale to millions of samples and establish reusable low-latency analysis patterns for petabyte-scale data.
- Deliver APIs, web interfaces, and natural-language or agentic query experiences for exploring genomic, cohort, phenotype, and association data.
- Operate and optimize workloads spanning millions of concurrent jobs, including orchestration, monitoring, debugging, fault tolerance, reproducibility, and cloud cost management.
- Partner with bioinformatics scientists, cloud infrastructure engineers, and product management to define product architecture and delivery plans.
- Set technical direction through design reviews, mentorship, rapid prototyping, and hands-on engineering across a globally distributed team.
Requirements
- Degree in Computer Science, Engineering, Bioinformatics, Mathematics, or a related field.
- At least 12 years of related experience with a bachelor's degree, 8 years with a master's degree, 5 years with a PhD, or equivalent experience.
- Experience delivering large-scale cloud data-analysis software using containerization, orchestration, infrastructure automation, and CI/CD.
- Deep knowledge of lakehouse technologies such as Apache Iceberg or Delta Lake, Parquet, Spark, Trino, or Databricks, including table internals, partitioning, compaction, catalog services, and schema evolution.
- Strong systems engineering skills in C/C++ and Python, including concurrency, memory management, SIMD, vectorized processing, and tools such as Apache Arrow, DuckDB, or Velox.
- Experience with distributed workflows, cloud batch or Kubernetes, APIs, numerical validation, reproducibility, technical leadership, and AI tooling for software development.
Nice to have
- Knowledge of statistical genetics, population genomics, association studies, genomic formats, and quality-control practices.
- Experience with TileDB, TileDB-VCF, Hail, GLOW, GenomicsDB, or other genomics-native storage and query stacks.
- Experience commercializing or delivering clinical-grade genomics or life-science data products.
- Experience with agentic workflows, MCP-style integrations, machine learning, deep learning, data governance, and large-scale visualization.
- Experience with cloud cost engineering, including spot capacity, storage tiering, and egress management.
Culture & Benefits
- Inclusive environment focused on belonging, openness, collaboration, and diverse perspectives.
- Employee Resource Groups provide career development, cultural awareness, and social-responsibility opportunities.
- Commitment to equal employment opportunity and pay equity.
- Background checks may be conducted after a conditional offer.
- This position is not eligible for visa sponsorship.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →