Назад
Company hidden
28 Π΄Π½Π΅ΠΉ Π½Π°Π·Π°Π΄

Data Scientist (Machine Learning)

Π€ΠΎΡ€ΠΌΠ°Ρ‚ Ρ€Π°Π±ΠΎΡ‚Ρ‹
onsite
Π’ΠΈΠΏ Ρ€Π°Π±ΠΎΡ‚Ρ‹
fulltime
Π“Ρ€Π΅ΠΉΠ΄
senior
Английский
b2
Π‘Ρ‚Ρ€Π°Π½Π°
Peru
Вакансия ΠΈΠ· списка Hirify.GlobalВакансия ΠΈΠ· Hirify Global, списка ΠΌΠ΅ΠΆΠ΄ΡƒΠ½Π°Ρ€ΠΎΠ΄Π½Ρ‹Ρ… tech-ΠΊΠΎΠΌΠΏΠ°Π½ΠΈΠΉ
Для мэтча ΠΈ ΠΎΡ‚ΠΊΠ»ΠΈΠΊΠ° Π½ΡƒΠΆΠ΅Π½ Plus

ΠœΡΡ‚Ρ‡ & Π‘ΠΎΠΏΡ€ΠΎΠ²ΠΎΠ΄

Для мэтча с этой вакансиСй Π½ΡƒΠΆΠ΅Π½ Plus

ОписаниС вакансии

ВСкст:
/
TL;DR
Data Scientist (Machine Learning): Building and evaluating entity-matching models for messy, large-scale company data with an accent on embeddings, LLMs, NLP, classification, and experimental rigor. Focus on designing ranking and similarity approaches, analyzing model behavior and error cases, and balancing model quality, inference cost, and scalability.

Location: Lima, Peru

Company

hirify.global designs, builds, and scales AI-powered solutions by combining data, artificial intelligence, cloud, design, and development.

What you will do

  • Build and evaluate machine learning approaches for company and entity matching.
  • Develop embedding- and LLM-based matching, scoring, ranking, and similarity methodologies.
  • Work with messy, multilingual data including names, aliases, domains, websites, firmographic attributes, and data hierarchies.
  • Define benchmark datasets, baselines, metrics, test sets, and error-analysis processes.
  • Design experiments, analyze model behavior and trade-offs, and compare LLM-assisted solutions with lower-cost alternatives.
  • Communicate recommendations to engineering and business stakeholders and document successful and unsuccessful experiments.

Requirements

  • 5+ years of professional Data Science or Machine Learning experience.
  • Strong applied machine learning fundamentals and experience with supervised and unsupervised learning, classification, NLP, embeddings, semantic similarity, and LLMs.
  • Excellent Python and SQL skills.
  • Working knowledge of neural networks and transformer architectures.
  • Hands-on experience with TensorFlow, PyTorch, PyCaret, or equivalent machine learning frameworks.
  • Strong English communication skills.

Nice to have

  • Entity resolution, record linkage, deduplication, ranking, or similarity-scoring experience.
  • Retrieval, clustering, or candidate-generation experience.
  • Experience with Spark, Snowflake, Databricks, or BigQuery.
  • Experience with company, domain, website, or firmographic data.
  • Experience working with multilingual datasets.

Culture & Benefits

  • High-performance culture built around empowering excellence, collaborative teamwork, respect, transparency, and efficient communication.
  • Opportunity to work on AI-native transformation and scalable products with cross-functional teams.
  • Environment focused on learning quickly, taking ownership, and modern ways of working.

Π‘ΡƒΠ΄ΡŒΡ‚Π΅ остороТны: Ссли Ρ€Π°Π±ΠΎΡ‚ΠΎΠ΄Π°Ρ‚Π΅Π»ΡŒ просит Π²ΠΎΠΉΡ‚ΠΈ Π² ΠΈΡ… систСму, ΠΈΡΠΏΠΎΠ»ΡŒΠ·ΡƒΡ iCloud/Google, ΠΏΡ€ΠΈΡΠ»Π°Ρ‚ΡŒ ΠΊΠΎΠ΄/ΠΏΠ°Ρ€ΠΎΠ»ΡŒ, Π·Π°ΠΏΡƒΡΡ‚ΠΈΡ‚ΡŒ ΠΊΠΎΠ΄/ПО, Π½Π΅ Π΄Π΅Π»Π°ΠΉΡ‚Π΅ этого - это мошСнники. ΠžΠ±ΡΠ·Π°Ρ‚Π΅Π»ΡŒΠ½ΠΎ ΠΆΠΌΠΈΡ‚Π΅ "ΠŸΠΎΠΆΠ°Π»ΠΎΠ²Π°Ρ‚ΡŒΡΡ" ΠΈΠ»ΠΈ ΠΏΠΈΡˆΠΈΡ‚Π΅ Π² ΠΏΠΎΠ΄Π΄Π΅Ρ€ΠΆΠΊΡƒ. ΠŸΠΎΠ΄Ρ€ΠΎΠ±Π½Π΅Π΅ Π² Π³Π°ΠΉΠ΄Π΅ β†’