Назад
Company hidden
4 дня назад

Intern, AI Data Developer

Формат работы
onsite
Тип работы
fulltime
Грейд
trainee
Английский
b2
Страна
Canada
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Intern, AI Data Developer (AI/data engineering): Building and scaling telemetry data pipelines, feature engineering workflows, models, and platform integrations for compliance and analytics capabilities with an accent on data quality, model evaluation, and platform extensibility. Focus on designing production-grade ETL/ELT systems, monitoring drift and reliability, reverse-engineering system outputs, and prototyping end-to-end extensions with AI-assisted development.

Location: Toronto, ON, Canada. The 2027 Canada internship runs for 16 weeks from January 4 to April 23, 2027. The role may require in-person onboarding and/or identity verification.

Company

hirify.global develops software for designing and making buildings, machines, factories, cars, and films.

What you will do

  • Design, build, and maintain ETL/ELT pipelines for large-scale telemetry and operational data.
  • Engineer features and evaluate scoring, classification, and ranking models against benchmark datasets.
  • Build data models, schemas, and versioned datasets for engineering and analytics users.
  • Monitor data and model accuracy, drift, reliability, and regressions through validation and alerting.
  • Analyze shared platforms, identify architecture and separation-of-concerns gaps, and prototype end-to-end extensions or integrations.
  • Document architectures, data contracts, benchmarks, and technical recommendations for engineering and business stakeholders.

Requirements

  • Currently enrolled in a full-time undergraduate program in Computer Science, Engineering, Data Science, Statistics, or a related field, with expected graduation in April 2027 or later.
  • Proficiency in Python and SQL, including pandas, Scikit-learn, or PySpark.
  • Understanding of ETL/ELT design, data validation, schema design, dataset versioning, machine learning workflows, precision/recall, and experimental design.
  • Experience with Git, cloud environments such as AWS or Azure, and distributed data processing such as Spark.
  • Hands-on experience using AI coding assistants such as Cursor, Claude Code, or GitHub Copilot for development, debugging, refactoring, or analysis.
  • Comfort working on ambiguous, outcome-based problems and explaining AI-assisted technical decisions to stakeholders.

Nice to have

  • Experience with PySpark, Spark SQL, Airflow, dbt, data warehousing, anomaly detection, or feature engineering.
  • Familiarity with GenAI or LLM-powered platforms, RAG-based tools, multi-tenant architecture, and separation-of-concerns design.
  • Experience reverse-engineering system outputs, producing architecture documentation, or working in compliance, licensing, fraud, or risk detection.

Culture & Benefits

  • Paid 16-week internship with mentorship from industry leaders.
  • Participation in technical talks and professional development activities.
  • Hands-on work on production-grade pipelines and models supporting business decisions.
  • hirify.global’s Flexible Workplace approach supports office, remote, and hybrid work preferences, subject to business needs.
  • Opportunity to build a portfolio of benchmarked models, measurable outcomes, and reusable technical documentation.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →