8 часов назад
Data Scientist - NLP
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Data Scientist - NLP (Python/Transformers): Developing scalable, production-ready NLP solutions for challenging public-sector problems with an accent on preprocessing, feature engineering, statistical modeling, validation, and visualization. Focus on building classification and deep learning models, applying transformer architectures such as BERT and RoBERTa, and explaining model results for federal clients.
Location: Fully remote within the United States; support is provided for long-term federal client engagements in the DC Metro area.
Company
is a US consulting and information technology solutions provider supporting public-sector health, civilian, and national security missions.
What you will do
- Collect, clean, preprocess, and prepare datasets for NLP and computational modeling using Python.
- Develop NLP features and evaluate attributes using methods such as TF-IDF, word2vec, GloVe, and FastText.
- Select and build classification, regression, neural network, deep learning, and other machine learning models.
- Develop NLP use cases including document summarization, named entity recognition, sentiment analysis, and topic modeling.
- Validate, investigate, report, and justify model results while presenting insights to stakeholders.
- Coordinate and maintain user stories for project delivery.
Requirements
- Master’s degree in Statistics, Mathematics, Computer Science, or a related field; PhD preferred.
- At least four years of experience developing scalable, production-ready NLP solutions.
- Experience with SAS, R, or Python and frameworks including scikit-learn, Keras, TensorFlow, PyTorch, or Spark NLP.
- Experience with transformer architectures and models such as BERT, GPT-J, RoBERTa, or T5.
- Experience with open-source NLP packages such as Gensim, spaCy, or NLTK, and with Git/GitHub version control.
- Must be a US citizen and able to obtain and maintain a Public Trust security clearance.
Nice to have
- Experience with GenAI and prompt engineering.
- Experience with Databricks and MLflow.
- Experience with machine translation and transcription of foreign-language documents using Microsoft Azure translation services.
- Experience with AWS services including Bedrock and Textract.
Culture & Benefits
- Temporary position with potential conversion to a regular full-time role based on business needs, project funding, and individual performance.
- Employer-paid health care and a 401(k) match.
- Training and development funds.
- Potential bonus opportunities.
- Work is fully remote while supporting public-sector projects.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →