яндекс

NLP-разработчик в команду качества претрейна Alice LLM (Python)

Ищем ML-разработчика, который будет улучшать качество претрейна Alice AI LLM. Это направление, отвечающее за умность базовой модели и качество данных.

яндекс · На сервисе с: 22.09.26 10:08

285k–436k ₽РоссияМоскваОфис

Ищем ML-разработчика, который будет улучшать качество претрейна Alice AI LLM. Это направление, отвечающее за умность базовой модели и качество данных.

Гонка LLM ускоряется: модели решают задачи в математике, коде, агентах, медицине, которые ещё недавно требовали участия человека. При этом крупные улучшения LLM происходят при обновлении и масштабировании претрейна: именно на этом этапе закладываются базовые способности модели, от которых зависит качество на всех последующих срезах.

Мы считаем, что побеждает в этой гонке тот, кто эффективнее и быстрее улучшает претрейны, и вкладываемся в это направление: исследуем способы улучшения качества моделей, ставим тяжёлые вычислительные эксперименты, обрабатываем и синтезируем данные огромного масштаба.

Результат нашей работы — это модели семейства Alice AI, которые доступны в одноименном сервисе. Делая нашу модель умнее, вы улучшите главного и самого массового AI-ассистента в России, которым уже пользуется более 24 млн человек еженедельно.

Что мы уже сделали:
Наши текущие претрейн-модели уже успешно показали себя в предыдущих релизах Алисы. Например, мы писали на Хабре, что показываем лучшее качество на фактовых и культурных срезах, замерах и не только. Об исследованиях, которые легли в основу этих улучшений, мы говорили на Practical ML Conf 2025.

Чем вам предстоит заниматься?
Ризонинг-способности модели имеют доказанный профит в таких областях, как математика, код и агенты, но передовые игроки получают хорошие приросты качества за счёт рассуждений на более широком спектре задач: от финансов до компьютерных игр. Наша цель — сделать рассуждающую модель Алисы сильнейшей в РФ и не только.

Какие задачи вас ждут

Эксперименты с претрейн-моделями
Ставить и анализировать вычислительно тяжёлые эксперименты с претрейн-стадией. Проверять гипотезы об источниках прироста ризонинг-способностей внутри различных срезов умности. Постановка и выбор сетапа эксперимента — это довольно творческий этап.

Пайплайны данных
Строить эффективные пайплайны сбора обучающих данных, их синтеза, обработки и фильтрации в новых доменах. Построенные пайплайны предстоит масштабировать и делать более эффективными.

Изменение процесса обучения
Модифицировать процесс обучения модели для задач улучшения ризонинга. Внедрять изменения в обучающий пайплайн и оценивать их влияние на качество.

Экспертные модели
Обучать модели-верификаторы для проверки качества данных и ответов. Обучать модели для генерации обучающих данных, которые являются SOTA в своём узком домене. Обучать модели-фильтры для отсева некачественных данных.

Исследование скейлинга
Исследовать законы скейлинга качества модели с точки зрения данных и вычислений. Определять, как масштабирование влияет на рост рассуждающих способностей в разных доменах.

Больше об ML в Яндексе — в канале Yandex for ML

Мы ждём, что вы

  • Глубоко разбираетесь в NLP
  • Знаете Python и разрабатывали на нём

Будет плюсом, если вы

  • Обучали претрейн-модели в любой из модальностей (NLP, CV, звук, рекомендации)
  • Строили эффективные CPU/GPU-пайплайны
  • Погружены хотя бы в одну из областей исследования языковых моделей (RL, архитектуры или данные для LLM)
Эта вакансия также есть на:Сайт компании

Похожие вакансии Data Science & ML

hh.ru
онтаргет лабс

Computer Vision engineer

онтаргет лабсНа сервисе с: 25.08.26 23:52↑ Вакансия с автоподнятием
Зарплата не указанаРоссияГрузияСанкт-ПетербургТбилисиУдалёнка
Локации (удалёнка):Санкт-ПетербургТбилиси

Удалённая работа — географически не привязана. Щёлкни по любой из локаций, чтобы открыть оригинальную карточку.

OnTarget Labs is a leading international software product development company.
We create next generation of world class product lines.

The company is looking for a Computer Vision/AI engineer to join our innovative product team as a full-time member working REMOTELY.
Lots of opportunities for professional growth and business trips abroad are offered.
Join our friendly team of IT professionals now!

Product description

We are building AI-powered tennis video analysis from smartphone-recorded match footage.

Role description

We are looking to bring in a Computer Vision + AI Engineer to support machine learning and video analytics work across multiple models and product capabilities.

The role would support our internal team across the full CV/ML pipeline, from dataset quality and model evaluation through production-oriented model improvement, custom tracking/interpretation logic, and selective edge/mobile deployment.

The engineer will operate under Head of Engineering / CTO guidance but should be skilled enough to independently assess problems, recommend experiments, implement improvements, and communicate technical tradeoffs clearly.

Key skillset / experience requirements:

  • Strong practical computer vision experience, especially with real-world video
  • Strong Python and PyTorch experience
  • Experience with object detection, object tracking, classification, and/or keypoint-style models
  • Ability to evaluate model outputs, diagnose failure modes, and recommend targeted improvements
  • Strong understanding of dataset strategy, annotation quality, validation design, and metric interpretation
  • Experience designing or improving custom tracking, smoothing, temporal decoding, or model interpretation layers
  • Experience improving models toward production-level robustness across diverse real-world conditions
  • Experience deploying and optimizing cloud-based video processing pipelines
  • Experience with on-device edge/mobile inference, and skilled with deployment workflows, such as ONNX, Core ML, TensorFlow Lite, quantization, pruning, latency optimization, or mobile performance tuning
  • Bachelor’s degree in Information Systems, Computer Science, or a related field
  • Excellent verbal and written communication skills in English

Nice-to-have experience:

  • Sports video analytics
  • Ball/racket/puck tracking or other fast-moving object detection
  • Smartphone-recorded or consumer-quality video

We offer

  • Competitive compensation to be defined upon the interview results
  • Full time REMOTE WORK
Другие площадки
S

Staff Data Scientist, Watchlist

socureНа сервисе с: 24.09.26 18:12
Зарплата не указанаСШАУдалёнка

Why Socure?

Socure is building the identity trust infrastructure for the digital economy — verifying 100% of good identities in real time and stopping fraud before it starts. The mission is big, the problems are complex, and the impact is felt by businesses, governments, and millions of people every day.

We hire people who want that level of responsibility. People who move fast, think critically, act like owners, and care deeply about solving customer problems with precision. If you want predictability or narrow scope, this won’t be your place. If you want to help build the future of identity with a team that holds a high bar for itself — keep reading.

WHY SOCURE?

Socure is building the identity trust infrastructure for the digital economy — verifying 100% of good identities in real time and stopping fraud before it starts. The mission is big, the problems are complex, and the impact is felt by businesses, governments, and millions of people every day.

We hire people who want that level of responsibility. People who move fast, think critically, act like owners, and care deeply about solving customer problems with precision. If you want predictability or narrow scope, this won't be your place. If you want to help build the future of identity with a team that holds a high bar for itself — keep reading.

ABOUT THE ROLE

We are looking for a Staff Data Scientist to join Socure's Watchlist Data Science team. Watchlist sits at the heart of global AML compliance — our platform screens hundreds of millions of entities in real time across sanctions lists, PEP databases, and adverse media sources for banks, fintechs, and payment companies worldwide.

As a Staff Data Scientist, you will work on the hardest problems in entity matching and classification: scaling our patented real-time matching engine, building advanced Natural Language Processing (NLP) models for Named Entity Recognition (NER) and Information Extraction, and bringing next-generation research to production. This is a senior individual contributor role with broad technical ownership and direct impact on a product that helps the world's financial institutions manage sanctions and AML risk.

WHAT YOU'LL DO

Data Quality & Enrichment

  • Improve the quality, coverage, and freshness of Watchlist's underlying data through next-generation ingestion pipelines.

  • Design and execute rigorous data quality analysis pipelines to identify anomalies, evaluate dataset health, and ensure high-fidelity inputs for downstream model training.

  • Apply NLP and AI to classify and enrich raw source data into normalized schemas — extracting structured entity attributes from unstructured sanctions, PEP, adverse media, and enforcement sources.

  • Expand multilingual capabilities to support global screening across Latin and non-Latin scripts.

Entity Resolution

  • Build and improve NLP systems that consolidate how watchlist identities are represented. Developing Information Extraction and Named Entity Recognition (NER) pipeline to deduplicate entities across lists and resolve aliases into canonical profiles..

  • Develop approaches to handle how entity profiles change over time as names, aliases, and sanctions status evolve.

  • Measure and benchmark entity resolution quality, driving continuous improvement in coverage and accuracy.

Match Engine & Risk Scoring

  • Design and scale advanced NLP models and algorithms that perform real-time name matching and identity classification across diverse, multilingual unstructured data sources.

  • Build multi-signal risk scoring that combines name similarity, entity type, geography, list type, and other attributes into unified, calibrated risk scores.

  • Maintain and improve benchmarking frameworks, golden datasets, and regression tests that keep the match engine at the highest levels of recall and precision.

Analytics, Tuning & Evaluation

  • Build models and analytics that help customers tune their screening thresholds to the right operating point for their risk appetite and entity mix.

  • Develop backtesting and counterfactual analysis capabilities so customers and internal teams can understand how model or threshold changes would affect screening outcomes.

  • Design evaluation frameworks for AI-powered autonomous decision systems — defining correct behavior, calibrating confidence thresholds, and monitoring for drift in production.

AML Risk Detection

  • As Watchlist expands into payment screening, build the mathematical analysis and feature engineering needed to detect AML risk patterns across transaction data and payment message fields.

  • Develop and maintain the AML taxonomy and risk signal library that underlies Watchlist's classification and detection capabilities.

  • Apply graph-based methods to surface indirect risk exposure — identifying entities connected to sanctions risk even when they are not directly listed.




Research & Technical Leadership

  • Lead technical initiatives across Watchlist Data Science and shape the team's long-term approach to entity matching, enrichment, and AI.

  • Collaborate closely with Product and Engineering to translate research into production-grade systems at scale.

  • Stay current with advances in NLP, large language models, and entity resolution; prototype and deploy relevant techniques (e.g., advanced NER, LLM-based extraction) to AML use cases.

  • Mentor peers and contribute to a culture of technical rigor and continuous improvement.

WHAT YOU BRING

  • Master's or PhD in Computer Science, Computational Linguistics, Statistics, Applied Mathematics, or a related field; or equivalent professional experience.

  • 7+ years of experience in data science or machine learning, with meaningful work in NLP, entity resolution, or information extraction.

  • Experience in AML, sanctions screening, adverse media, or financial crime detection is strongly preferred.

  • Hands-on experience building and deploying NLP pipelines for entity extraction, named entity recognition, and record linkage at production scale.

  • Familiarity with multilingual NLP and non-Latin script processing is a strong plus.

  • Experience with LLMs and agentic AI frameworks (e.g., LangChain/LangGraph) is a plus.

  • Strong proficiency in Python and major ML libraries (PyTorch, spaCy, HuggingFace Transformers).

  • Strong SQL proficiency and experience with large-scale data pipelines and production ML systems.

  • Excellent communication skills — able to translate model performance tradeoffs into compliance and business language for non-technical audiences.

Note: We cannot provide Sponsorship at this time.

You must be located in one of our talent hubs: New York, San Francisco, Seattle, or Miami.

Socure is an equal opportunity employer that values diversity in all its forms within our company. We do not discriminate based on race, religion, color, national origin, gender, sexual orientation, age, marital status, veteran status, or disability status. If you need an accommodation during any stage of the application or hiring process — including interview or onboarding support — please reach out to your Socure recruiting partner directly.

Socure is an equal opportunity employer that values diversity in all its forms within our company. We do not discriminate based on race, religion, color, national origin, gender, sexual orientation, age, marital status, veteran status, or disability status.
If you need an accommodation during any stage of the application or hiring process—including interview or onboarding support—please reach out to your Socure recruiting partner directly.



Follow Us!

YouTube | LinkedIn | X (Twitter) | Facebook

Соц.сети
mayflower

Data Scientist (Search & Recommendations) и Data Scientist (Moderation)

mayflowerНа сервисе с: 24.09.26 22:02
Зарплата не указанаКипрLimassolОфис

🔹 ••••••••
🔹 ••••••••
в Mayflower — технологическая компания, создающая высоконагруженные продукты, которыми пользуются миллионы людей по всему миру.
Офис (Лимасол, Кипр). Помощь с переездом.
Ищет Степан Широбоков, его ••••••••.

HireSeeker собирает вакансии со всех площадок и присылает только релевантные. Бесплатно.