S

Senior Machine Learning Engineer (FastAPI, Java)

We are looking for an experienced Machine Learning Engineer to build, deploy, and own end-to-end ML solutions that combat sophisticated fraud, deepfakes, and identity theft. This role is an exciting opportunity to have true product autonom…

sumsub · На сервисе с: 08.07.26 11:08

Зарплата не указанаНе указана странаЛокация не указанаУдалёнка

We are looking for an experienced Machine Learning Engineer to build, deploy, and own end-to-end ML solutions that combat sophisticated fraud, deepfakes, and identity theft. This role is an exciting opportunity to have true product autonomy, see your models protect millions of users in real time, and scale a platform that handles millions of verifications a day.

What You Will Be Doing:

  • Own the full ML lifecycle: from data gathering and training to optimization, production rollout, and continuous quality assessment.

  • Train SOTA CV/ML models to detect deepfakes, segment documents, and uncover complex anomaly patterns in real-time.

  • Balance research and speed: know when to build a sophisticated deep learning model and when to deploy a simple, clever heuristic to solve a problem fast.

  • Serve models efficiently using FastAPI and PyTorch, ensuring seamless integration with our high-load Java backend.

  • Build and monitor ML pipelines and dashboards using Airflow, ClickHouse, and Superset to keep our models reliable under extreme loads.

About You:

  • 4+ years of experience as an ML / CV Engineer in fast-paced, high-load product environments.

  • Strong background in Deep Learning frameworks (PyTorch is our go-to) and solid production-grade engineering skills.

  • Up-to-date with SOTA and practical baselines, specifically in image segmentation, text extraction, or anomaly detection.

  • An autonomous product owner: capable of taking an idea from inception to implementation under the pressure of a competitive market.

  • Proactive communicator: you easily align with PMs and backend teams, and you don't go quiet if something breaks - you find the solution.

Nice to have:

  • Experience with big data and stream processing tools (Kafka, Flink, Trino, or ClickHouse).

  • Hands-on experience with Generative AI (GANs, Diffusion models) or adversarial attack simulation.

  • Background in the fintech, anti-fraud, or KYC/KYB domain.

What We Offer:

  • Remote-first, trust-based culture. Work from the place that works best for you. No mandatory office days, no attendance trackers. In some locations, we provide offices or coworking spaces, but the choice is yours.

  • True flexibility. We do not fix you to a 9-to-5 schedule. You can adjust your working hours when needed, as long as your day stays productive and in sync with the team.

  • Extra time off. Your birthday is a holiday here. Add to that 10 personal days each year, seven sick days without paperwork, and extra time to enjoy Christmas and New Year. Time to rest is part of the deal.

  • Work that matters. Our mission is to build a digital world that is secure, accessible and inclusive for everyone. From fighting fraud to making online services easier and safer to use, your work will have a real impact on how people experience trust online.

  • Compensation. We offer fair and transparent pay, benchmarked to the market.

  • Truly global. We work across continents and time zones, with teammates and customers from all over the world. You will run campaigns that cross borders, cultures, and languages, and see your ideas land worldwide.

  • Growth built in. Clear goals, open feedback and personal development plans. We support your progress with learning opportunities and by covering role-specific events, from design conferences to marketing forums.

  • Team offsites. Sometimes just Slack is not enough. That is why we meet in person a few times a year. Trips are fully covered, so you can meet, collaborate, and recharge together.

  • Getting you set up. We make sure you have access to the tools and hardware you need to do your work well.

  • Friendly by design. Our logo is a dog for a reason. We keep things human, open and kind. We welcome individuality, quirks and different perspectives, because that is what makes our work smarter and more fun.

The hiring stages: TA screening -> Hiring Manager Interview -> Assignment -> Final Interview.

Sounds like a great opportunity for your career development? Then go ahead and apply! We are a global community of innovators, creators, and thinkers, and we believe that diversity fuels our innovation. Sumsub is proud to be an equal opportunity employer, committed to building a diverse and inclusive workforce. We welcome applications from people of all backgrounds, cultures, genders, experiences, abilities and perspectives. Join us in shaping the future inclusively.

We believe in fair and transparent compensation. Salaries vary depending on work location and applicable local market conditions.

This role is open to candidates in multiple countries, and compensation differs based on the candidate’s confirmed work location. For this reason, a single universal salary range is not included in this job posting. The applicable work location will be confirmed early in the recruitment process, and the relevant compensation information for that location will be shared in writing, in line with local requirements.

Похожие вакансии Data Science & ML

Соц.сети
Н

Senior ML Engineer / MLOps Engineer

Неизвестный работодательНа сервисе с: 24.09.26 16:27
Зарплата не указанаПольшаУдалёнка

#вакансия #vacancy
🚀 Senior ML Engineer / MLOps Engineer — 100% Remote
Must-have:
• Strong Python
AWS / SageMaker
MLOps, CI/CD
Docker + Kubernetes
• ML lifecycle tools: MLflow / Kubeflow / SageMaker Pipelines
• Production experience with ML & LLM applications
Generative AI / LLMOps
• PySpark / Apache Spark
• FastAPI / Flask
• Model training, fine-tuning, evaluation & optimization
📍 100% Remote | 🇬🇧 English B2+
📅 Duration: 3 months + extension
💰 Rate: TBD
⏰ Full-time
📩 Для подачи присылайте CV в telegram: ••••••••

Соц.сети
D

NLP / LLM Data Scientist

dandelionНа сервисе с: 24.09.26 19:30
Зарплата не указанаСШАУдалёнка

NLP / LLM Data Scientist
Dandelion is a product development platform for clinical AI, focused on harnessing the immense power of healthcare data to create innovations that will benefit all patients and communities.
Remote work (US-based) and flexible hours.
••••••••.
••••••••.

Другие площадки
S

Staff Data Scientist, Watchlist

socureНа сервисе с: 24.09.26 18:12
Зарплата не указанаСШАУдалёнка

Why Socure?

Socure is building the identity trust infrastructure for the digital economy — verifying 100% of good identities in real time and stopping fraud before it starts. The mission is big, the problems are complex, and the impact is felt by businesses, governments, and millions of people every day.

We hire people who want that level of responsibility. People who move fast, think critically, act like owners, and care deeply about solving customer problems with precision. If you want predictability or narrow scope, this won’t be your place. If you want to help build the future of identity with a team that holds a high bar for itself — keep reading.

WHY SOCURE?

Socure is building the identity trust infrastructure for the digital economy — verifying 100% of good identities in real time and stopping fraud before it starts. The mission is big, the problems are complex, and the impact is felt by businesses, governments, and millions of people every day.

We hire people who want that level of responsibility. People who move fast, think critically, act like owners, and care deeply about solving customer problems with precision. If you want predictability or narrow scope, this won't be your place. If you want to help build the future of identity with a team that holds a high bar for itself — keep reading.

ABOUT THE ROLE

We are looking for a Staff Data Scientist to join Socure's Watchlist Data Science team. Watchlist sits at the heart of global AML compliance — our platform screens hundreds of millions of entities in real time across sanctions lists, PEP databases, and adverse media sources for banks, fintechs, and payment companies worldwide.

As a Staff Data Scientist, you will work on the hardest problems in entity matching and classification: scaling our patented real-time matching engine, building advanced Natural Language Processing (NLP) models for Named Entity Recognition (NER) and Information Extraction, and bringing next-generation research to production. This is a senior individual contributor role with broad technical ownership and direct impact on a product that helps the world's financial institutions manage sanctions and AML risk.

WHAT YOU'LL DO

Data Quality & Enrichment

  • Improve the quality, coverage, and freshness of Watchlist's underlying data through next-generation ingestion pipelines.

  • Design and execute rigorous data quality analysis pipelines to identify anomalies, evaluate dataset health, and ensure high-fidelity inputs for downstream model training.

  • Apply NLP and AI to classify and enrich raw source data into normalized schemas — extracting structured entity attributes from unstructured sanctions, PEP, adverse media, and enforcement sources.

  • Expand multilingual capabilities to support global screening across Latin and non-Latin scripts.

Entity Resolution

  • Build and improve NLP systems that consolidate how watchlist identities are represented. Developing Information Extraction and Named Entity Recognition (NER) pipeline to deduplicate entities across lists and resolve aliases into canonical profiles..

  • Develop approaches to handle how entity profiles change over time as names, aliases, and sanctions status evolve.

  • Measure and benchmark entity resolution quality, driving continuous improvement in coverage and accuracy.

Match Engine & Risk Scoring

  • Design and scale advanced NLP models and algorithms that perform real-time name matching and identity classification across diverse, multilingual unstructured data sources.

  • Build multi-signal risk scoring that combines name similarity, entity type, geography, list type, and other attributes into unified, calibrated risk scores.

  • Maintain and improve benchmarking frameworks, golden datasets, and regression tests that keep the match engine at the highest levels of recall and precision.

Analytics, Tuning & Evaluation

  • Build models and analytics that help customers tune their screening thresholds to the right operating point for their risk appetite and entity mix.

  • Develop backtesting and counterfactual analysis capabilities so customers and internal teams can understand how model or threshold changes would affect screening outcomes.

  • Design evaluation frameworks for AI-powered autonomous decision systems — defining correct behavior, calibrating confidence thresholds, and monitoring for drift in production.

AML Risk Detection

  • As Watchlist expands into payment screening, build the mathematical analysis and feature engineering needed to detect AML risk patterns across transaction data and payment message fields.

  • Develop and maintain the AML taxonomy and risk signal library that underlies Watchlist's classification and detection capabilities.

  • Apply graph-based methods to surface indirect risk exposure — identifying entities connected to sanctions risk even when they are not directly listed.




Research & Technical Leadership

  • Lead technical initiatives across Watchlist Data Science and shape the team's long-term approach to entity matching, enrichment, and AI.

  • Collaborate closely with Product and Engineering to translate research into production-grade systems at scale.

  • Stay current with advances in NLP, large language models, and entity resolution; prototype and deploy relevant techniques (e.g., advanced NER, LLM-based extraction) to AML use cases.

  • Mentor peers and contribute to a culture of technical rigor and continuous improvement.

WHAT YOU BRING

  • Master's or PhD in Computer Science, Computational Linguistics, Statistics, Applied Mathematics, or a related field; or equivalent professional experience.

  • 7+ years of experience in data science or machine learning, with meaningful work in NLP, entity resolution, or information extraction.

  • Experience in AML, sanctions screening, adverse media, or financial crime detection is strongly preferred.

  • Hands-on experience building and deploying NLP pipelines for entity extraction, named entity recognition, and record linkage at production scale.

  • Familiarity with multilingual NLP and non-Latin script processing is a strong plus.

  • Experience with LLMs and agentic AI frameworks (e.g., LangChain/LangGraph) is a plus.

  • Strong proficiency in Python and major ML libraries (PyTorch, spaCy, HuggingFace Transformers).

  • Strong SQL proficiency and experience with large-scale data pipelines and production ML systems.

  • Excellent communication skills — able to translate model performance tradeoffs into compliance and business language for non-technical audiences.

Note: We cannot provide Sponsorship at this time.

You must be located in one of our talent hubs: New York, San Francisco, Seattle, or Miami.

Socure is an equal opportunity employer that values diversity in all its forms within our company. We do not discriminate based on race, religion, color, national origin, gender, sexual orientation, age, marital status, veteran status, or disability status. If you need an accommodation during any stage of the application or hiring process — including interview or onboarding support — please reach out to your Socure recruiting partner directly.

Socure is an equal opportunity employer that values diversity in all its forms within our company. We do not discriminate based on race, religion, color, national origin, gender, sexual orientation, age, marital status, veteran status, or disability status.
If you need an accommodation during any stage of the application or hiring process—including interview or onboarding support—please reach out to your Socure recruiting partner directly.



Follow Us!

YouTube | LinkedIn | X (Twitter) | Facebook

HireSeeker собирает вакансии со всех площадок и присылает только релевантные. Бесплатно.