N

Data Engineering Team Lead (Agentic Search) (Python)

Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without t…

nebius · На сервисе с: 24.06.26 13:46

Зарплата не указанаНе указана странаЛокация не указанаУдалёнка

About Nebius:


Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-house AI/ML infrastructure.


Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across compute, storage, networking and applied AI.


Listed on Nasdaq (NBIS) and headquartered in Amsterdam, we have a global footprint with R&D hubs across Europe, the UK, North America and Israel. Our team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software and AI R&D.


The Product:


In a rapidly evolving world, trust in AI depends on AI agents being grounded in fresh, verified real-world data. Search is the foundation that makes this possible.


We are building an agent-native search platform designed specifically for AI systems rather than human users. Our product provides programmatic, low-latency, and observable search APIs that AI agents use to retrieve, filter, and reason over real-world information at scale.


Behind every search request is a rich stream of signals - query patterns, retrieval decisions, crawling outcomes, ranking quality, usage and revenue events. Turning that stream into a trustworthy, queryable data platform is what makes the product improvable, the business measurable, and the models trainable.


The Role:


We are looking for a Data Engineering Team Lead to lead our data platform - both the data behind our search quality and ML pipelines, and the analytics that drive product and business decisions.


In this role, you will lead a team of data engineers and own the end-to-end data lifecycle: ingestion from production services, helping model and architect our data warehouse, and exposing clean, well-documented data to researchers, engineers, and analysts across the company.


The platform spans tens of terabytes and ingests from tens of proprietary and third-party sources - our own search engine and its components, CRM, billing, identity, and product analytics across multi-region production environments. Around 100 internal users rely on it daily.


You will lead our data platform, hire and grow the team, and stay hands-on enough to design and review the systems your team ships.


In this position, your responsibility will be to:


  • Lead and architect Tavily's data platform - from real-time ingestion through data warehouse medallion layers to consumer-facing datasets and dashboards
  • Lead, hire, mentor, and grow a team of data engineers; set engineering standards for code quality, testing, documentation, and on-call
  • Work closely with engineers across the company to make sure batch and streaming pipelines are done correctly
  • Define and implement observability for the data platform: data quality checks, freshness monitors, lineage, schema evolution, and cost controls
  • Partner with researchers, engineers , analysts, finance, and product managers to deliver trustworthy datasets for product & gtm analytics.
  • Define the objects, entities, and relationships that model Tavily's search domain - agent inputs, URLs, chunks, agent sessions, crawls, and the connections between them - and turn that mental model into a clean, queryable data model that the rest of the company can reason about.
  • Data Governance: Can ensure the highest standards of data quality, integrity, and security across all environments.


You may be a good fit if you:


  • 5+ years of Data Engineering experience, with a focus on designing and implementing scalable, analytics-ready data models and cloud data warehouses (e.g., BigQuery, Snowflake).
  • Have hands-on experience with Snowflake (or a comparable cloud data warehouse) and a strong grasp of data warehouse architecture preferably medallion schema.
  • Deep knowledge of databases (schema design, query optimization) and familiarity with NoSQL use cases.
  • Expertise in modern data orchestration and transformation frameworks (e.g., Airflow, DBT).
  • Solid understanding of cloud data services (e.g., AWS, GCP) and streaming platforms (e.g., Kafka, Pub/Sub).
  • Have hands-on experience with the Spark / MapReduce paradigm and understand when distributed processing is the right tool
  • Are fluent in Python and SQL for production data work
  • Have operated data systems in production: debugged them under pressure, recovered from data incidents, and understand what it means to backfill a corrupted table.


Benefits & Perks:


  • Competitive compensation
  • Career growth and learning opportunities
  • Flexibility and work-life balance
  • Collaborative and innovative culture
  • Opportunity to work on impactful AI projects
  • International environment and talented teams


What's it like to work at Nebius:


Fast moving - Bold thinking - Constant growth - Meaningful impact - Trust and real ownership - Opportunity to shape the future of AI


Equal Opportunity Statement:


Nebius is an equal opportunity employer. We are committed to fostering an inclusive and diverse workplace and to providing equal employment opportunities in all aspects of employment. We do not discriminate on the basis of race, color, religion, sex (including pregnancy), national origin, ancestry, age, disability, genetic information, marital status, veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by applicable law.


Applicants must be authorized to work in the country in which they apply and will be required to provide proof of employment eligibility as a condition of hire.


If you need accommodations during the application process, please let us know.

Похожие вакансии Data Engineer

Соц.сети
D

Senior Data Engineer (Fintech Data)

delivery_heroНа сервисе с: 24.09.26 18:12
Зарплата не указанаГерманияBerlinГибрид

As the world’s pioneering local delivery platform, our mission is to deliver an amazing experience, fast, easy, and to your door. We operate in around 65 countries worldwide, powered by tech, designed by people. As one of Europe’s largest tech platforms, headquartered in Berlin, Germany, Delivery Hero has been listed on the Frankfurt Stock Exchange since 2017 and is part of the MDAX stock market index. We push hard, learn quickly and stay human along the way. It’s this wonderful mix of high performance and real community that makes Delivery Hero a place where ambition and belonging grow side by side. If you’re curious, collaborative and ready to dive deep into meaningful work, you’ll fit right in.

Job Description

We are on the lookout for a Senior Data Engineer (Fintech) to help build the next-generation data foundation for fraud prevention.

Be part of building the financial backbone of Delivery Hero. You’ll develop products that empower millions of customers and merchants, from seamless payments to innovative financial solutions like wallets and credit. Your work will support our path to profitability by creating financial flexibility for users and enabling smooth transactions across our markets.

Our Flink-based streaming platform already powers real-time fraud signals. In this role, you will build and enhance historization capabilities, batch pipelines, data quality framework, and observability tooling to make fraud data reliable, traceable, and easy to use for rule evaluation, model training, inference, and backtesting.

You will work closely with Data Science and Engineering teams and own the full lifecycle of these data solutions, from design and development to operation and continuous improvement.

  • Build accurate, point-in-time-correct datasets for model training, backtesting, and analysis.
  • Design and build scalable batch pipelines that complement our existing Flink-based streaming platform.
  • Develop robust controls for data completeness, freshness, anomaly detection, reconciliation, and business-logic accuracy.
  • Build observability capabilities that surface pipeline failures, data drift, and quality issues before they affect downstream models or fraud rules.
  • Partner with Data Scientists and Fraud Ops to translate fraud-prevention use cases into scalable, resilient data products while maintaining consistency between batch and real-time processing.

Qualifications

Cloud Data Engineering: experience building production data pipelines using Python, Airflow, dbt, and BigQuery on GCP or AWS.

  • Stream Processing: hands-on experience with Apache Flink or a similar framework, including event time, state, checkpointing, and late-arriving data.
  • Data Science Enablement: a strong understanding of feature preparation, model training, backtesting, and inference workflows.
  • Data Quality and Observability: experience building validation frameworks, monitoring, alerting, and data lineage for production systems.
  • Engineering Ownership: a track record of owning data solutions from design and implementation through deployment and operations.

Additional Information

Ensuring you and all our Heroes are looked after, happy, and healthy is always on the menu. Because if you’re in good shape, then we’re in good shape.

  • Make the most of our hybrid working model and join the team for face-to-face connection and collaboration in our beautiful Berlin campus 2 days a week
  • We offer 27 days holiday with an extra day on 2nd and 3rd year of service
  • We will support you in developing yourself and your career growth opportunities: 1.000 € Educational Budget, Language Courses, Parental Support and access to the Udemy Business platform to explore a variety of online courses.
  • Get moving and release those wonderful, mind-boosting endorphins: Health Checkups, Meditation, Gym Subsidy.
  • Cash. Dough. Cheddar. Whatever you call it, we’ll help you with it: Employee Share Purchase Plan, Sabbatical Bank, Public Transportation Ticket Discount, Life & Accident Insurance, Corporate Pension Plan
  • The power of getting together over some food is unrivaled. Here are a few ways to help you do that. All the yum: Digital Meal Vouchers, Food Vouchers, Corporate Discounts. Courses.
  • Wondering what relocating to Berlin is like? In this article, we’ve put together 10 things you should know about moving to Berlin and how Delivery Hero can support you. You can also visit our relocation hub and check out more information about moving to Berlin.
  • Ready to prepare for your interview? Check out the list of the 5 most common interview questions and answers created in collaboration with our recruiters.

Ready to join our team? If you’re excited to grow, collaborate and be part of the world’s leading delivery platform, we’d love to hear from you. Apply today!

We believe diversity and inclusion are key to creating not only an exciting product, but also an amazing customer and employee experience. Fostering this starts with hiring - therefore we do not discriminate on the basis of racial identities, religious beliefs, color, national origin, gender identities or expressions, sexual orientations, age, marital or disability statuses, or any other aspect that makes you, you.

We encourage you to let us know if you need any accommodations or specific accessibility support to ensure a smooth interview experience—just let us know with an email to our Inclusion Officer at inclusion@deliveryhero.com.

Severely disabled applicants with equal qualifications will be given preferential consideration.

You're welcome to share your pronouns (he/she/they) right from the start so we can address you respectfully from our first contact.

hh.ru
рсхб-интех

Middle/Middle+ разработчик ETL/ELT (Антифрод)

рсхб-интехНа сервисе с: 24.09.26 16:57
Зарплата не указанаРоссияМоскваГибрид

Проект реализация DWH и BI в области противодействия мошенническим операциям.

Цель проекта: обеспечение целевых подразделений антифрода инструментами аналитики и подготовки принятия решений для задач предотвращения мошенничества.

ЧТО ВАС ЖДЕТ:

  • Высокая автономность: вы будете исследовать и развивать сквозной процесс от множественных источников данных до форм представлений непосредственным пользователям.
  • Разнообразие данных: для задач домена используются данные практически всех направлений внутри банка и много внешних специфичных данных.
  • Передовой технологический стек и нетривиальные задачи: Графовая аналитика и Graf RAG; Feature Store с возможностью его переиспользования для in-line задач, интеграции результатов AI в бизнес аналитику и так далее.
  • Влияние и бизнес-импакт: у нас высокий спрос на качественные данные как в аналитических задачах, так и в области машинного обучения.

СТЕК:

  • Greenplum (GPDB)
  • Data Vault 2.0
  • Data Warehouse
  • Data Architecture
  • PowerDesigner
  • SQL
  • Data Modeling
  • Banking Data Model
  • General Ledger
  • Data Governance
  • MPP Database
  • ETL Architecture

ЧЕМ ПРЕДСТОИТ ЗАНИМАТЬСЯ:

  • Принимать участие в проработке архитектуры ETL и интеграций аналитического хранилища данных (DWH);
  • Проектировать и обеспечивать реализацию потоков ETL в соответствии с выбранной методологией;
  • Управлять физической моделью данных корпоративного хранилища;
  • Формировать стандарты в области данных и контролировать их соблюдение, обеспечивать историзацию и правила качества данных;
  • Взаимодействовать с владельцами данных, бизнес и системными аналитиками;
  • Оптимизировать архитектуру хранилища и обеспечивать ее масштабируемость;
  • Готовить и согласовывать обязательную документацию.

НАШИ ОЖИДАНИЯ ОТ КАНДИДАТА:

  • Опыт работы главным/ведущим инженером ETL/ELT от 3 лет;
  • Глубокое понимание архитектуры корпоративных хранилищ данных;
  • Практический опыт проектирования хранилищ с использованием Greenplum (GPDB), Postgres, Hadoop;
  • Понимание архитектуры Greenplum: распределение, партиционирование, планы выполнения, оптимизация хранения и обработки больших объемов данных;
  • Результативный опыт анализа проектных решений, моделей, ETL/ELT-процессов, BI-витрин и технической документации;
  • Практический опыт работы с отечественными BI решениями;
  • Отличное знание SQL и принципов оптимизации запросов в MPP-системах;
  • Экспертиза и применение архитектурных паттернов загрузки, инкрементальности, CDC, дедупликации, историчности, отслеживаемости и разграничения доступа;
  • Опыт в проектировании и применение Spark, Impala, Trino, Airflow, Java/Python в промышленном ландшафте данных;
  • Понимание принципов CI/CD, работа с системами контроля версий;
  • Уверенное владение инструментами проектирования:
    - разработка физических моделей;
    - управление репозиторием моделей;
    - поддержка корпоративных стандартов моделирования.

БУДЕТ ПЛЮСОМ:

  • Опыт работы с данными в антифрод-подразделении банка или фин. организации;
  • Опыт проектирования CDM/LDM/PDM в крупных банках;
  • Опыт построения Data Lake / Lakehouse совместно с DWH;
  • Знание подходов Data Governance и Data Quality;
  • Понимание принципов MDM;
  • Опыт работы с StarRocks;
  • Опыт работы с PXF;
  • Знание Agile-подходов к разработке;
  • Экспертиза в банковской предметной области. Вы должны хорошо ориентироваться в банковских данных и понимать структуру основных предметных областей.

ЧТО МЫ ПРЕДЛАГАЕМ:

  • Обучение за счет компании (посещение конференций, курсов, помощь в написании статей на Хабр и т.д.);
  • Вертикальное и горизонтальное развитие: регулярные тренинги, вебинары, митапы;
  • Забота о вашем здоровье: ДМС с первого месяца работы, куда входит стоматология;
  • Прозрачный доход: оклад (по итогам интервью) + ежеквартальные премии по результатам KPI;
  • Гибридный формат работы (2 дня офис, 3 дня удаленно) в Сколково;
  • Дополнительные бонусы от Россельхозбанка для сотрудников группы компаний (Скидки на спортзалы, рестораны, маркетплейсы и т.д.).
Соц.сети
Т

Middle Data Engineer

телекомНа сервисе с: 24.09.26 15:56
Зарплата не указанаРоссияУдалёнка

ID 3659 - Middle Data Engineer

🌍 Локация: РФ
💼 Удаленно
🕔 Занятость: фулл тайм

🏢 Проект: Телеком

💡 Требования:
• Опыт построения ETL-пайплайнов от 3 лет
• Знание Python, SQL
• Опыт работы с озёрами данных и витринами данных
• Понимание качества данных и валидации
• Опыт создания промышленных витрин на Hadoop, PostreSQL
• Готовность прохождения очного интервью с решением задач на Python и SQL на темы: создание ETL-процессов; оркестрация ETL-процессов ; подключение к интерфейсам, для захвата/передачи данных; оптимизация запроса к СУБД; использование очередей; использование LLM

📨 Отклик — через форму: •••••••• или напрямую рекрутеру ••••••••

❗️Откликайтесь только при релевантном опыте.

❗️При первичном отклике:
ID вакансии / ФИО / локация / возраст / занятость (работаете/нет) / формат работы (удаленка, гибрид, офис) / стек / опыт / резюме / сверка с требованиями

❗️Повторный отклик: ID вакансии + сверка.

#Data #Engineer #Удаленно #вакансия

HireSeeker собирает вакансии со всех площадок и присылает только релевантные. Бесплатно.