профи.ру

Data Engineering Team Lead (BI Platform) (Python, TypeScript)

У нас много изменений в архитектуре и большая зона ответственности. Ищем тимлида, которому интересно разбираться в сложных системах и влиять на то, какими они станут.

профи.ру · На сервисе с: 30.09.26 11:06

400k–600k ₽РоссияМоскваГибрид

У нас много изменений в архитектуре и большая зона ответственности. Ищем тимлида, которому интересно разбираться в сложных системах и влиять на то, какими они станут.

Привет! Наш сервис помогает клиентам находить профессионалов для разных задач, а специалистам — зарабатывать на том, что они любят и умеют делать.

Ищем тимлида в команду BI. Несмотря на название, это скорее инженерная BI-платформа, чем бизнесовая команда. Задача ребят — сделать так, чтобы аналитики и бизнес получали нужные данные и инструменты быстро, стабильно и с понятным качеством.

В зоне команды много разных систем: realtime-аналитика, CDC из основной MySQL-базы, ETL и витрины данных, ClickHouse и Greenplum, калькулятор результатов A/B-экспериментов, OpenMetadata, Metabase и внутренние сервисы.

Сейчас у нас сразу несколько больших инженерных изменений: перенос realtime-аналитики на новый стандартизированный пайплайн, переход с RabbitMQ на Kafka, запуск CDC и обогащения событий в проде, развитие калькулятора A/B-тестов и системы расчёта зарплат поддержки.

В команде пять человек: три data-инженера, DWH-архитектор и BI-аналитик. Тебе предстоит быть их пипл-менеджером и отвечать за то, как команда проходит через эти изменения.

Работаем почти полностью удалённо. Регулярно встречаемся на онлайн-стендапе, иногда выбираемся куда-нибудь вместе, но обязательных офлайн-встреч нет.

Технологии

  • Python — основной язык команды. На нём пишем большую часть сервисов и внутренних инструментов, которые сами поддерживаем.
  • TypeScript — используем для сервисов, которые выходят за пределы команды, например админки и MCP-серверов.
  • SQL — работаем с ClickHouse, Greenplum и MySQL.
  • ClickHouse — realtime-аналитика, маркетинговые данные, затраты рекламных агентств и реплики витрин.
  • Greenplum — ETL и формирование витрин по нескольким доменам данных.
  • Kafka — стриминг аналитических событий и CDC.
  • Debezium — CDC из основной MySQL-базы компании.
  • Airflow 3 — оркестрация ETL-процессов.
  • OpenMetadata — метаданные объектов баз данных и бизнес-глоссарий сущностей и метрик.
  • Metabase — BI-инструмент.
  • Есть немного легаси на PHP и Java, но новые решения на них не пишем.

Зачем тебе к нам

  • Определять, какой будет BI-платформа Профи.ру дальше. Команда создаёт инфраструктуру и инструменты, которыми пользуются аналитики и бизнес. У тебя здесь большая зона технического влияния: от принципов доставки и хранения данных до того, как устроены ключевые витрины и доступ к ним.
  • Прийти в момент, когда можно заметно изменить архитектуру. Мы переводим realtime-аналитику на новый пайплайн, переезжаем с RabbitMQ на Kafka, запускаем CDC и near-realtime-обогащение событий. Многие важные решения ещё предстоит принять и довести до продакшена.
  • Работать с платформой целиком, а не с одним её слоем. В одной команде сходятся стриминг, ETL, ClickHouse и Greenplum, A/B-инфраструктура, метаданные и внутренние сервисы. Можно влиять на полный путь данных — от появления события в продукте до момента, когда аналитик или бизнес получает на его основе ответ.

Чем предстоит заниматься

  • Быть пипл-менеджером для команды: работать с тремя data-инженерами, DWH-архитектором и BI-аналитиком.
  • Провести команду через миграцию на новый пайплайн realtime-аналитики: перенести отправку событий и помочь перевести пользователей данных на новые структуры в ClickHouse.
  • Перевести стриминговую аналитику с RabbitMQ на Kafka.
  • Довести до продакшена CDC из основной MySQL-базы и near-realtime-обогащение аналитических событий.
  • Развивать калькулятор результатов A/B-экспериментов: у аналитиков накопился запрос на существенные изменения в этом инструменте.
  • Развивать систему расчёта зарплат поддержки, сделать расчёты прозрачнее и обеспечить отображение оплачиваемых событий в реальном времени.
  • Поддерживать и развивать платформу данных: ETL и витрины, ClickHouse и Greenplum, OpenMetadata, Metabase и внутренние сервисы команды.
  • Вместе с командой разбираться в логике ключевых витрин и следить за качеством и своевременностью предоставления данных.
  • Писать код вместе с командой — разработка будет занимать примерно 30–40% рабочего времени.

Что нужно, чтобы к нам присоединиться

  • Хорошая база в data engineering. В команде сильные самостоятельные инженеры, поэтому не нужно быть главным техническим экспертом. Но важно говорить с ребятами на одном языке, уверенно обсуждать пайплайны, ETL, хранилища и стриминг, вместе принимать решения.
  • Опыт проектирования и развития data-платформ. Уметь видеть систему целиком, задавать правильные вопросы, замечать риски и помогать команде разбирать сложные технические развилки.
  • Понимание realtime- и near-realtime-сценариев. В работе много стриминга, CDC и доставки событий, поэтому пригодится опыт с Kafka, ClickHouse, Debezium или похожими технологиями.
  • Умение вести технические обсуждения и принимать решения в условиях неопределённости. Здесь не всегда будет один очевидный путь: важно уметь сравнивать варианты, объяснять выбор и договариваться с сильными инженерами.
  • Готовность быть пипл-менеджером команды. В подчинении будут data-инженеры, DWH-архитектор и BI-аналитик, поэтому важно уметь держать общий фокус команды, помогать людям развиваться и выстраивать рабочее взаимодействие.

Работай с комфортом

Online&Offline

Работай, где и когда тебе удобно.

ДМС

С первого дня работы ходи в лучшие клиники России, пользуйся телемедициной.

Обучение

Оплачиваем курсы, конференции, митапы. Проводим тренинги и хакатоны.

Зарплата

Платим в рынке или выше. Пересматриваем зарплату и грейд два раза в год.

Завтраки в офисе

Каждое утро омлет, запеканка, каша, авокадо, лосось, сыр и другая вкуснятина.

Книги

Пользуйся электронной библиотекой «Альпина» или бери с книжной полки в офисе.

Скидка на Профи.ру

Компенсируем до 3450 рублей gross от стоимости заказа на Профи.ру.

Тимбилдинги

Встречайся с профийцами в любом городе. Компания компенсирует веселье :)

На выбор с компенсацией 50/50

Психотерапия

Оплатим часть стоимости сеансов с психотерапевтом.

Спорт

Частично компенсируем любой спорт: фитнес-клуб, занятия с тренером, студии.

Коворкинг

Выбирай коворкинг в любом городе, кроме Москвы (там уже есть классный офис).

Рабочее место

Оплатим нужную мебель и аксессуары для домашнего рабочего места.

Пиши, мы на связи.

Похожие вакансии Data Engineer

Соц.сети
simbirsoft

Data Engineer

simbirsoftНа сервисе с: 04.10.26 13:17
Зарплата не указанаНе указана странаЛокация не указанаУдалёнка

#senior #удаленка

SimbirSoft
Data Engineer
Формат работы: можно удалённо

☑️ Чем предстоит заниматься
-Организация пайплайнов потоков данных (конвейера движения данных в компании)
-Разработка, поддержка и оптимизация производительности Корпоративного Хранилища Данных (EDWH)
-Разработка и настройка ETL/ELT-процессов в Корпоративное Хранилище Данных (сбор, структурирование и обеспечение сохранности данных)
-Настройка инфраструктуры для обеспечения качества данных
-Разработка BI дашбордов, визуализации данных

☑️ Наши пожелания к кандидатам
-Обязателен опыт работы с технологиями: Greenplum или другая MPP СУБД, ClickHouse, Airflow
-Хорошее знание SQL и реляционных баз данных (желательно, Greenplum или PostgreSQL), опыт написания сложных запросов
-Хорошее знание Python
-Знание современных технологий обработки больших данных
-Высшее образование (IT/математика/математическая статистика/физика)

Контакты: •••••••• / Telegram ••••••••

🔥 •••••••• / •••••••• / ••••••••

Сайты компаний
A

Data Analytics Engineer

AvrideНа сервисе с: 03.10.26 15:30↑ Вакансия с автоподнятием
Зарплата не указанаСШАAustinУдалёнка

About the Team

Here at Avride, we're building the future with self-driving vehicles and delivery robots. As you can imagine, this creates a massive amount of data, and we need someone to help us connect the dots. This is a great opportunity to join a growing team and have a real, tangible impact on our technology and our success.

About the Role

We’re looking for a Data Analytics Engineer who is skilled in both data engineering and data analysis. Your main goal will be to take the messy, raw data coming from our vehicles, APIs, and databases, and transform it into clean, reliable datasets that everyone from our engineers to our CEO can actually use. You'll be responsible for building the data pipelines that make this happen, and the dashboards that bring the insights to life.

What You'll Do

  • You'll tame our raw data streams, building the ETL pipelines that pull information from all corners of the company into our data warehouse.
  • You'll work with some truly unique datasets—everything from vehicle sensor telemetry and logistics data to the results of our internal system tests.
  • You'll write the smart, efficient SQL queries needed to shape our data and prepare it for analysis.
  • You'll create and manage the Grafana dashboards that our teams rely on to track performance and spot issues.
  • You'll work closely with other teams to figure out what data they need, and then you'll deliver it.
  • You'll also get to experiment with our internal AI and LLM-based tools to find new ways to analyze data and automate insights.

What You'll Need

  • Strong, practical experience with Python and its data libraries (like Pandas, Polars, etc.).
  • Expert-level SQL. You should be very comfortable with complex joins, window functions, and query optimization.
  • A solid background in building and maintaining ETL pipelines using modern tools.
  • Hands-on experience with a BI tool like Grafana, Tableau, or Looker.
  • Familiarity with workflow orchestrators like Airflow, Dagster, or Prefect.
  • A good high-level understanding of how Large Language Models (LLMs) work and an interest in applying them to data problems.
  • A strong sense of ownership and a passion for making sure the data is right.

Nice to Have

  • You've worked with ClickHouse or other modern analytical databases (Snowflake, BigQuery, Redshift).
  • You have experience with vehicle, sensor, or logistics data.
  • You've worked in the autonomous vehicle or robotics industry before.

#LI-MS1

Candidates are required to be authorized to work in the U.S. The employer is not offering relocation, sponsorship, and remote work options are not available.

Avride is an equal opportunity employer and committed to providing reasonable accommodations to qualified applicants and employees with disabilities to ensure they have equal access to employment opportunities. Avride complies with the Americans with Disabilities Act (ADA), if you need a reasonable accommodation to assist with the application or hiring process, or to perform the essential functions of a job, please email jobs@avride.ai.

Сайты компаний
V

Director, Graph Databases

veeamНа сервисе с: 03.10.26 21:43
Зарплата не указанаКанадаSan Jose

Veeam is the Data and AI Trust Company, specializing in helping organizations ensure their data and AI are fully understood, secured, and resilient to enable the acceleration of safe AI at scale. As the market leader in both data resilience and data security posture management, Veeam is built for the convergence of identity, data, security, and AI risk. Headquartered in Seattle with offices in more than 30 countries, Veeam protects over 550,000 customers worldwide, who trust Veeam to keep their businesses running. Join us as we go fearlessly forward together, growing, learning, and making a real impact for some of the world’s biggest brands.

About the Role

You’ll lead two core systems inside Veeam Data Command Center: the Knowledge Graph and the hyperscale data lake integrations. Together, they help customers understand where sensitive data lives, who can access it, how it moves, and whether AI models trained on it can be trusted. 

You’ll lead multiple teams building a searchable, security-aware graph that works at enterprise scale. This role is for a hands-on technical leader who can set clear direction, grow strong teams, and deliver reliable systems—while building an AI-first engineering culture with high standards for quality and security. 

What You’ll Do

  • Set the technical vision and end-to-end architecture for the Knowledge Graph, including the data model, storage engine, and query layer at very large scale 
  • Guide the evolution of the graph schema for data sources, identities, access, classifications, and lineage (property graph and/or RDF) using Amazon Neptune and/or Neo4j 
  • Own the strategy for hyperscale lake and lakehouse integrations, including connectors and scanning engines that ingest metadata and lineage from Delta Lake, Iceberg, Parquet/Avro, and platforms like Azure Data Lake, AWS S3/Glue, and BigQuery without disrupting production 
  • Drive performance and reliability, including standards for indexing, partitioning, and query planning, and tuning traversals and queries (Gremlin, Cypher/openCypher, SPARQL) 
  • Build and scale an AI-first engineering approach where teams use tools like Claude Code, Cursor, and Copilot responsibly, with guardrails for security, maintainability, and code quality 
  • Invest in reusable engineering building blocks (including “Claude skills” and agent workflows) that make teams faster and more consistent 
  • Own delivery outcomes: roadmap execution, operational readiness, incident learning, and cross-team alignment 
  • Hire, coach, and develop leaders, including engineering managers and senior/staff engineers, with clear expectations and growth paths 

What You’ll Bring

  • 10+ years of software engineering experience in data infrastructure, graph systems, or distributed data platforms 
  • 4+ years of engineering leadership experience, including leading through managers and scaling multiple teams 
  • Strong production experience with Amazon Neptune and/or Neo4j, including scaling, operations, and trade-offs (property graph vs. RDF) 
  • Proven ability to lead graph modeling for complex domains, including lineage and permissions at enterprise scale 
  • Deep knowledge of Gremlin, Cypher/openCypher, and/or SPARQL, including performance tuning and query design best practices 
  • Experience with data lakes/lakehouses (Delta Lake, Iceberg, Parquet) across major cloud platforms (Azure Data Lake, AWS S3/Glue, BigQuery) 
  • Experience designing and operating distributed systems using tools like Spark, Flink, or Presto/Trino, with strong judgement on scalability and cost 
  • Strong backend background in Go and/or Python, with the ability to review designs, guide decisions, and unblock teams 
  • Practical experience using AI-assisted development tools and the ability to set standards that keep AI-assisted code secure and high quality 

Bonus Skills

  • Experience operating graph systems at massive scale 
  • Background in data security, access governance, and policy controls 
  • Experience with AI/ML governance tools and practices (e.g., MLflow, Databricks Mosaic AI) 
  • Experience building custom agents, MCP-based workflows, or reusable engineering automation 
  • Infrastructure-as-Code experience (e.g., Terraform or Pulumi) 
  • Contributions to graph standards or communities (GQL, openCypher, SPARQL) 

 

What you'll get

  • Unlimited paid time off, 12 paid holidays including 4 global VeeaMe Days for self-care and 24 paid volunteer hours annually through Veeam Cares
  • Paid parental leave: 8 weeks for all parents, 16 weeks for birthing parents
  • Medical, dental, and vision coverage starting on your first day
  • Mental health support, therapy sessions, and digital wellness tools via our Employee Assistance Program
  • 401(k) retirement plan with company matching contributions
  • Fertility, adoption, and surrogacy support through Maven, plus paid volunteer time
  • AirVet: 24/7 virtual veterinary care at no cost
  • Legal services, identity protection, and supplemental health insurance options
  • Tax-advantaged spending accounts for healthcare, dependent care, and commuting
  • Opportunities to learn and grow through on-demand libraries (LinkedIn Learning, O’Reilly), mentoring, workshops, and learning events like our annual Global Day of Learning

Pay Transparency

Veeam is committed to pay transparency and equitable compensation. For this role, the compensation range below reflects the expected total target compensation (TTC), inclusive of base pay and a competitive performance-based bonus. For roles with a commission plan, the compensation range represents On Target Earnings (OTE), which includes base salary plus variable commission. When determining compensation, Veeam takes into consideration factors such as experience, education, skills, and geographic zone. Offers are typically made below the midpoint of the range.

In addition to compensation, Veeam provides a comprehensive benefits package, including health coverage, retirement plans, and unlimited time off.

Compensation Range (TTC / OTE)
$382,560—$710,400 USD

Veeam Software is an equal opportunity employer and does not tolerate discrimination in any form on the basis of race, color, religion, gender, age, national origin, citizenship, disability, veteran status or any other classification protected by federal, state or local law. All your information will be kept confidential.

Personal data collected during the recruitment process will be processed in accordance with our Recruiting Privacy Notice, which explains how your information is collected, used, and handled in connection with hiring activities. By applying for this position, you consent to this processing. 

By submitting your application, you confirm that the information provided, including any supporting documents, is complete and accurate to the best of your knowledge. Any misrepresentation, omission, or falsification may result in disqualification from consideration or, if discovered after employment begins, termination of employment.

HireSeeker собирает вакансии со всех площадок и присылает только релевантные. Бесплатно.