andersen

Data Engineer in Poland (Python)

Andersen is hiring a Data Engineer in Poland for a project developing large-scale analytics solutions and data platforms that transform audience and media data into actionable insights.

andersen · На сервисе с: 03.10.26 14:53

Зарплата не указанаПольшаWarsawГибрид

Andersen is hiring a Data Engineer in Poland for a project developing large-scale analytics solutions and data platforms that transform audience and media data into actionable insights.

The company is a global technology provider delivering data analytics and AI-powered solutions that help organizations better understand customer behavior, optimize marketing performance, and make informed business decisions. By transforming large volumes of data into actionable insights, it enables enterprises to improve audience engagement, measure the effectiveness of their initiatives, and enhance decision-making. The company works with a diverse range of customers, supporting them in navigating an increasingly data-driven and rapidly evolving digital landscape.

The project is focused on developing data-driven audience measurement and advertising analytics solutions powered by connected-device and media consumption data. It leverages machine learning and large-scale data processing to help media companies, advertisers, and content providers better understand viewer behavior and engagement across digital channels.

  • Serving as the primary DE point of contact for report-related questions, including data availability, metric discrepancies, and pipeline behavior.
  • Writing and executing SQL queries (BigQuery, PostgreSQL) to validate report outputs, isolate data anomalies, and confirm fixes.
  • Reproducing and documenting data issues, coordinating with the core DE team for resolution.
  • Supporting ad hoc data pulls and validation requests tied to reporting products.
  • Setting up and configuring cloud environments required for partner data delivery, including AWS S3 buckets, IAM policies, and SFTP endpoints.
  • Writing and maintaining Airflow DAGs (Python) to orchestrate ingestion workflows, including scheduling, alerting, and dependency management.
  • Coordinating with external partners on file format specifications, delivery schedules, and test dataset requests.
  • Executing match tests to validate identity resolution across ingestion pipelines.
  • Identifying and flag data quality issues in partner files before they propagate to production pipelines.
  • Experience as a Data Engineer for 3+ years.
  • Hands-on production experience on AWS, or Databricks.
  • Strong SQL (BigQuery and/or PostgreSQL) — complex analytical queries in production.
  • Python at production quality.
  • Apache Airflow DAG authoring.
  • Distributed processing in production; workflow orchestration.
  • API-first services; working experience with streaming or event-driven system.
  • Bachelor's degree in Computer Science, Software Engineering, or a related technical field, or equivalent.
  • Level of English – from Upper-Intermediate and above.
  • Background in ad tech, audience measurement, CTV/OTT, ACR, or programmatic data flows.
  • Hands-on production experience on GCP.
  • Experience in teamwork with leaders in FinTech, Healthcare, Retail, Telecom, and others. Andersen cooperates with such businesses as Samsung, Siemens, Johnson & Johnson, BNP Paribas, Ryanair, Mercedes, TUI, Verivox, Allianz, T-Systems, etc..
  • The opportunity to change the project and/or develop expertise in an interesting business domain.
  • Job conditions – you can work both fully remotely and from the office or can choose a hybrid variant.
  • Guarantee of professional, financial, and career growth! The company has introduced systems of mentoring and adaptation for each new employee.
  • The opportunity to earn up to an additional 1,000 USD per month, depending on the level of expertise, which will be included in the annual bonus, by participating in the company's activities.
  • Access to the corporate training portal, where the entire knowledge base of the company is collected and which is constantly updated.
  • Bright corporate life (parties / pizza days / PlayStation / fruits / coffee / snacks / movies).
  • Certification compensation (AWS, PMP, etc).
  • Referral program.
  • Private health insurance and compensation for sports activities.

Join us!

Похожие вакансии Data Engineer

Сайты компаний
indrive

Data Engineer

indriveНа сервисе с: 03.10.26 16:07
Зарплата не указанаCairo

About the role

We are looking for a Data Engineer to join one of the Data Platform teams that works with the Marketing, Growth, partner, and financial data domains.
You will be working with cutting edge cloud technologies (GCP, AWS, BigQuery, Databricks, K8s) and building a large scale data infrastructure for analytics, machine learning, and streaming/CDC data delivery.

Responsibilities

  • Build and operate batch and streaming ingestion into a layered BigQuery DWH (raw → ODS → data marts) using Airflow, Debezium CDC over Kafka with protobuf, Pub/Sub, and Dataflow
  • Integrate external data sources end-to-end — marketing platforms (GA4, AppsFlyer, TikTok/Meta/Google Ads), payment providers, S3 buckets, and third-party APIs — including schema contracts, backfills, and reconciliation
  • Engineer the data platform itself in Python: custom Airflow operators and connectors in a shared ETL framework, Kafka Connect on Strimzi (K8s), Cloud Functions, and API integrations with external providers
  • Build CI/CD and change-management tooling for BigQuery: GitHub-based test-and-approval flows, SQL migration engines (Liquibase/Flyway/Bytebase), sandbox validation, backup and rollback
  • Own reliability and correctness of pipelines: idempotency, deduplication, late-data handling, backfill and replay, freshness monitoring and alerting; write integration and unit tests
  • Drive data governance and compliance: ITGC-compliant change management for BigQuery, IAM and least-privilege access, PII policy tags and DLP, Unity Catalog on Databricks, column-level lineage (OpenMetadata/Dataplex), and disaster-recovery planning
  • Build internal data tools and platform services for agentic workflows with data — Streamlit apps, Slack bots, LLM-based agents and MCP servers that help teams find and use data
  • Support analysts and business teams with data requests, fostering data-driven decision-making across the company
  • Contribute to system design and architecture with the development team

Qualifications

  • Strong practical Python: clean, well-structured, and tested code for services, tooling, and data pipelines
  • Solid software design skills (OOP, modularity, design patterns) — we build platform tools for agentic workflows with data and plan to develop data-related backend services, so well-designed code is highly valued
  • Experience building and operating services in a cloud environment (GCP, AWS or similar): CI/CD, containerization, monitoring and alerting
  • Familiarity with Kubernetes and Terraform — our infrastructure runs on GCP/K8s
  • Hands-on experience with DWH-related tasks (BigQuery or another cloud warehouse) and confident working SQL
  • Clear communication with non-engineering stakeholders — a meaningful share of the work is data requests from analysts and business teams
  • Demonstrated ability to take ownership of technologies or services and proactively contribute ideas to the team
Nice to have
  • Advanced SQL: complex queries, window functions, partitioning, clustering, and cost optimization
  • Experience building reliable pipelines around CDC (e.g., Debezium): idempotency, schema evolution, backfills, and reconciliation
  • Analytical data modeling skills: table grain, facts vs dimensions, slowly changing dimensions, and metric definitions
  • Experience with stream processing frameworks such as Flink or Apache Beam/Dataflow
  • Exposure to data governance and audit compliance (ITGC/SOX), Databricks Unity Catalog, or lineage/catalog tooling (OpenMetadata, Dataplex)
  • Interest in building LLM-based agents and AI tooling for data

Benefits

  • Help us challenge injustice by creating fair choices for millions of people across 47 countries.
  • Develop your professional skills with access to mentoring, career consulting, and learning programs.
  • Collaborate with teams around the world and gain international experience through our Global Talent Exchange Program.
  • Engage in company-wide challenges, awards, sports activities, employee-led social impact and volunteering projects.
  • Work alongside people who take initiative, speak openly, and challenge themselves to grow.
  • Improve your language skills through co-financed courses and internal speaking clubs.
Final benefits may vary depending on the location.

Соц.сети
О

DataOps

ооо_7_красных_линийНа сервисе с: 02.10.26 20:22
200k–250k ₽Не указана странаЛокация не указанаУдалёнка

Публикатор: ••••••••
Обсуждение: ••••••••
#вакансия #devops #data #инженер #DWH #ETL #девопс #AI #ИИ

Должность: DataOps
Занятость: полная
Компания: ООО «7 Красных Линий»
Оклад на руки: 200-250К
Формат работы: удалённо
Оформление: ТК РФ
Заработная плата: обсуждается по итогам собеседования
Контактная информация
Telegram: ••••••••
Эл. почта: ••••••••
О компании
7 КРАСНЫХ ЛИНИЙ — AI-интегратор нового поколения. С 2016 года мы реализуем сложные IT-проекты для крупных компаний, а сегодня активно развиваем направления AI, Data и интеллектуальной автоматизации.
Чем предстоит заниматься
• Развивать и поддерживать IT-инфраструктуру, обеспечивать стабильность и отказоустойчивость систем.
• Настраивать и сопровождать процессы CI/CD, автоматизировать развёртывание и эксплуатацию сервисов.
• Работать с контейнеризацией и оркестрацией на базе Docker и Kubernetes.
• Автоматизировать инфраструктурные и эксплуатационные задачи с использованием Ansible, Bash и Python.
• Администрировать базы данных, обеспечивать резервное копирование, восстановление и мониторинг.
• Поддерживать и оптимизировать процессы загрузки, преобразования и обработки данных (ETL).
• Участвовать в развитии хранилищ данных (DWH) и витрин данных.
• Настраивать мониторинг, алертинг и диагностику инфраструктуры и сервисов.
• Анализировать и устранять проблемы производительности, загрузки и обработки данных.
• Использовать AI-инструменты и AI-агентов для автоматизации рутинных DevOps-задач и повышения эффективности работы.
Наши ожидания
• Опыт работы в DevOps и/или Data Engineering от 4 лет с практическими навыками в обоих направлениях.
• Уверенное владение Linux (Ubuntu/Debian, CentOS/RHEL).
• Опыт работы с Docker, Kubernetes и Helm.
• Практический опыт работы с Ansible и CI/CD на базе GitLab.
• Навыки автоматизации с использованием Bash и Python.
• Уверенное владение SQL и опыт работы с PostgreSQL.
• Опыт работы с одной или несколькими СУБД: PostgreSQL, ClickHouse, Greenplum.
• Понимание принципов построения хранилищ данных (DWH) и организации ETL-процессов.
• Опыт работы с Apache NiFi и/или Apache Airflow.
• Опыт настройки мониторинга и алертинга (Grafana, Prometheus или аналоги).
• Понимание сетевых технологий и протоколов TCP/IP, HTTP, DNS.
• Умение самостоятельно решать технические задачи и взаимодействовать с проектной командой.
Будет плюсом
• Практический опыт разработки ETL-процессов, хранилищ и витрин данных.
• Опыт работы с Python-библиотеками SQLAlchemy и Pandas.
• Знакомство с продуктами Arenadata (ADB, ADS, ADH).
• Опыт работы с Hadoop, PySpark или Apache Kafka.
• Опыт работы с Terraform и OpenStack.
• Опыт настройки высокодоступных и масштабируемых систем.
• Понимание принципов Data Quality и опыт автоматизации проверок качества данных.
• Практический опыт использования AI/LLM-инструментов и AI-агентов для автоматизации инженерных задач.

Сайты компаний
andersen

Data Engineer (Azure Databricks) in the EU

andersenНа сервисе с: 03.10.26 15:07
Зарплата не указанаКипрГерманияИспанияПортугалияЧехияЛюксембургЛатвияЛитваЭстонияВенгрияФинляндияФранцияАвстрияДанияИталияГибрид

Andersen is hiring a Data Engineer (Azure Databricks) in the EU for a project enhancing a cloud-based data platform and delivering scalable data solutions for drug development.

The customer is an international company operating in the life sciences sector. It focuses on developing and commercializing specialized products and solutions, supported by research, technology, and partnerships across multiple markets.

The project is focused on developing and enhancing data products for scientific and business stakeholders within the drug development domain. It aims to strengthen the cloud-based data platform and delivery capabilities, enabling scalable data processing, reliable analytics, and effective use of data across the organization.

  • Building and maintaining multi-layer pipelines (Landing → Raw → Enriched → Curated) on Azure Databricks (Spark, Delta Lake),
  • Orchestrating pipelines via Databricks Jobs / Workflows: dependencies, retries, alerting.
  • Implementing batch and API-based data ingestion from multiple source systems.
  • Applying schema enforcement and handle schema drift; run data quality checks at ingestion.
  • Working hands-on with dbt for pipeline observability and data governance tasks.
  • Monitoring job failures, data freshness, and pipeline health against SLA-based delivery targets.
  • Setting up alerts, logging, and retry strategies for production pipelines.
  • Working within a controlled Dev/Test/Prod environment: access control, audit trails, pipeline documentation, traceability to requirements.
  • Contributing to CI/CD and MLOps/AI pipelines on Databricks where the project needs it.
  • Commercial experience in Data Engineering for 4+ years.
  • Strong hands-on experience with Azure Databricks, including Spark and Delta Lake.
  • Experience building and maintaining multi-layer data pipelines following the Landing → Raw → Enriched → Curated architecture.
  • Experience with schema enforcement and schema-drift handling.
  • Ability to implement data quality checks at ingestion.
  • Experience orchestrating pipelines using Databricks Jobs / Workflows, including dependencies, retries, alerts, and failure handling.
  • Experience with both batch and API-based data ingestion.
  • Strong hands-on experience with dbt.
  • Ability to monitor job failures, data freshness, and overall pipeline health.
  • Experience delivering data in line with defined SLAs.
  • Experience implementing alerting, logging, and retry strategies.
  • Experience designing and maintaining access control models.
  • Understanding of audit trails, data lineage, and end-to-end traceability.
  • Experience with logging and evidence generation for audit and compliance purposes.
  • Understanding of segregation of duties.
  • Experience defining and maintaining data contracts.
  • Ability to create and maintain clear pipeline documentation.
  • Experience working across controlled Dev, Test, and Prod environments.
  • Ability to ensure traceability between requirements and implemented data solutions.
  • Level of English – from Upper-Intermediate and above.
  • Experience with Databricks.
  • Experience with SAP S/4HANA; Microsoft 365; Azure; SharePoint; Enterprise integrations.
  • Experience with CI/CD and DevOps practices for data, including Git-based development and branching strategies.
  • Experience building and maintaining CI/CD pipelines for data workflows.
  • Experience promoting code and data solutions across Dev, Test, and Prod environments.
  • Ability to integrate automated testing into deployment pipelines.
  • Experience with MLOps and AI pipelines on Databricks.
  • Exposure to regulated domains such as GxP, clinical trials, and pharmacovigilance.
  • Strong SQL engineering skills, including the development of modular and reusable dbt models, tests, and documentation at Analytics Engineer depth.
  • Experience with dimensional data modeling, including star schemas, One Big Table (OBT) approaches, and curated data marts.
  • Familiarity with Tabular Editor and metadata-as-code concepts.
  • Experience with Power BI dataset design, performance optimization, and maintaining KPI and semantic consistency.
  • Experience in the pharmaceutical or biotechnology industry.
  • Experience in teamwork with leaders in FinTech, Healthcare, Retail, Telecom, and others. Andersen cooperates with such businesses as Samsung, Siemens, Johnson & Johnson, BNP Paribas, Ryanair, Mercedes, TUI, Verivox, Allianz, T-Systems, etc..
  • The opportunity to change the project and/or develop expertise in an interesting business domain.
  • Job conditions – you can work both fully remotely and from the office or can choose a hybrid variant.
  • Guarantee of professional, financial, and career growth! The company has introduced systems of mentoring and adaptation for each new employee.
  • The opportunity to earn up to an additional 1,000 EUR per month, depending on the level of expertise, which will be included in the annual bonus, by participating in the company's activities.
  • Access to the corporate training portal, where the entire knowledge base of the company is collected and which is constantly updated.
  • Bright corporate life (parties / pizza days / PlayStation / fruits / coffee / snacks / movies).
  • Certification compensation (AWS, PMP, etc).
  • Referral program.
  • Private health insurance and sports compensation, depending on the type of employment.

Join us!

HireSeeker собирает вакансии со всех площадок и присылает только релевантные. Бесплатно.