Senior Data Engineer (AWS)
⚲ Warszawa, Rzeszów, Gdańsk, Wrocław
21 000 - 33 000 PLN (B2B)
Wymagania
- AWS
- Python
- SQL
- graph databases
- dbt
- Airflow
Opis stanowiska
O projekcie:
Who We Are
While Xebia is a global tech company, our journey in CEE started with two Polish companies – PGS Software, known for world-class cloud and software solutions, and GetInData, a pioneer in Big Data. Today, we’re a team of 1,000+ experts delivering top-notch work across cloud, data, and software. And we’re just getting started.
What We Do
We work on projects that matter – and that make a difference. From fintech and e-commerce to aviation, logistics, media, and fashion, we help our clients build scalable platforms, data and AI solutions, and cutting-edge applications to shape the future of tech. Our clients include McLaren, Aviva, Deloitte, Spotify, Disney, ING, UPS, Tesco, Truecaller, AllSaints, Volotea, Schmitz Cargobull, Allegro, InPost, and many, many more.
We value smart tech, real ownership, and continuous growth. We use modern, open-source stacks, and we’re proud to be trusted partners of Databricks, dbt, Snowflake, Azure, GCP, and AWS. Fun fact: we were the first AWS Premier Partner in Poland!
Beyond Projects
What makes Xebia special? Our community. We support tech communities, organize meetups (Software Talks, Data Tech Talks), and have a culture that actively support your growth via Guilds, Labs, and personal development budgets — for both tech and soft skills. It’s not just a job. It’s a place to grow.
What sets us apart?
Our mindset. Our vibe. Our people. And while that’s hard to capture in text – come visit us and see for yourself.
Wymagania:
Your profile:- 5+ years of experience in Data Engineering,- strong Python development skills,- advanced SQL proficiency,- hands-on experience with AWS-based data platforms (Lake Formation, Athena, Glue, DynamoDB),- experience building and maintaining data pipelines in Apache Airflow,- commercial experience with dbt,- exposure to graph databases, preferably Amazon Neptune (Neo4j, CosmosDB),- solid knowledge of Git and CI/CD practice and Terraforn,- experience with Apache Iceberg in production environments,- knowledge of openLineage or similar lineage frameworks,- experience with observability and monitoring frameworks for data platforms is a plus.
Work from the European Union region and a work permit are required.
Recruitment Process: CV review – HR Interview – Technical Interview - Client Interview – Decision
Codzienne zadania:
- designing and implementing scalable data pipelines ingesting data from DynamoDB, Aurora PostgreSQL, and Neptune,
- building and maintaining data lake layers, including Raw, Canonical, and Curated zones on Amazon S3 using Apache Iceberg,
- developing ingestion frameworks supporting both CDC and batch processing patterns,
- contributing to the implementation and ongoing maintenance of the AWS Glue Data Catalog,
- developing and maintaining Airflow workflows for orchestration of data pipelines,
- implementing CI/CD practices for data platform development and deployment,
- automating testing, validation, and deployment processes across environments,
- ensuring reliable operation of development, staging, and production environments,
- optimizing query performance, storage layouts, and Iceberg table design,
- troubleshooting and resolving production issues while continuously improving platform performance and stability.
Who We Are
While Xebia is a global tech company, our journey in CEE started with two Polish companies – PGS Software, known for world-class cloud and software solutions, and GetInData, a pioneer in Big Data. Today, we’re a team of 1,000+ experts delivering top-notch work across cloud, data, and software. And we’re just getting started.
What We Do
We work on projects that matter – and that make a difference. From fintech and e-commerce to aviation, logistics, media, and fashion, we help our clients build scalable platforms, data and AI solutions, and cutting-edge applications to shape the future of tech. Our clients include McLaren, Aviva, Deloitte, Spotify, Disney, ING, UPS, Tesco, Truecaller, AllSaints, Volotea, Schmitz Cargobull, Allegro, InPost, and many, many more.
We value smart tech, real ownership, and continuous growth. We use modern, open-source stacks, and we’re proud to be trusted partners of Databricks, dbt, Snowflake, Azure, GCP, and AWS. Fun fact: we were the first AWS Premier Partner in Poland!
Beyond Projects
What makes Xebia special? Our community. We support tech communities, organize meetups (Software Talks, Data Tech Talks), and have a culture that actively support your growth via Guilds, Labs, and personal development budgets — for both tech and soft skills. It’s not just a job. It’s a place to grow.
What sets us apart?
Our mindset. Our vibe. Our people. And while that’s hard to capture in text – come visit us and see for yourself.
Wymagania:
Your profile:- 5+ years of experience in Data Engineering,- strong Python development skills,- advanced SQL proficiency,- hands-on experience with AWS-based data platforms (Lake Formation, Athena, Glue, DynamoDB),- experience building and maintaining data pipelines in Apache Airflow,- commercial experience with dbt,- exposure to graph databases, preferably Amazon Neptune (Neo4j, CosmosDB),- solid knowledge of Git and CI/CD practice and Terraforn,- experience with Apache Iceberg in production environments,- knowledge of openLineage or similar lineage frameworks,- experience with observability and monitoring frameworks for data platforms is a plus.
Work from the European Union region and a work permit are required.
Recruitment Process: CV review – HR Interview – Technical Interview - Client Interview – Decision
Codzienne zadania:
- designing and implementing scalable data pipelines ingesting data from DynamoDB, Aurora PostgreSQL, and Neptune,
- building and maintaining data lake layers, including Raw, Canonical, and Curated zones on Amazon S3 using Apache Iceberg,
- developing ingestion frameworks supporting both CDC and batch processing patterns,
- contributing to the implementation and ongoing maintenance of the AWS Glue Data Catalog,
- developing and maintaining Airflow workflows for orchestration of data pipelines,
- implementing CI/CD practices for data platform development and deployment,
- automating testing, validation, and deployment processes across environments,
- ensuring reliable operation of development, staging, and production environments,
- optimizing query performance, storage layouts, and Iceberg table design,
- troubleshooting and resolving production issues while continuously improving platform performance and stability.
🔍 Dekoder Ogłoszenia
🔴
We work on projects that matter – and that make a difference.
Może to oznaczać projekty o dużym wpływie, ale równie dobrze może być pustym hasłem marketingowym bez konkretów.
🔴
We value smart tech, real ownership, and continuous growth.
„Real ownership” może oznaczać dużą odpowiedzialność i samodzielność, ale też brak jasnego podziału zadań i konieczność samodzielnego rozwiązywania problemów.
🔴
It’s not just a job. It’s a place to grow.
Podkreśla rozwój, ale może sugerować, że oczekuje się od pracownika zaangażowania wykraczającego poza standardowe obowiązki.
🔴
Our mindset. Our vibe. Our people.
Bardzo ogólne stwierdzenia, które mają budować atmosferę, ale nie dostarczają konkretnych informacji o kulturze pracy czy oczekiwaniach.
🔴
come visit us and see for yourself.
Zachęta do osobistej wizyty, która może być próbą pokazania pozytywów, ale też maskowania potencjalnych minusów, których nie da się opisać w tekście.