NoFluffJobs Praca zdalna Senior

Senior Data Platform Engineer (AWS)

Xebia sp. z o.o.

⚲ Warszawa, Rzeszów, Gdańsk, Wrocław

20 900 - 30 900 PLN (B2B)

Wymagania

  • AWS
  • Terraform
  • Apache Iceberg
  • dbt
  • CI/CD
  • Snowflake (nice to have)

Opis stanowiska

O projekcie:
Who We Are

While Xebia is a global tech company, our journey in CEE started with two Polish companies – PGS Software, known for world-class cloud and software solutions, and GetInData, a pioneer in Big Data. Today, we’re a team of 1,000+ experts delivering top-notch work across cloud, data, and software. And we’re just getting started.

What We Do

We work on projects that matter – and that make a difference. From fintech and e-commerce to aviation, logistics, media, and fashion, we help our clients build scalable platforms, data and AI solutions, and cutting-edge applications to shape the future of tech. Our clients include McLaren, Aviva, Deloitte, Spotify, Disney, ING, UPS, Tesco, Truecaller, AllSaints, Volotea, Schmitz Cargobull, Allegro, InPost, and many, many more.

We value smart tech, real ownership, and continuous growth. We use modern, open-source stacks, and we’re proud to be trusted partners of Databricks, dbt, Snowflake, Azure, GCP, and AWS. Fun fact: we were the first AWS Premier Partner in Poland!

Beyond Projects

What makes Xebia special? Our community. We support tech communities, organize meetups (Software Talks, Data Tech Talks), and have a culture that actively support your growth via Guilds, Labs, and personal development budgets — for both tech and soft skills. It’s not just a job. It’s a place to grow.

What sets us apart? 

Our mindset. Our vibe. Our people. And while that’s hard to capture in text – come visit us and see for yourself.

Wymagania:
Your profile:- strong experience with AWS data and platform services: S3, Glue (Data Catalog and Jobs), Athena, Lake Formation, IAM, VPC networking, DMS, DynamoDB, Aurora PostgreSQL,- experience with infrastructure as Code with Terraform, including modular design and YAML-driven configuration,- knowledge about Apache Iceberg and open table formats on S3,- experience with security and access control,- knowledge about CI/CD engineering with GitHub Actions, trunk-based development, and environment promotion,- solid Python and SQL skills,- experience with dbt on AWS (dbt-athena, dbt-glue adapters) and analytics-engineering workflows,- previous exposure to Airflow orchestration, ideally with Astronomer Cosmos,- data lineage and observability (OpenLineage, dbt-elementary) and data quality tooling (dbt tests, dbt-expectations) skills,- experience with CDC and streaming ingestion patterns.
Work from the European Union region and a work permit are required.
Nice to have:- exposure to BI serving layers such as QuickSight or Sigma,- snowflake experience.
Recruitment Process: CV review – HR Interview – Technical Interview - Client Interview – Decision

Codzienne zadania:
- Designing and provisioning the lakehouse foundation on AWS: S3 lake zones (raw, canonical, curated), Apache Iceberg tables, Glue Data Catalog, Athena, and Lake Formation,
- delivering all infrastructure as code with Terraform, reproducible from the client's GitHub repositories and CI/CD, across isolated environments (ci, dev, staging, prod),
- building and enforcing the tenant-isolation security gate: Lake Formation row-level security and physical partitioning by CompanyId, separate IAM roles for customer versus internal access, and fail-closed handling that quarantines and alerts on rows with missing or unresolved CompanyId,
- implementing the upstream entitlement mapping (principal to allowed CompanyIds) that drives access control,
- setting up ingestion infrastructure: CDC paths from DynamoDB Streams, AWS DMS extracts from Aurora PostgreSQL, and bulk export to Parquet, supporting the near-real-time (5 min) and batch (4 hr) SLAs,
- standing up workflow orchestration with Apache Airflow and Astronomer Cosmos, including profile-based connections, model-level retries, and lineage emission,
- building CI/CD pipelines (GitHub Actions): dbt compile and slim builds, SQLFluff linting, DAG validation, branch protection, environment promotion, and Git-revert rollback,
- configuring end-to-end governance and observability: catalog and data contracts, OpenLineage capture, and org-wide audit through CloudTrail,
- owning cost governance and monitoring: resource tagging, usage alerting, and right-sizing of compute,
- collaborating with the Data Platform Architect, Data Engineers, and Analytics Engineer, and supporting knowledge transfer to the client team.

🔍 Dekoder Ogłoszenia

🟡
We work on projects that matter – and that make a difference.
Może oznaczać projekty o dużym wpływie, ale też projekty, które są po prostu bardzo złożone i wymagające.
🔴
We value smart tech, real ownership, and continuous growth.
„Real ownership” może oznaczać dużą odpowiedzialność i samodzielność, ale też brak jasnych ram i konieczność samodzielnego rozwiązywania problemów.
🟡
It’s not just a job. It’s a place to grow.
Podkreśla rozwój, ale może też sugerować, że praca jest bardzo wymagająca i pochłaniająca, a rozwój jest priorytetem nad work-life balance.
🔴
Our mindset. Our vibe. Our people. And while that’s hard to capture in text – come visit us and see for yourself.
Jest to bardzo ogólne stwierdzenie, które może maskować brak konkretnych benefitów lub niejasną kulturę organizacyjną, którą trzeba osobiście zweryfikować.