Brak wyników spełniających kryteria wyszukiwania.

🟣 You will be:
Designing and provisioning the lakehouse foundation on AWS: S3 lake zones (raw, canonical, curated), Apache Iceberg tables, Glue Data Catalog, Athena, and Lake Formation,
delivering all infrastructure as code with Terraform, reproducible from the client's GitHub repositories and CI/CD, across isolated environments (ci, dev, staging, prod),
building and enforcing the tenant-isolation security gate: Lake Formation row-level security and physical partitioning by CompanyId, separate IAM roles for customer versus internal access, and fail-closed handling that quarantines and alerts on rows with missing or unresolved CompanyId,
implementing the upstream entitlement mapping (principal to allowed CompanyIds) that drives access control,
seting up ingestion infrastructure: CDC paths from DynamoDB Streams, AWS DMS extracts from Aurora PostgreSQL, and bulk export to Parquet, supporting the near-real-time (5 min) and batch (4 hr) SLAs,
standing up workflow orchestration with Apache Airflow and Astronomer Cosmos, including profile-based connections, model-level retries, and lineage emission,
building CI/CD pipelines (GitHub Actions): dbt compile and slim builds, SQLFluff linting, DAG validation, branch protection, environment promotion, and Git-revert rollback,
configuring end-to-end governance and observability: catalog and data contracts, OpenLineage capture, and org-wide audit through CloudTrail,
owning cost governance and monitoring: resource tagging, usage alerting, and right-sizing of compute,
collaborating with the Data Platform Architect, Data Engineers, and Analytics Engineer, and support knowledge transfer to the client team.
🟣 Your profile:
strong experience with AWS data and platform services: S3, Glue (Data Catalog and Jobs), Athena, Lake Formation, IAM, VPC networking, DMS, DynamoDB, Aurora PostgreSQL,
experience with infrastructure as Code with Terraform, including modular design and YAML-driven configuration,
knowledge about Apache Iceberg and open table formats on S3,
experience with security and access control,
knowledge about CI/CD engineering with GitHub Actions, trunk-based development, and environment promotion,
solid Python and SQL skills,
experience with dbt on AWS (dbt-athena, dbt-glue adapters) and analytics-engineering workflows,
previous exposure to Airflow orchestration, ideally with Astronomer Cosmos,
data lineage and observability (OpenLineage, dbt-elementary) and data quality tooling (dbt tests, dbt-expectations) skills,
experience with CDC and streaming ingestion patterns.
Work from the European Union region and a work permit are required.
Candidates must have an active VAT status in the EU VIES registry: https://ec.europa.eu/taxation_customs/vies/
🟣 Nice to have:
exposure to BI serving layers such as QuickSight or Sigma,
snowflake experience.
🟣 Recruitment Process:
CV review – HR call – Interview (with Live-coding) – Client Interview (with Live-coding) – Hiring Manager Interview – Decision
🎁 Benefits 🎁
✍ Development:
development budgets of up to 6,800 PLN,
we fund certifications e.g.: AWS, Azure,
access to Udemy, O'Reilly (formerly Safari Books Online) and more,
events and technology conferences,
technology Guilds,
internal training,
Xebia Upskill.
🩺 We take care of your health:
private medical healthcare,
multiSport card - we subsidise a MultiSport card,
mental Health Support.
🤸♂️ We are flexible:
B2B or employment contract,
contract for an indefinite period.
Zainteresowany ofertą?
Aplikuj już teraz!