~/antoniocali
guest@antoniocali:~

        

guest@antoniocali

─────────────────────

NameAntonio Davide Calì

RoleStaff Data Engineer

HostZego

Uptime10+ years

LocationLondon, UK

LanguagesPython, Scala, Go, Java

StackFlink, Iceberg, Kafka, dbt

SpeakerFlink Forward 2025

Emailantoniodavidecali@gmail.com


      

Antonio Davide Calì

Staff Data Engineer // Streaming & Distributed Systems

10+ years building large-scale streaming and batch data platforms across fintech, retail, insurance, and consumer tech. Specialised in Apache Flink, Kafka, and Iceberg lakehouse architecture, with a recurring focus on self-service data infrastructure and GDPR/PII governance built directly into the platform. Speaker at Flink Forward 2025.

$ cat skills.yaml

languages:
  - Python
  - Scala
  - Go
  - Java
  - SQL
frameworks:
  - Apache Flink (JVM & Python)
  - Apache Kafka
  - Apache Spark
  - Apache Airflow
  - Apache Iceberg
data_platforms:
  - Databricks
  - BigQuery
  - Snowflake
  - dbt
cloud_infra:
  - AWS (ECS, S3, Glue, MSK, MSF)
  - GCP
  - Kubernetes
  - Kubebuilder
  - Terraform
  - Docker

$ git log --graph --stat experience

e28d4f1 HEAD -> main Apr 2026 — Present

Staff Data Engineer @ Zego

London, UK

Own the Data Engineering platform end-to-end — Iceberg lakehouse infrastructure and Snowflake/dbt analytics — leading the introduction of Apache Flink and Protobuf-based PII/GDPR governance across the platform.

  • Fully scoped and led the introduction of Apache Flink to Zego's data platform, driving the migration of streaming ingestion from a legacy Kinesis/Scala/Protobuf pipeline to a Flink-based forwarder.
  • Leading GDPR/PII governance by extending the company's Protobuf schema contract with field-level PII annotations, driving automatic PII derivation and masking across the data lake and Snowflake's B2B/B2C domains.
  • Own Iceberg lakehouse infrastructure (Terraform + AWS S3/IAM, Polaris catalog) and its Snowflake integration across staging and production environments.
  • Evaluated OLake (Debezium-based database-to-Iceberg replication) and a Kubernetes CRD for Polaris catalog management as ingestion PoCs, contributing upstream fixes to the open-source olake-helm Helm chart.
+Python+Scala +Protobuf+Flink +Iceberg+Snowflake +dbt+Airflow +Terraform+Kubernetes +AWS
a1f3e9c Nov 2024 — Apr 2026

Staff Data Engineer @ Just Eat Takeaway.com

London, UK

Lead the Streaming Ingestion team, building self-service tooling that enables product teams to stream and manage real-time data across a multi-cloud estate. Drive the "left-shift" initiative: a declarative DSL on top of Apache Flink that lets domain teams design, deploy, and own their own data products.

  • Architected a self-service streaming platform abstracting Flink, Kafka, and multi-cloud connectivity behind a unified DSL — cut onboarding time for new data products from weeks to days.
  • Lead schema management and data connectivity strategy across GCP and AWS, standardising contracts between producers and consumers.
  • Selected as speaker at Flink Forward 2025 to share the team's approach to productionising self-service streaming.
+Scala+Python +Java+Go +Flink+Kafka +Beam+GCP +AWS+K8s +Terraform+BigQuery
7c2b8d1 Mar 2024 — Nov 2024

Senior Data Engineer @ Dojo

London, UK

Member of the Data Streaming Platform team, expanding Apache Kafka capabilities and introducing Apache Flink and Apache Iceberg into the company's data stack.

  • Designed and shipped Kubernetes CRDs giving product teams a self-service path to ingest streaming data, abstracting Kafka/Flink config behind declarative manifests.
  • Improved platform scalability and developer experience for streaming ingestion, cutting manual operational toil for the platform team.
+Go+Java +Kubernetes+GCP +Kafka+Flink +Iceberg
4e9a017 Jun 2023 — Feb 2024

Senior Data Engineer @ Raft AI

London, UK

Led infrastructure refactoring and the rollout of a new data platform architecture across the company.

  • Designed and implemented the company's new data platform on dbt, Airbyte, dlthub, Airflow, and BigQuery, replacing a fragmented ad-hoc stack.
  • Introduced Apache Spark and Apache Kafka to the organisation; uplifted Airflow patterns and back-end SQL across teams.
+Python+Spark +dbt+Airflow +GCP+Terraform
2f6c5aa Apr 2022 — May 2023

Senior Data Engineer @ The LEGO Group

Copenhagen, Denmark

Designed and built a secure data platform for storing, processing, and analysing PII data under GDPR, then led its migration to Databricks with fine-grained access control.

  • Built end-to-end PII detection, retention, and analytics on a cloud-agnostic microservice baseline, ensuring GDPR compliance and legal sign-off.
  • Led the migration of the platform to Databricks and implemented fine-grained access control via Unity Catalog.
+Python+Scala +AWS+Databricks +Unity Catalog+Terraform
9b1d340 May 2021 — Mar 2022

Data Engineer @ The LEGO Group

Billund, Denmark

Built the next-generation LEGO Data Platform from scratch, replacing legacy architecture with a microservice-based stack.

  • Led the introduction and full implementation of Apache Airflow pipelines for the platform's orchestration layer.
  • Re-wrote the Data API in Scala and built a new REST API in Flask, modernising data access for downstream consumers.
+Python+Scala +AWS+Airflow +Terraform
6a80cf2 Jul 2019 — Apr 2021

Data Engineer @ IKEA

Malmö, Sweden

Part of the Inspirational Feed team, productionising ML algorithms that personalised the IKEA.com homepage for every visitor.

  • Led the design, productionisation, and A/B testing of the ML-powered inspirational feed serving the IKEA website.
  • Built tooling and platform features giving data scientists a reliable environment to develop and ship multiple algorithm variants.
  • Scaled the team and automated the model deployment lifecycle.
+Python+Scala +Spark+Beam +Airflow+GCP
3d47b18 Jul 2017 — Jul 2019

Data Engineer (Consultant) @ Teradata

Prague, Czech Republic

  • Delivered streaming and batch data engineering projects for enterprise clients across a broad technology surface.
+Kafka+NiFi +Spark+Scala +Groovy+Python
0000001 root-commit Feb 2015 — Jun 2017

Data Engineer @ Exprivia Telco & Media

Milan, Italy

  • Developed Java SE applications and optimised Scala/Spark jobs, improving data processing throughput and reliability.
+Java+Scala +Spark

$ ls -la ~/projects

$ cat education.json

{
  "degrees": [
    {
      "degree": "MSc, Computer Engineering",
      "institution": "University of Bologna, Italy",
      "period": "Sep 2013 – Sep 2015"
    },
    {
      "degree": "BSc, Computer Engineering",
      "institution": "University of Padua, Italy",
      "period": "Sep 2008 – Aug 2013"
    }
  ]
}

$ curl -X POST calific.io/contact

response — 200 OK
{
  "status": "open_to_conversations",
  "email": "antoniodavidecali@gmail.com",
  "github": "github.com/antoniocali",
  "linkedin": "linkedin.com/in/antoniodavidecali",
  "location": "London, UK"
}