Currently — Data Engineer at Dusens Research, building the company's first modern data platform (medallion architecture on PostgreSQL, MinIO & Metabase) for market research analytics, based in Algiers, Algeria
Previously shipped 16M+ records/day Big Data pipelines at Algeria Telecom and 50K+ events/day Kafka streaming at Djezzy — see Experience below
Learning Distributed Compute, Data Modeling (Kimball), Data Quality, Advanced SQL, Cloud (AWS/GCP)
Ask me about Hadoop, Spark, Kafka, Flink, Airflow, Superset, Podman
Reach me at zakariaalizouaoui.dev@gmail.com
Open to full-time Data Engineering roles & freelance projects
Here's my Resume
Data Engineer · Dusens Research — Nov 2025 – Present
- Designing the company's first modern data platform: medallion architecture on PostgreSQL, MinIO & Metabase for market research analytics
- Building ETL pipelines that normalize 500+ column survey exports into dimensional models across Brand Health Tracker & Mystery Shopper studies
Big Data Engineering Intern · Algeria Telecom — Feb 2025 – Jul 2025
- Orchestrated a Hadoop/Spark/PySpark/Hive platform processing 16M+ records/day (40GB) — cut query time 45min → 25min (44% faster)
- Built Airflow ELT pipelines (400+ fields → 40 features) and delivered 8 automated Superset/Power BI dashboards with Prometheus alerting
Data Engineering Intern · Djezzy — Sep 2023 – Jan 2024
- Architected a Kafka + Flink streaming pipeline processing 50K+ events/day, optimizing latency 2.8s → 2.3s (18% faster)
- Built Grafana/Prometheus dashboards (4 NPS metrics, 30s refresh) contributing to a 1.5-point CSAT improvement
Earlier: Full-Stack Development Intern @ TalabaStore (MERN delivery platform) · Software Engineering Intern @ Deltalog (Laravel + Matter.js simulations)
| Project | Description | Tech | ★ Stars |
|---|---|---|---|
| Telecom CDR Data Engineering Project | End-to-end telecom CDR Data Engineering Pipeline with Trend analysis & anomaly detection | Python · Spark · Hadoop · Hive · Airflow · Kafka | |
| Telecom NPS Data Engineering Project - Event Stream Processing | Real-time NPS analytics with Flink & Kafka | Java · Flink · Kafka · Prometheus · Grafana | |
| Machine Learning Project | Deep learning model for generating image captions | Python · TensorFlow · scikit-learn · Pandas |
Verified against the code in my own repos — not just a badge wishlist.
GitHub buckets Jupyter Notebooks as their own "language," separate from Python — since most of my data engineering and ML work lives in notebooks, the stock language chart on GitHub understates how much Python I actually write. Here's what it looks like counted by project instead of raw bytes:
| Project | Primary Language | Notes |
|---|---|---|
| cdr-telecom-bigdata-platform | Python | PySpark, Hive, Airflow DAGs, Kafka producers |
| realtime-nps-analytics | Java + Python | Flink job in Java, Kafka data generator in Python |
| Image-Captioning-Project | Python | TensorFlow, scikit-learn, pandas |
| zoomcamp-docker-terraform-workshop | Python | Dockerized data pipeline (uv-managed) |
| AI_Attrition | Python | Notebook-based ML analysis |
| cafe-zack / my-portfolio | TypeScript | Frontend/web projects |
(Real numbers from my Experience section above — not generic trophy icons.)
Thanks for stopping by — let's build something great together!
Credit: Zakaria Alizouaoui






