{"$schema":"https://raw.githubusercontent.com/jsonresume/resume-schema/v1.0.0/schema.json","basics":{"name":"Hamza Ben Marzouk","label":"Data & AI Engineer","email":"contact@hamzabenmarzouk.com","url":"https://hamzabenmarzouk.com","summary":"Data & AI Engineer with 7+ years of experience, including 5 in consulting (OCTO Technology), across demanding industries: banking, insurance, energy, media and cloud. I design and harden data platforms (Snowflake, dbt, Databricks, Spark) and put AI into production: agents, LLM evaluation, adoption by business teams. I bridge business and engineering, upskill the teams I work with, and leave behind systems that are documented, tested and maintainable.","location":{"city":"Paris","countryCode":"FR"},"profiles":[{"network":"LinkedIn","username":"bmhamza","url":"https://www.linkedin.com/in/bmhamza/"}]},"availability":{"open":false,"text":"Available for freelance engagements"},"offers":[{"name":"Modern data platforms","summary":"Design, take over or modernise a Snowflake or Databricks platform: ingestion, dbt modelling, orchestration, infrastructure as code and CI/CD."},{"name":"Reliability, security & governance","summary":"Data quality, observability, data contracts, access management and GDPR compliance: make the platform trustworthy and auditable."},{"name":"AI in production","summary":"AI agents, LLM evaluation, chatbots and AI adoption by business teams: move from proof of concept to measured production use."}],"work":[{"name":"Wakam","position":"Data Platform Engineer","url":"https://www.wakam.com","location":"Paris","startDate":"2026-04","summary":"Back on the Data Platform to industrialise the platform's security, governance and reliability.","highlights":["Designed a GitOps system for temporary Snowflake access: request by pull request, approval, automatic revocation and audit log, including for sensitive (GDPR) data","Moved Snowflake configuration to infrastructure as code (Terraform): network policies, warehouses, service accounts, security tasks; addressed CIS benchmark recommendations","Built a regulatory data mirror (UK data residency) to Azure Blob Storage, with an independent daily integrity check","Hardened Prefect orchestration: dev/prod isolation, late-run monitoring, Slack alerting","Migrated CI/CD from Azure DevOps to GitHub Actions, with ephemeral review environments per pull request, a SonarCloud quality gate and a Python 3.13 / uv migration","Agentic tooling for the team: Claude Code skills automating PRs, tickets, diagnostics and routine operations"],"keywords":["Snowflake","Terraform","Prefect","dbt","Azure","GitHub Actions","Python","Claude Code"]},{"name":"Wakam","position":"AI Engineer — KamAI team","url":"https://www.wakam.com","location":"Paris","startDate":"2025-08","endDate":"2026-03","summary":"Built and deployed AI solutions with a dual goal: industrialise AI agent evaluation and maximise adoption of AI tools across the company.","highlights":["Designed and deployed an end-to-end evaluation system for business-critical AI agents","Automated generation of synthetic evaluation datasets, validated by business experts in a Retool app","Automated evaluation pipelines with centralised result tracking in Langfuse","Shipped a conversational chatbot on wakam.com for policyholders (agentic workflows on Dify)","Drove adoption of the Dust platform; ran bi-monthly hackathons (average satisfaction ≥ 4/5) and coached business teams on high-value use cases"],"keywords":["Dust","Dify","Langfuse","Retool","LLMs","Python"]},{"name":"Wakam","position":"Data Engineer — Data Platform","url":"https://www.wakam.com","location":"Paris","startDate":"2024-02","endDate":"2025-07","summary":"Two strategic workstreams, PDX (Partner Data eXchange) and DPF (Data Platform Foundation), ensuring the data platform's reliability, scalability and operability.","highlights":["Technical lead on data contracts: brought 3 partner teams to autonomy in defining and applying the standards","Delivered a hardened end-to-end pipeline (partner exchange → Snowflake), with documentation, knowledge transfer and formalised processes","Optimised the monthly run: 50% of data support requests handled self-service","Full observability with Datadog (alerting, dashboards); rebuilt the data quality framework and introduced Elementary","Designed ETL / reverse ETL pipelines and restructured the dbt codebase by domain; significantly reduced data incidents"],"keywords":["Snowflake","dbt","dlt","Databricks","Elementary","Datadog","Python"]},{"name":"Scaleway","via":"OCTO Technology","position":"Data Ops Consultant","location":"Paris","startDate":"2023-09","endDate":"2024-01","summary":"Launch of Scaleway's first data product: a managed Spark offering (Spark as a Service).","highlights":["Validated technical prerequisites: scalability, connectivity, resilience","Implemented and validated the first use cases (proof of concept)"],"keywords":["Apache Spark","Spark as a Service","POC"]},{"name":"Mobilize Financial Services (Renault Group)","via":"OCTO Technology","position":"Backend Developer","location":"Paris","startDate":"2022-10","endDate":"2023-02","summary":"Creation of the Mobilize Pay neobank, alongside Accenture: built and shipped the mobile banking app.","highlights":["Implemented banking features (layered architecture inspired by clean architecture)","Upheld engineering best practices and drove the team's continuous improvement"],"keywords":["GCP","Java","Spring Boot","Flutter","PostgreSQL","GitLab CI"]},{"name":"RelevanC","via":"OCTO Technology","position":"Data Engineer","location":"Paris","startDate":"2022-01","endDate":"2022-05","summary":"Data lake overhaul (optimisation and remediation).","highlights":["PySpark data cleansing and processing pipelines; CI/CD and infrastructure as code","Upskilled the team through pair and mob programming and code reviews"],"keywords":["GCP","BigQuery","Apache Spark","Airflow","Dataproc","Terraform"]},{"name":"SACEM","via":"OCTO Technology","position":"Tech Lead","location":"Paris","startDate":"2021-04","endDate":"2021-12","summary":"Automated reconciliation of music rights contract updates submitted by publishers.","highlights":["Designed ingestion, enrichment and business-rule workflows, plus process tracking metrics","Technical leadership: team facilitation, pair programming, code reviews, architecture design with the architects"],"keywords":["AWS","Java","Python","PostgreSQL","Lambda","ECS","Terraform"]},{"name":"Engie Digital","via":"OCTO Technology","position":"Backend Developer","location":"Paris","startDate":"2019-09","endDate":"2021-03","summary":"Livin' smart city platform: real-time air quality, street lighting, traffic.","highlights":["Designed and built product features, bringing data expertise on ingestion and processing","Drove DevOps culture and Accelerate practices; ran event storming workshops"],"keywords":["AWS","Java","Spring Boot","PostgreSQL","MongoDB","Apache Spark","CQRS"]},{"name":"BNP Paribas BDDF","via":"OCTO Technology","position":"Data Engineer","location":"Paris","startDate":"2019-02","endDate":"2019-09","summary":"Rebuilt client file processing pipelines: migration from a DB2 mainframe to a Big Data stack.","highlights":["Spark aggregation and transformation jobs, orchestrated with Oozie","Set up and maintained the CI/CD pipeline"],"keywords":["Scala","Apache Spark","HDFS","Hive","Oozie","Jenkins"]}],"projects":[{"name":"Code & IT architecture audit","entity":"Banque de France","type":"audit","description":"Following service degradations: mapped and audited the code of the corporate credit-rating components, with remediation recommendations.","keywords":["Distributed architecture","Coupling","Resilience"]},{"name":"Document management data model & data use cases","entity":"AXA France","type":"audit","description":"Studied the document management system's data model during its overhaul; recommendations to optimise its use, then extract value from its data.","keywords":["Unified model","Service contracts","Governance"]},{"name":"Vendor due diligence","entity":"Argos","type":"audit","description":"Assessed the development and delivery practices and the organisation of a software vendor ahead of its sale.","keywords":["Organisation","Software delivery"]},{"name":"Deployment practices audit","entity":"Monoprix","type":"audit","description":"Assessed in-store software delivery processes and made recommendations tailored to the teams.","keywords":["Deployment","Tooling"]},{"name":"Data organisation audit","entity":"Rexel","type":"audit","description":"Analysed how customer data is produced and consumed; recommendations on organisation, data products and governance.","keywords":["Data products","Governance","Service contracts"]},{"name":"Frugal architectures","entity":"OCTO Technology","type":"research","description":"Framework for assessing software architectures through their carbon footprint, with actionable recommendations."},{"name":"CPU cache","entity":"OCTO Technology","type":"talk","description":"Tech talk: cache lines, spatial and temporal locality, cache coherence."},{"name":"Databricks platform","entity":"OCTO Technology","type":"research","description":"In-depth exploration of the Databricks ecosystem and internal knowledge sharing."}],"education":[{"institution":"Université Paris Dauphine","area":"Artificial Intelligence, Systems & Data","studyType":"Master's degree","startDate":"2017-09","endDate":"2018-09"},{"institution":"INSAT, Tunis","area":"Computer networks & telecommunications","studyType":"Engineering degree","startDate":"2012-09","endDate":"2017-09"}],"certificates":[{"name":"Databricks Certified Data Engineer Associate","issuer":"Databricks","url":"https://credentials.databricks.com/02d0f055-f1a8-42de-925e-cbec188a7a5e"},{"name":"Databricks Certified Associate Developer for Apache Spark 3.0","issuer":"Databricks","url":"https://credentials.databricks.com/dbf2d5b1-65be-4689-8e83-598386934ddf"},{"name":"Databricks Partner Training — Solutions Architect Essentials","issuer":"Databricks","url":"https://credentials.databricks.com/d8048b44-5798-4fb3-bd8b-e404d338425b"},{"name":"AWS Certified Solutions Architect — Associate","issuer":"Amazon Web Services"}],"skills":[{"name":"Data engineering","keywords":["Snowflake","dbt","dlt","Databricks","Apache Spark","Delta Lake","Kafka","BigQuery","Airflow","Prefect"]},{"name":"Data quality & observability","keywords":["Elementary","Datadog","Data contracts","Alerting"]},{"name":"AI & LLMs","keywords":["AI agents","LLM evaluation","Synthetic datasets","Dust","Dify","Langfuse","Retool","Claude Code"]},{"name":"Cloud & infrastructure","keywords":["AWS","Azure","GCP","Terraform","GitHub Actions","GitLab CI","Docker"]},{"name":"Programming","keywords":["Python","SQL","Java","Scala"]},{"name":"Architecture & practices","keywords":["Data Mesh","Data modelling","Clean Architecture","Software Craftsmanship","Accelerate","Agile"]}],"languages":[{"language":"French","fluency":"Bilingual"},{"language":"English","fluency":"Professional"}],"interests":[{"name":"Sharing","keywords":["Management","Training","Continuous improvement","Sustainable IT"]},{"name":"Sport & hobbies","keywords":["Running (marathon finisher)","Mountain biking","Photography (nature, sunrises and sunsets)"]}]}