Databricks Certified · Available for new missions

Mouhammad
Diakhate

Data & Cloud Solutions Architect

I design scalable Lakehouse architectures — from customer discovery to production delivery — across AWS, Azure and GCP.

Portrait of Mouhammad Diakhate
0 Years in customer-facing
data & cloud delivery
0 Missions across energy, logistics,
transport, retail & telecom
0 Databricks
certifications
0 Clouds mastered
AWS · Azure · GCP

01 — About

Turning complex data problems
into architecture that ships

I'm a Data & Cloud Solutions Architect who has spent the last five years inside customer-facing missions — translating messy business problems into Lakehouse and cloud architectures that actually ship. My path has taken me through energy grids, container logistics, transit systems, retail platforms and streaming media, always the same way: understand the problem before the platform, then build something that scales.

Databricks has been at the center of that work since 2023, from designing medallion architectures to leading full Lakehouse migrations. I'm now looking to bring that experience into more architectural and pre-sales oriented roles — while staying hands-on with the platform I know best.

Based in Paris & Lyon, France
Languages French (native),
English (fluent)
Focus Lakehouse & Cloud Architecture
Currently Senior Data Engineer @ Bedrock Streaming

02 — Skills

Core skillset

Lakehouse & Data Platforms

DatabricksApache SparkDelta Lake Delta Live TablesApache IcebergUnity Catalog Databricks WorkflowsLakeflow Declarative Pipelines

Cloud Platforms

AWS S3AWS GlueAWS AthenaAWS Lambda DynamoDBRedshiftStep FunctionsKinesis Azure DatabricksAzure Data FactoryAzure SQL Google Cloud (GKE, Cloud Run)

Architecture & Modeling

Data VaultMedallion ArchitectureSCD2 Change Data FeedData LineageGDPR-compliant Design

IaC, DevOps & CI/CD

TerraformDockerKubernetesSpark Operator HelmArgo CDGitLab CIGitHub ActionsJenkins

Languages

PythonScalaSQL

BI & Reporting

Power BISSRS

04 — Experience

Professional experience

Bedrock Streaming Current

Senior Data Engineer · Freelance

01/2026 – Now

Analyze and improve the data platform for European customers across France, Hungary, Germany and the Netherlands.

  • Collaborate closely with customers to migrate existing data workloads to Databricks
  • Leading migration strategy for 300+ batch legacy workloads to Databricks streaming pipelines
  • Maintain legacy workloads and make existing Data Platform IaC more user-friendly
DatabricksLakeflow Declarative Pipelines TerraformSparkAWS S3 GlueAthenaAirflow

50Hertz (Elia Group)

Senior Data Engineer · Freelance

08/2025 – 01/2026

Proof of concept: challenged against another team on how to build Elia's settlement application — led the rebuild from scratch with Spark to prove feasibility while ensuring reliability.

  • Worked closely with tech and business stakeholders to validate accuracy, scalability and performance
  • Designed and built the Spark-based settlement engine from scratch
  • Ran it all on a low-budget on-prem setup — a single 4-core / 16GB node
  • Delivered 25 settlement products (5 Levies + 20 ADREA) with complex business rules
  • Stress-tested on 4M+ records: under 15-minute total execution time
  • Reached 78% functional and architectural coverage within PoC scope
Apache SparkPythonGitLab CI Argo CDKubernetesSpark Operator DockerHelmMinIO

Sogelink / Groupe SEB

Data Architect · Freelance

11/2024 – 12/2025

Led Sogelink's architecture improvement while helping Groupe SEB maintain and evolve their data platform.

Sogelink

  • Advised stakeholders on migration strategy from a legacy Data Lake to a modern Lakehouse
  • Led the migration of the Data Lake to Apache Iceberg
  • Integrated new CRM/ERP data sources from Dutch and Norwegian subsidiaries
  • Refactored Terraform IaC to streamline developer experience

Groupe SEB

  • Developed API ingestions for new online and retail marketplaces across six countries
  • Maintained and enhanced Retail pipelines — monitoring, incident management, stakeholder alerts
Apache SparkTerraformPython PandasAWSApache Iceberg GitLab CIAWS GlueJenkins

CMA-CGM

Data Engineer Consultant

05/2024 – 09/2024

Set the direction for redesigning legacy EDI preprocessing into a robust data processing tool for container stowage optimization.

  • Built an EDI parser and data model, applying SOLID design patterns
  • Refactored legacy business rules, cutting code volume by 66%
  • Designed a "restow" algorithm to optimize heavy-container placement
  • Prepared a migration strategy and usage guidelines to transfer knowledge internally
Apache SparkPythonPandas AWS LambdaS3DynamoDB Design Patterns

ENGIE

Data Engineer Consultant

01/2023 – 04/2024

Migrated legacy BI from SQL Server (AWS RDS) to a Databricks Lakehouse architecture.

  • Designed a hybrid data vault and medallion Lakehouse architecture on Databricks
  • Built a 3-hop architecture with an SCD2 reliability layer and Change Data Feed
  • Migrated legacy SQL Server workloads to production-ready gold tables for Power BI and SSRS
  • Delivered a data lineage layer and full functional documentation for all migrated jobs
  • Optimized production performance through Spark UI monitoring and tuning
DatabricksApache SparkDelta Lake Databricks WorkflowsDelta Live TablesS3 SQL ServerPower BI

Capgemini

Data Engineer Consultant

02/2022 – 11/2022

Advised on new KPI development for the Transilien T-REX project.

  • Designed cost-effective Azure cloud architecture and secured cloud resources
  • Managed Databricks and Data Factory pipeline scheduling, execution and monitoring
  • Built a star-schema Data Warehouse on Azure, orchestrated via Data Factory
  • Migrated the CD pipeline artifact repository from Nexus to JFrog Artifactory
ScalaAzure DatabricksSpark Azure Data FactoryAzure SQLPower BI

Orange

Junior Data Engineer & DevOps

11/2020 – 08/2021

Automated Cloud infrastructure provisioning and observability of application logs.

  • Automated Cloud infrastructure provisioning with Terraform, Bash and Nginx scripts
  • Built a serverless Kafka pipeline for log ingestion, with source/sink connectors
  • Delivered a self-service gateway and template registry for infra provisioning requests
ElasticsearchKnativeOpenFaaS DockerKubernetesApache Kafka GCP

05 — Education

Education

Université Claude Bernard Lyon 1

Master Degree — Data Science

2022

École Polytechnique de Thiès

Engineering Degree — Computer Design Engineering

2021

École Polytechnique de Thiès

Preparatory Classes

2017

06 — Contact

Let's build something
that scales

Open to Solutions Architecture and pre-sales oriented roles. Reach out directly, or grab a copy of my CV below.

Email mouhammad.diakhate12@gmail.com
Location Paris & Lyon, France
Languages French (native),
English (fluent)

Or send a message directly