Skip to content
sajidalam

London, United Kingdom

Sajid Alam

Senior Software Engineer at QuantumBlack, AI by McKinsey. I maintain Kedro, an open-source Python framework on the Linux Foundation’s LF AI & Data.

Five years on one codebase, in the open. I write the framework internals, the Spark and Databricks integrations, the React tool that draws your pipeline, and the proposals that decide where it all goes next, then I cut the release.

maintaining Kedro
5 yrsmaintaining Kedro
merged pull requests
200+merged pull requests
stars on the framework
10.9kstars on the framework
downloads per month
860kdownloads per month

Open source

Where five years went

Since October 2021 I have worked almost entirely in public, inside the kedro-org GitHub organisation. This is the breakdown.

  • kedro-org/kedro

    The framework itself

    Runner internals, the parameter validation subsystem, the `kedro new` tools flow, micropackaging, Rich logging, cloud `--conf-source`, and release engineering.

    86 merged PRs ·643 commits

  • kedro-org/kedro-plugins

    Datasets & integrations

    SparkDatasetV2, Databricks and Spark Connect utilities, the vector-store abstraction, and the Langfuse/Opik LLM observability family.

    44 merged PRs ·458 commits

  • kedro-org/kedro-viz

    The React visualisation tool

    Flowchart model decomposition, DataCatalog 2.0 lazy loading, modular pipeline expansion, and the Redux preferences layer.

    35 merged PRs ·487 commits

  • kedro-org/kedro-starters

    Project templates

    The spaceflights, pyspark-iris and databricks-iris starters, including the canonical “initialise Spark via hooks” pattern.

    17 merged PRs ·57 commits

  • kedro-org/vscode-kedro

    The editor extension

    Refactored the Python language-server validation architecture into a pluggable validator package with live catalog diagnostics.

    9 merged PRs ·74 commits

  • kedro-org/kedro-academy

    Teaching material

    Observability provider as an environment switch, and evaluation-pipeline training content.

    3 merged PRs ·2 commits

Technical Steering Committee

Named maintainer on the Kedro TSC, one of nine people with a vote on the direction of a Linux Foundation graduate project.

See the maintainer list →

Releases cut personally

kedro
0.17.7 · 0.18.12 · 0.19.2 · 1.3.0 · 1.3.1
kedro-viz
9.2.0 · 10.2.0
vscode-kedro
0.6.0 · 0.7.0

I also moved Kedro’s release process off CircleCI and onto GitHub Actions.

Talks & community

I host the Kedro Coffee Chats

A public technical livestream for the Kedro community. I took over hosting it, which means booking speakers, shaping each session and running the broadcast. I present on it too.

Next up

Dataset Validation is coming to Kedro

Kedro Coffee Chat · August 2026

Presenting the KEP-10 design and a live before/after demo: a supplier feed with a duplicate ID, a malformed rating and an invalid flag, caught at the I/O boundary instead of deep inside a node.

  • Kedro: The Toolbox for Production-Ready Data Pipelines

    May 2026

    GOSIM Paris 2026 · Station F

    Invited talk at the Own Your Data Science & AI Workshop, hosted by scikit-learn, Probabl and Fraunhofer IAIS.

  • GraphRAG with Kedro

    June 2026

    Kedro Coffee Chat · with Laura Couto

    On GraphRAG, agents, and the unglamorous pipeline work that makes agentic systems reproducible. The most-watched Coffee Chat of the series.

  • Host, Kedro Coffee Chats

    2026 – present

    Kedro community livestream

    Took over hosting the community’s public technical stream: booking speakers, shaping each session, and running the broadcast. Recent episodes have covered the Data Catalog, spec-driven development, Kedro in VS Code, and evaluation pipelines with Langfuse.

All talks and streams →

Contact

Open to interesting problems

The work I like best sits where research meets production: frameworks, developer tooling, and the unglamorous plumbing that makes data and AI systems reproducible. If that is what you are building, I would like to hear about it.