Subject
3 entries
Data Pipelines
Bookmarks
Ploomber: Data Pipelines from Dev to Production
Ploomber is a Python framework for building data pipelines that can develop in Jupyter notebooks and deploy to Kubernetes, Airflow, or AWS Batch without rewriting code. Solves the notebook-to-production gap by treating notebooks as first-class pipeline tasks.
Migrate Kedro Pipeline to Vertex AI
A walkthrough by Ivan Nardini on migrating a Kedro data science pipeline to run on Google Vertex AI Pipelines — covering the Kedro-Vertex plugin, pipeline conversion, and deployment. Shows how Kedro's portability story works in practice against a major cloud ML platform.
Kedro: Production-Ready Data Science Pipelines
Kedro is an open-source Python framework for building reproducible, maintainable, and modular data science pipelines — applying software engineering principles (catalogs, pipelines, project templates) to ML workflows. The answer to 'how do data science teams write production-grade code.'
