Ross ROSS = Recommend OSS · open-source software intelligence for agents

ploomber/ploomber

The fastest ⚡️ way to build data pipelines. Develop iteratively, deploy anywhere. ☁️ observed · 2026-08-28

github.com/ploomber/ploomber · homepage · Python · Apache-2.0 (permissive) · archived observed · 2026-08-28

Health v2 · maintenance only

10/100

  • Activity 24
  • Release rhythm 35
  • Longevity 100

Flags: no_releases archived

How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-03. Adoption (stars, forks) is never an input.

  • gap_med: n/a
  • age_days: 2417
  • days_rel: n/a
  • days_push: 461
  • n_releases_24m: 0

Full methodology

Adoption not part of the score

3622 stars · 244 forks observed · 2026-08-28

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-29, confidence not recorded

Ploomber is a Python framework for building maintainable data pipelines from scripts and Jupyter notebooks, with iterative local development and deployment to Kubernetes, Airflow, AWS Batch, or SLURM without code changes. It can also refactor legacy notebooks into modular pipelines with a single command.

Use cases

  • build data pipelines from jupyter notebooks
  • convert notebooks into modular pipelines
  • deploy data pipelines to airflow or kubernetes
  • orchestrate machine learning workflows in python
  • develop data pipelines iteratively in vscode or pycharm
  • run pipelines on aws batch or slurm

When to choose

  • your pipelines are built from Python scripts or Jupyter notebooks
  • you want interactive, iterative development with production deployment
  • you need to deploy the same pipeline code to multiple backends like Airflow, Kubernetes, or SLURM
  • you have legacy notebooks that need refactoring into modular pipelines

When to avoid

  • you need a general-purpose workflow scheduler for non-Python tasks
  • your team is standardized on another orchestrator like Dagster or Prefect
  • you only need simple one-off scripts with no pipeline structure

Facets

framework · maturity active

workflow-automation etl machine-learning data-science scheduling data-science machine-learning developer-tools python cross-platform cloud data-pipelines jupyter-notebooks mlops airflow papermill notebook-pipelines data-engineering kubernetes

2 sources

Member repositories

RepositoryRoleHealth v2
ploomber/ploombermain10

For agents

markdown · JSON · MCP: product_card(name="ploomber/ploomber")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem