etl
Data pipeline for cleaning, verifying, and publishing global statistics
All Skills in This Repository (2)
Pure Emerald Level IndicatorsFrequently Asked Questions
FAQPage SchemaHow to install owid/etl skills?โผ
Run `npx skills add owid/etl --all -g -y` in your terminal to install all skills in this repository globally.
What does the OWID ETL pipeline do?โผ
It moves external datasets through snapshot, meadow, garden, and grapher stages so raw source files become clean, documented, chart-ready data. The skills automate each stage, from creating snapshots to publishing updates.
How do I update a dataset to a new version?โผ
The update-dataset skill runs the whole refresh end to end: bumping versions, rebuilding steps, checking metadata, and preparing the PR. It pauses only when a decision genuinely needs a human.
Can it catch errors in the source data itself?โผ
Yes. The adversarial-data-review skill verifies metadata claims against the producer's own documentation and cross-checks suspicious values against independent sources, catching unit errors and wrong-year values.
Do I need to know the ETL codebase to use these skills?โผ
No. Skills like create-dataset are designed for non-experts: provide a data file or link, and the agent infers the structure, builds the pipeline, and asks you to review the result.
Related Repositories in Data & Analytics
View All in Data & AnalyticsโPaddleOCR
Extract text, tables, and formulas from PDFs and images
Scrapling
Scrape any website and bypass anti-bot protection with AI
last30days-skill
Research any topic across Reddit, X, YouTube, and the web