ClawData avatar

ClawData

Official

@clawdata

0Followers
|
1Public Repos
|
22Published Skills

Offers modular orchestration and management interfaces for enterprise data pipelines, cloud storage, and relational database warehouse environments.

Skills Distribution
DomainData Systems...Data Engineering &.. (40%)Warehouse Management (30%)Pipeline Orchestra.. (20%)Database Administr.. (10%)

Agent Skills by ClawData

Showing 22 vetted skills indexed across 1 GitHub repositories.

clawdataclawdata
24

fivetran

Manage Fivetran connectors and data syncs via the Fivetran API.

Official
Intermediate
clawdataclawdata
24

s3

Automate AWS S3 bucket listing, file transfers, and directory syncs.

Official
Intermediate
clawdataclawdata
24

dlt

Automate data ingestion pipelines from APIs, databases, and files into destinations.

Official
Advanced
clawdataclawdata
24

azure

Automate Azure resource provisioning and management with guided az CLI workflows.

Official
Advanced
clawdataclawdata
24

test-skill

Validate AI prompt handling with repeatable test cases using YAML frontmatter.

Official
Basic
clawdataclawdata
24

dbt

Manage dbt projects by orchestrating models, tests, docs, and packages via the dbt CLI.

Official
Advanced
clawdataclawdata
24

data-quality

Generate schema YAML tests and custom checks for dbt models.

Official
Intermediate
clawdataclawdata
24

dbt-model-gen

Generate dbt models across bronze, silver, gold, and semantic layers from source definitions.

Official
Intermediate
clawdataclawdata
24

databricks

Execute SQL statements on Databricks warehouses via the REST API.

Official
Intermediate
clawdataclawdata
24

bigquery

Query and manage Google BigQuery datasets using the bq CLI.

Official
Intermediate
clawdataclawdata
24

snowflake

Query and manage Snowflake warehouses, schemas, and tables via the snow CLI.

Official
Intermediate
clawdataclawdata
24

dagster

Orchestrate data pipelines with assets, jobs, schedules, and sensors using Dagster.

Official
Advanced
clawdataclawdata
24

postgres

Manage and query PostgreSQL databases from the command line via psql.

Official
Intermediate
clawdataclawdata
24

github

Manage GitHub repositories, issues, pull requests, releases, and Actions via the gh CLI.

Official
Advanced
clawdataclawdata
24

kafka

Manage Kafka clusters, topics, and messages with CLI tools.

Official
Intermediate
clawdataclawdata
24

data-analysis

Run SQL queries to explore datasets and generate inline chart blocks.

Official
Advanced
clawdataclawdata
24

duckdb

Query local DuckDB warehouses with SQL and ingest CSV, Parquet, and JSON files.

Official
Intermediate
clawdataclawdata
24

weather

Fetch current weather and multi-day forecasts via Open-Meteo with geocoding.

Official
Intermediate
clawdataclawdata
24

metabase

Configure and query Metabase dashboards and questions via API.

Official
Intermediate
clawdataclawdata
24

sql-reviewer

Analyze SQL and dbt models for anti-patterns and performance bottlenecks.

Official
Intermediate
clawdataclawdata
24

airflow-dag-gen

Generate Airflow DAGs from pipeline specifications using project templates.

Official
Intermediate
clawdataclawdata
24

spark

Automate building and running Apache Spark batch and streaming data processing jobs.

Official
Intermediate

Frequently Asked Questions About ClawData

FAQPage Schema
What specific data engineering tasks are supported?β–Ό

ClawData enables end-to-end ingestion from external sources, schema validation for dbt models, SQL execution across BigQuery, Snowflake, and Databricks, and the generation of Airflow DAGs for pipeline scheduling.

Which technical personas benefit from these capabilities?β–Ό

Data engineers, analytics engineers, and database administrators utilize these functions to manage warehouse infrastructure, maintain data quality through automated testing, and streamline the deployment of complex processing jobs.

What are the primary dependencies for running these operations?β–Ό

Execution requires access to target cloud environments such as AWS S3, Azure, or Google Cloud, alongside configured credentials for warehouse platforms like Snowflake, Databricks, or PostgreSQL and local environment support for dbt and Dagster.