data-strategy-architect

Evaluate data sources and design scalable data pipelines with cost-benefit analysis.

Updated Aug 23, 2026
One-click install
npx skills add https://github.com/AIBPM42/hodgesfooshee-site-spark --skill data-strategy-architect
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: data-strategy-architect
Source: https://github.com/AIBPM42/hodgesfooshee-site-spark/tree/main/.claude/skills/data-strategy-architect
Command: npx skills add https://github.com/AIBPM42/hodgesfooshee-site-spark --skill data-strategy-architect

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

Strategic thinking and architecture planning for data acquisition, workflow optimization, and business intelligence systems. Use when deciding data sources, scraping strategies, API integrations, cost-benefit analyses, or pipeline architecture. Helps identify opportunities, evaluate trade-offs, and design scalable data systems.

Core Features & Use Cases

  • Data Source Evaluation: Compare data sources (free vs paid, API vs scraping) and assess reliability, quality, and cost
  • Workflow Architecture: Design efficient data pipelines with scale and fallback strategies
  • Cost-Benefit Analysis: ROI, build vs buy, and long-term cost considerations
  • Strategic Planning: Prioritize features, phased rollouts, and risk mitigation

Quick Start

Use this skill to evaluate data sources and design a scalable pipeline for a new data-driven project.

Frequently Asked Questions about data-strategy-architect

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I evaluate different data sources for a new pipeline?

Data source evaluation compares options—free vs. paid APIs, web scraping, or direct integrations—across reliability, quality, and cost. Assess each source's update frequency, accuracy, compliance requirements, and total-cost-of-ownership to select the optimal fit for your architecture.

What's the best way to design a scalable data pipeline?

Scalable pipeline design structures data workflows with growth capacity, fallback strategies, and phased rollout plans. Build modular stages that handle increasing volume, define redundancy for critical sources, and plan infrastructure cost against performance trade-offs.

How do I calculate ROI for a data pipeline project?

ROI calculation for data pipelines compares build vs. buy costs, infrastructure expenses, and maintenance overhead against business value gained. Include licensing, API fees, storage, compute, labor, and risk mitigation costs over the project lifecycle.

When should I use API integration versus web scraping for data collection?

API integration provides structured, reliable data with vendor support and compliance guarantees, while scraping suits undocumented sources but risks breakage and legal exposure. Choose APIs for critical workflows; reserve scraping for supplemental or legacy data sources.

What governance and compliance considerations apply to data pipeline architecture?

Pipeline governance addresses data lineage, access control, retention policies, and regulatory compliance. Design pipelines to document source provenance, enforce quality standards, and satisfy audit, privacy, and industry-specific requirements across all stages.