unstract
Turn unstructured documents into clean structured JSON data
All Skills in This Repository (3)
Pure Emerald Level Indicatorsconnector-ops
Manage Unstract connector lifecycles for databases, filesystems, and queues.
worktree
Create isolated Git worktrees with descriptive branches and Docker service setup.
adapter-ops
Manage LLM and embedding adapters in unstract/sdk1.
Frequently Asked Questions
FAQPage SchemaHow to install Unstract?โผ
Run `npx skills add Zipstack/unstract --all -g -y` in your terminal to install all skills in this suite globally.
How to extract data from PDFs with AI?โผ
Unstract lets you define what to extract using plain-English prompts, then returns clean JSON from PDFs, scans, and images via API or ETL pipelines.
Which LLM providers does Unstract support?โผ
It works with OpenAI, Anthropic, Azure, AWS Bedrock, Google Gemini, Mistral, and local models through Ollama, all pluggable without code changes.
Can Unstract load extracted data into my warehouse?โผ
Yes. Its ETL pipelines pull documents from S3, Google Drive, or Dropbox and load structured results into Snowflake, BigQuery, Redshift, PostgreSQL, and more.
Do I need coding skills to use Unstract?โผ
No. You define extraction schemas in the Prompt Studio using natural language, and the platform handles the LLM integration and deployment for you.
Related Repositories in Data & Analytics
View All in Data & AnalyticsโPaddleOCR
Extract text, tables, and formulas from PDFs and images
Scrapling
Scrape any website and bypass anti-bot protection with AI
last30days-skill
Research any topic across Reddit, X, YouTube, and the web