paperclip

Extract document metadata and generate YAML/JSON structured output.

Updated Aug 23, 2026
One-click install
npx skills add https://github.com/rodgemd1-lgtm/Startup-Intelligence-OS --skill paperclip-rodgemd1-lgtm
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: paperclip
Source: https://github.com/rodgemd1-lgtm/Startup-Intelligence-OS/tree/main/artifacts/paperclip/startup-intelligence-os-live-post-apply-snapshot/skills/paperclipai/paperclip/paperclip
Command: npx skills add https://github.com/rodgemd1-lgtm/Startup-Intelligence-OS --skill paperclip-rodgemd1-lgtm

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

Paperclip automates the extraction and organization of document metadata to speed up retrieval and analysis.

Core Features & Use Cases

  • Metadata extraction: pull title, author, date, topics, and references from documents.
  • Structured output: generate YAML/JSON representations suitable for indexing in knowledge bases or research management systems.
  • Reference linking: optionally attach references to related documents and assets for traceability.

Quick Start

Provide a sample document to Paperclip and request structured metadata extraction.

Frequently Asked Questions about paperclip

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I extract metadata from multiple documents automatically?

To extract metadata from multiple documents automatically, provide your batch of contracts, reports, or research notes to generate structured YAML or JSON output. The tool applies configurable parsing and entity extraction to organize titles, authors, dates, and topics.

What is the best way to generate structured YAML or JSON from research notes?

The best way to generate structured YAML or JSON from research notes is using automated metadata extraction. This process pulls titles, authors, dates, and topics from your documents to create structured representations suitable for indexing in knowledge bases or research management systems.

Does automated document metadata extraction support reference linking?

Yes, automated document metadata extraction supports optional reference linking. You can attach references to related documents and assets during the extraction process, ensuring full traceability across your indexed knowledge base or research management system.

Can I use metadata extraction for contracts and reports?

Yes, you can use metadata extraction for contracts and reports. The automated extraction process applies configurable parsing to these document types, pulling essential metadata like titles, authors, dates, topics, and references for structured output generation in YAML or JSON.

How does entity extraction work for document indexing?

Entity extraction for document indexing works by applying configurable parsing to identify and pull structured information like titles, authors, dates, and topics from your files. This generates structured YAML or JSON representations suitable for knowledge base indexing.