archive-commons-batch

Collect rights-aware metadata from Internet Archive, Wayback, and OSS sources.

3|5|Updated May 26, 2026
One-click install
npx skills add https://github.com/alex-place/lantern-os --skill archive-commons-batch
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: archive-commons-batch
Source: https://github.com/alex-place/lantern-os/tree/main/skills/archive-commons-batch
Command: npx skills add https://github.com/alex-place/lantern-os --skill archive-commons-batch

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill includes references (resource) components.

What problem does it solve?

This Skill removes the manual burden of finding, validating, and organizing rights-aware metadata for public-domain, Creative Commons, open-source, and free-culture content before any download decision is made.

Core Features & Use Cases

  • Metadata-first intake: Collect item records from Internet Archive, Wayback CDX, and OSS sources without jumping straight to bulk downloads.
  • Rights-aware filtering: Preserve license and rights signals, and clearly separate public-domain and CC-eligible items from held or review-needed items.
  • Source-specific workflows: Support archive search, Wayback capture lookup, and OSS repository discovery for music, movies, software, games, and other reusable media.
  • Use Case: A researcher needs a clean intake list of free music albums and public-domain films, with identifiers, creators, dates, and source URLs ready for review.

Quick Start

Ask the skill to gather rights-aware metadata for a specific Archive.org or Wayback search and return the results as Lantern OS-ready intake records.

Frequently Asked Questions about archive-commons-batch

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I collect Internet Archive metadata without downloading large files?

You can collect Internet Archive metadata without bulk downloads by using a metadata-first intake workflow. This gathers item records, identifiers, and rights signals from archive sources while explicitly rejecting bandwidth-heavy downloads and controlled lending items.

What is rights-aware metadata filtering for public domain and Creative Commons items?

Rights-aware metadata filtering preserves license, rights, and identifier data during archive intake. It separates public-domain and Creative Commons eligible items from held, controlled lending, or rights-unclear items requiring further review.

Can I gather Wayback CDX metadata for open-source software and public domain movies?

Yes, you can gather Wayback CDX metadata for open-source software and public domain movies. The workflow supports source-specific capture lookups and repository discovery for music, movies, software, and games while preserving creator, date, and source URL.

How do I batch filter archive.org search results by license and rights status?

You batch filter archive.org search results by applying rights-aware metadata collection that preserves license and rights signals. This process rejects controlled lending and unclear rights items, returning clean intake lists with titles, creators, dates, and source URLs.

What are the limitations of using Wayback CDX for archive intake workflows?

Limitations of using Wayback CDX for archive intake include rejecting controlled lending items, excluding content with unclear rights status, and avoiding bandwidth-heavy downloads. The workflow strictly prioritizes rights-aware metadata collection over direct media retrieval.