Data File Summarizer
Paste the following prompt into your AI chat to install this skill:
Install @user_a71e4ceb/summarize-skill-02 into your AI assistant according to https://skillhub.cn/install/skillhub.md.
About this skill
Problem
Summarizing web pages, local files, and YouTube content usually means juggling extraction, model choice, output length, and formatting. summarize data file reduces that to a single CLI workflow, which is useful for scripts, triage, and documentation cleanup.
How It Works
- Inputs: URLs, local files, and YouTube links;
--extract-onlyextracts URL content only. - Model setup: use
OPENAI_API_KEY,ANTHROPIC_API_KEY,XAI_API_KEY, orGEMINI_API_KEY; the default model isgoogle/gemini-3-flash-previewwhen none is set. - Output control:
--lengthselectsshortthroughxxl,--max-output-tokenscaps generation length, and--jsonproduces machine-readable output. - Fallback extraction:
--firecrawlsupportsauto,off, andalways;--youtubecan use the Apify fallback whenAPIFY_API_TOKENis set.
Boundaries
This skill fits engineers who already have a model API key and want repeatable summaries. FIRECRAWL_API_KEY is for blocked sites, and APIFY_API_TOKEN is for YouTube fallback; local-file summarization does not require those extra services.
Use Cases
- Summarize a long technical blog URL into a 200-word incident note
- Read local Markdown files in a script and generate JSON summary fields
- Compress a YouTube lecture link into a project background brief
- Extract page text with Firecrawl fallback when direct fetch fails
Best For
- Technical writers who turn web pages, docs, and YouTube links into report material
- DevOps engineers who generate structured summaries in automation scripts
- AI application engineers comparing summary quality across model providers
- Researchers who need fast digests of local technical documents
Related Skills
Turns Taleb's antifragility, barbell strategy, and via negativa into executable prompts for risk analysis in career, investing, health, and other domains.
Keyword news search over ZAKER’s corpus with optional time filtering, returning up to 20 titles, authors, and summaries.
A generic knowledge review skill that generates role-based checklists, rrule-based reminders, and on-demand reviews without storing personal data.
A knowledge-management skill for layered retrieval across information sources.