AI Agent Hub
Back to skills
📊

WeChat Article Fetcher

Data Analysis Updated 2026.08.30

Paste the following prompt into your AI chat to install this skill:

Please install @user_2015f712/wechat-article-fetcher-cn according to https://skillhub.cn/install/skillhub.md.

About this skill

Problem

WeChat Official Account articles are often served as mobile pages, and their content can be hidden by #js_content { visibility: hidden }. Batch fetching can also hit IMA limits such as 220021 quota errors or 220030 file-level rejections. This skill is intended for offline reading and archiving articles with their original layout preserved, not for converting them into Markdown.

How It Works

  • Supports a single URL or batch mode from an IMA knowledge base, with batch filtering preferably using media_id.startswith('wechatarticle_').
  • Fetches with a mobile MicroMessenger UA to avoid the minimal ~10 KB version of the page.
  • Saves each article as a standalone HTML file, keeping div structure, paragraphs, quotes, images, and code blocks.
  • Injects four idempotent fixes: visibility override, desktop layout adaptation, cover removal, and noise stripping.
  • Noise stripping has safe, normal, and strict levels for toolbar, QR code, ad, comment, and H5 component cleanup.

Boundaries

It fits articles where you already have a URL or IMA items with wechatarticle_ media IDs. If an article is not in IMA and you have no URL, source discovery is a separate problem. Use an OCR/document skill for PDFs, scans, or image-heavy articles. Before batch runs, test quota first, avoid retrying 220021 blindly, and do not aggressively bypass IMA limits.

Use Cases

  • After receiving a WeChat article URL, save the text, images, and code blocks as an offline HTML file.
  • Filter exported IMA knowledge-base items whose media_id starts with wechatarticle_ and batch archive Official Account articles.
  • Remove bottom toolbars, QR codes, ads, and comments while preserving the title, body, images, and original layout.
  • Debug a fetched page that lacks content or images by checking for the mobile UA and CSS fix injection.

Best For

  • Knowledge engineers who need to archive WeChat content with its original layout.
  • Content operations staff who maintain IMA knowledge bases and batch-export Official Account materials.
  • Front-end engineers who need to fix anti-crawler CSS in fetched article HTML.
  • Documentation engineers who require offline reading and compliant retention of articles.