Screenshot OCR Text Extractor
Paste the following prompt into your AI chat to install this skill:
Please install @user_af163eaa/z-ocr-grab using https://skillhub.cn/install/skillhub.md.
About this skill
Problem
Text in screenshots, photos, or system UIs is often hard to copy, especially in tables, multi-column layouts, or low-contrast images. Manual transcription can drop words, merge columns, or break hierarchy. Engineers usually need editable plain text they can paste into docs, notes, or code blocks.
How It Works
This skill turns screenshot OCR into a predictable workflow rather than a one-shot image description:
- Image intake: checks image type and confirms whether it contains tables, columns, code snippets, or dense annotations.
- OCR extraction: reads text directly when the platform supports image input; otherwise suggests local tooling such as tesseract or paddleocr.
- Layout proofing: restores paragraphs, removes noise, and fixes obvious misreads, line breaks, or ordering issues.
- Text output: returns text in the conversation by default and does not silently write local files; file saving requires the appropriate file access permission.
Boundaries
It processes locally by default and does not fetch external resources online. It avoids confidential or private screenshots and does not alter the source text. It is useful for first drafts, material organization, and format conversion, but policy decisions and professional proofing still require human review.
Use Cases
- Extract log text from a system error screenshot and turn it into paste-ready troubleshooting notes.
- Recognize multi-column action points from a meeting whiteboard photo and structure them into follow-up items.
- Convert customer feedback screenshots into plain text for easier archiving and search.
- Read table descriptions from product documentation screenshots and restore them into editable paragraphs.
Best For
- Frontend engineers who need to convert UI error screenshots into copyable log text.
- Product operations staff who need to archive and search customer feedback screenshots.
- Technical writers who need to restore table descriptions from documentation screenshots.
- Project managers who want to turn whiteboard photos into action items.
Related Skills
Turns files, references, or chat context into styled single-page HTML reports with built-in templates and preset themes.
An engineering-oriented email automation solution for batch sending, Jinja2 templates, attachments, scheduled sending, inbox monitoring, and rule-based auto-reply.
An engineering workflow for creating, editing, reviewing, analyzing, and image-converting .docx files using pandoc, docx-js, OOXML, and LibreOffice.
Pre-submission scanner for Word/PDF blind-bid files that checks margins, fonts, page numbers, and metadata against tender requirements and flags residual bidder identities.