Screen OCR Text Recognition
Paste the following prompt into your AI chat to install this skill:
Please install @user_25ab92ec/ocr-screen according to https://skillhub.cn/install/skillhub.md.
About this skill
Problem
Screen text is often visible but not editable: it appears in screenshots, remote desktop windows, shared apps, or rendered document pages. ocr-screen targets this workflow by converting visible screen text into copyable plain text, especially for mixed Chinese and English content, table fragments, and batch capture tasks.
Core behavior
- Chinese and English recognition: supports Chinese and English text, which is useful for bilingual UI, annotations, field labels, and body copy.
- Table recognition: recognizes tabular structures in captured screens, helping keep table-like content easier to organize.
- Batch processing: handles multiple screen regions or items in sequence instead of one small crop at a time.
- Export-friendly output: returns recognized text as copyable plain text, with optional export to
Word/Excelfor downstream document or spreadsheet work.
Boundaries
- Accuracy depends on screen clarity, font size, contrast, and the integrity of table lines.
- It is oriented toward visible screen text; if the source file is natively editable, using the original text will usually be more reliable.
- Dependencies must be installed on first use, and the runtime environment can affect stable operation.
Use Cases
- During remote support, convert Chinese and English screen prompts into copyable text.
- Organize screenshots containing bilingual tables and export recognized text to Excel.
- Batch recognize UI fields from multiple shared windows into plain text for requirement checks.
- Turn bilingual on-screen instructions into plain text for tickets or notes.
Best For
- Remote support engineer who needs copyable text from error messages and UI labels on a client screen.
- Documentation specialist who needs table text from screenshots exported to Excel.
- QA engineer who needs batch OCR of multiple window fields for defect reports.
- Delivery specialist who needs bilingual UI instructions turned into archive-ready plain text.
Related Skills
Generate and edit .pptx decks with python-pptx, applying structured layouts, design rules, native charts, and visual QA to reduce template-like output.
Tencent Cloud Table Recognition V3 is an OCR skill for detecting and recognizing tables in images or PDFs, supporting various table types like linear and borderless tables, with Excel export.
The complete set of online document operation tools provided by Tencent Docs MCP, supporting creation, querying, and editing of smart docs, Excel, PPT, mind maps, and more.
Generates structured and consistently styled academic presentation PPTX files from paper PDFs for graduate seminars.