🦀 ClawHub
General OCR Struct
by @9penny
Offline OCR extracting and structuring Chinese/English screenshot text into raw or cleaned rows and fields for receipts, tables, and statements.
TERMINAL
clawhub install general-ocr-struct📖 About This Skill
name: general-ocr-struct description: General-purpose offline OCR and post-processing for Chinese/English screenshots, scanned images, receipts, tables, chat screenshots, statement screenshots, and other text-heavy images. Use when you need to: (1) extract text from an image locally, (2) return raw OCR text before interpretation, (3) clean broken OCR lines into structured content, (4) reorganize recognized text into rows/fields for downstream use, or (5) separate recognition from later table entry, summarization, or document drafting.
General OCR Struct
Use this skill to separate OCR recognition from downstream content整理.
Workflow
1. Run the local OCR script on the image first.
2. Return the raw OCR text before making business interpretations when accuracy matters.
3. If the image is a transaction-detail screenshot, run structuring mode to group rows into fields.
4. Mark uncertain fields explicitly as 待确认; do not guess missing content.
5. Only after the user confirms recognition quality, use the result for tables, summaries, or documents.
Commands
Raw OCR
python3 scripts/general_ocr.py raw /path/to/image.jpg
Structured transaction extraction
python3 scripts/general_ocr.py transactions /path/to/image.jpg
JSON output
python3 scripts/general_ocr.py transactions /path/to/image.jpg --json
Output rules
待确认 instead of inferring.card_last4, date, time, currency, merchant, amount.