← ClaudeAtlas

ocrlisted

Extract text from screenshots, scanned documents, image files, and PDFs. Use when a user asks to OCR, read text from an image/screenshot/scan, transcribe visible text, extract text from a scanned PDF, compare OCR text, or convert image/PDF text into Markdown/JSON/plain text.
OpenCoven/coven · ★ 47 · Data & Documents · score 76
Install: claude install-skill OpenCoven/coven
# OCR Use this skill to extract text from local screenshots, images, scans, and PDFs with a repeatable workflow. ## Quick start Prefer the bundled helper when the user needs exact-ish text extraction from a local file: ```bash python3 ~/.openclaw/workspace/skills/ocr/scripts/ocr.py <path> --format text ``` Useful variants: ```bash # Structured output with per-line confidence/bounding boxes python3 ~/.openclaw/workspace/skills/ocr/scripts/ocr.py screenshot.png --format json # Markdown output from a scanned PDF, OCRing up to the first 5 pages python3 ~/.openclaw/workspace/skills/ocr/scripts/ocr.py scan.pdf --force-ocr --max-pages 5 --format markdown # Multiple languages for macOS Vision OCR python3 ~/.openclaw/workspace/skills/ocr/scripts/ocr.py receipt.jpg --languages en-US,es-ES --format text ``` ## Workflow 1. Locate the file. If the image/PDF is attached in chat, use the local attachment path the runtime provides. 2. For a single image or scanned PDF, run `scripts/ocr.py`. 3. For text-native PDFs, let `ocr.py` use `pdftotext` first; only use `--force-ocr` when the PDF is a scan or the extracted text is wrong. 4. Inspect output before relying on it. OCR can confuse punctuation, columns, totals, handwriting, low contrast, and small text. 5. If the user needs high confidence, run JSON output and mention low-confidence lines or ambiguous characters. 6. If the user asks for a clean transcript, lightly normalize whitespace but do not silently rewrite wording. ## Tool c