← All creators

Aidenwu0209

User

PaddleOCR agent skills for text recognition and document parsing. PaddleOCR 文本识别与文档解析技能仓库。

3 indexed · 0 Featured · 35 stars · avg score 81

Categories

Indexed Skills (3)

Data & Documents Solid

paddleocr-doc-parsing

Use this skill to extract structured Markdown/JSON from PDFs and document images—tables with cell-level precision, formulas as LaTeX, figures, seals, charts, headers/footers, multi-column layout and correct reading order. Trigger terms: 文档解析, 版面分析, 版面还原, 表格提取, 公式识别, 多栏排版, 扫描件结构化, 发票, 财报, 复杂 PDF, PDF转Markdown, 图表, 阅读顺序; reading order, formula, LaTeX, layout parsing, structure extraction, PP-StructureV3, PaddleOCR-VL.

34 Updated 3 weeks ago
Aidenwu0209
Data & Documents Solid

paddleocr-text-recognition

Use this skill whenever the user wants text extracted from images, photos, scans, screenshots, or scanned PDFs. Returns exact machine-readable strings with line-level text and optional bbox coordinates. Strong accuracy for CJK, small print, and handwritten text. Trigger terms: OCR, 文字识别, 图片转文字, 截图识字, 提取图中文字, 扫描识字, 识字, 纯文字, plain text extraction, 坐标, 检测框, bbox, bounding box, image to text, screenshot, photo scan, recognize text.

34 Updated 3 weeks ago
Aidenwu0209
Data & Documents Listed

unlimited-ocr-document-parsing

Parse long images, PDFs, OFD, Office documents, and text files into complete Markdown with Baidu Unlimited-OCR cloud API, or parse local images/PDFs through an SGLang/OpenAI-compatible server. Use for long-document OCR, tables, formulas, reading order, and structured document extraction.

1 Updated 3 weeks ago
Aidenwu0209

Bio shown is the top-scored skill's repo description as a fallback — real GitHub bios land in a future update.