跳转到主要内容
此内容尚未提供您的语言版本,正在以英文显示。

PaddleOCR Document Parsing

Skill 已验证
98

Use this skill to extract structured Markdown/JSON from PDFs and document images—tables with cell-level precision, formulas as LaTeX, figures, seals, charts, headers/footers, multi-column layout and correct reading order. Trigger terms: 文档解析, 版面分析, 版面还原, 表格提取, 公式识别, 多栏排版, 扫描件结构化, 发票, 财报, 复杂 PDF, PDF转Markdown, 图表, 阅读顺序; reading order, formula, LaTeX, layout parsing, structure extraction, PP-StructureV3, PaddleOCR-VL.

AI 摘要

This skill leverages the PaddleOCR API to parse complex documents, extracting text, tables, formulas, and layout information into structured Markdown or JSON. It supports both local files and URLs, with options for output customization and error handling.

Versioning

  • warning:Release ManagementNo manifest version (SKILL.md, package.json, etc.) or GitHub release tags are present, and installation instructions reference HEAD.

安装

npx skills add aidenwu0209/paddleocr-skills

通过 npx 运行 Vercel skills CLI(skills.sh)— 需要本地安装 Node.js,以及至少一个兼容 skills 的智能体(Claude Code、Cursor、Codex 等)。前提是仓库遵循 agentskills.io 格式。

2 days ago
20 stars
Apache-2.0
更新于 2 days ago
查看源代码