PaddleOCR Document Parsing

CleanCK: Clean

by Lin Manhui

9.9K downloads43 starsv3.0.022 versions

Description

Use this skill to extract structured Markdown/JSON from PDFs and document images—tables with cell-level precision, formulas as LaTeX, figures, seals, charts,...

Changelog

- Significant update: migrated from custom scripts and detailed workflow to usage of the official PaddleOCR CLI. - Removed all helper scripts and schema/reference files. - Updated instructions for document parsing using the new paddleocr api command-line interface. - Simplified configuration: only requires PADDLEOCR_ACCESS_TOKEN and paddleocr CLI. - Added quick-start usage examples with key CLI options and new output format. - Clarified error handling and preprocessing recommendations.

Published Feb 5, 2026Updated Jun 5, 2026

Security Analysis

ClawHub:
Clean
Clawkeeper:
CK: Clean
Risk Score0.0 / 10

No significant risk signals detected. Score is the sum of weighted findings below.

Findings

2 files analyzed

No issues detected

ClawHub scanned: Jun 5, 2026

Rescanned every 6 hours by Clawkeeper

Deploy Securely

Free to scan. Pro to deploy.