openakita/skills@baidu-paddleocr-text

SkillDocs & knowledge

Lets your agent read text out of images, including photos, scanned documents, and handwriting.

Instructions available. Your AI can read the instructions. Execution depends on the setup they require.

Add ahel to your AI once: Claude, ChatGPT, Cursor, Claude Code or Codex. Then ask it to use this.

Then ask your AI: use the openakita/skills@baidu-paddleocr-text skill

About this skill

PaddleOCR text recognition skill using PP-OCRv5 lightweight model. Supports natural scene and complex document text detection and recognition. Use when user needs OCR text extraction from images.

What this skill tells your AI

The instructions your AI receives, as published by openakita/openakita in skills/baidu-paddleocr-text/SKILL.md and read by ahel’s review.

集成 SOTA 级轻量化 OCR 模型 PP-OCRv5,支持自然场景及复杂文档的文字检测与识别。

功能

  • 自然场景文字识别
  • 复杂文档 OCR
  • 多语言支持
  • 轻量化推理

预置脚本

scripts/baidu_ocr_text.py

百度通用文字 OCR 识别,需设置 BAIDU_OCR_AK 和 BAIDU_OCR_SK。

python3 scripts/baidu_ocr_text.py general /path/to/image.jpg
python3 scripts/baidu_ocr_text.py accurate /path/to/image.jpg
python3 scripts/baidu_ocr_text.py handwriting /path/to/note.jpg

Signals

GitHub stars
2k
Forks
280
Last commit
Sep 2026
Advanced
Item type
skill
Key
openakita-skills-baidu-paddleocr-text
Source
github.com/openakita/openakita