ocr

OCR skill for extracting text from images and PDFs. Use when you need to read text from screenshots, photos, scanned documents, or any image file. Supports Chinese, English, and 100+ languages.

By mr-shaper · 493 installs

npx skills add mr-shaper/opencode-skill-hybrid-ocr --skill ocr

Source repository · Upstream listing

OCR Skill Usage To extract text from an image or PDF, run: Options Option Description prompt "text" Custom prompt (e.g., "Extract table as markdown") fast Use faster PaddleOCR instead of DeepSeek OCR json Output as JSON format Examples Supported Formats Images: PNG, JPG, JPEG, BMP, GIF, WEBP, TIFF Documents: PDF