Site audit, content writing, and competitor analysis for organic search rankings.
Design & media
Table Ocr
Try itOCR and extract tables from scanned PDFs and images using MinerU. Recognizes table structures in image-based documents and converts them to structured Markdo...
What it does
Convert and extract content from .pdf / images (.png/.jpg/.jpeg/.webp) using MinerU ().
The skill document
Table Ocr
Convert and extract content from .pdf / images (.png/.jpg/.jpeg/.webp) using MinerU (mineru-open-api).
Install
npm install -g mineru-open-api
# or via Go (macOS/Linux):
go install github.com/opendatalab/MinerU-Ecosystem/cli/mineru-open-api@latest
Quick Start
# Extract tables from PDF (requires token)
mineru-open-api extract report.pdf -o ./out/
# With explicit table flag and OCR for scanned docs
mineru-open-api extract scanned.pdf --ocr --table -o ./out/
Authentication
Token required for extract and crawl:
mineru-open-api auth # Interactive token setup
export MINERU_TOKEN="your-token" # Or via environment variable
Create token at: https://mineru.net/apiManage/token
Capabilities
- Supports local files and URLs
- Requires token (
mineru-open-api authorMINERU_TOKENenv) - Supported input: .pdf / images (.png/.jpg/.jpeg/.webp)
- Language hint with
--language(default:ch, useenfor English) - Page range with
--pages(where applicable)
Notes
- Table recognition requires
extractwith token. Use--ocrfor scanned content and--tablefor table detection (both enabled by default in extract). - Output goes to stdout by default; use
-oto save to file - Binary formats (docx) require
-oflag (cannot stream to stdout) - All progress/status messages go to stderr
- MinerU is an open-source project by OpenDataLab (Shanghai AI Lab): https://github.com/opendatalab/MinerU
Related skills
Join a video meeting as an AI bot with voice, avatar, and screenshare across four operating modes.
Handle PDF tasks in Python: merge, split, rotate, extract text/tables/images, OCR scans, and create new PDFs.
chatgpt-apps
OfficialScaffold ChatGPT Apps SDK projects with docs-aligned MCP servers, widgets, and tool plans.
figma-generate-design
OfficialReuse the target Figma file's published design system to build or update full-page screens from code or description.
hatch-pet
OfficialGenerate Codex-compatible animated pets and 9-state atlases from a concept, brand cue, or reference images.
More from mzlzyca
Browse all skillsConvert PDF documents to Word (.docx) format using MinerU. Transforms PDF files into editable Word documents preserving layout, text, tables, and formatting....
Parse and extract structured content from Word documents (.doc, .docx) into well-organized Markdown using MinerU. Preserves the full document hierarchy: head...
Parse academic papers and research documents from PDF using MinerU. Extracts structured content including title, abstract, sections, figures, tables, formula...
OCR for photos and images using MinerU. Extract text from photographs, screenshots, camera captures, and image files with high accuracy. Features: image OCR...
Professional-grade OCR for PDFs and images using MinerU. Advanced text recognition with VLM (Vision Language Model) support for complex layouts, mixed conten...
Convert Word documents (.doc, .docx) to clean, well-structured Markdown using MinerU's document processing engine. Ideal for turning Microsoft Word files int...