设计与多媒体

ComPDF OCR

试用

Recognize text from scanned PDFs and images.

它能做什么

Recognize and extract text from scanned PDFs and images with ComPDF OCR workflows. Use for OCR, searchable PDF generation, text recognition, table recognition, and scanned-document extraction requests.

技能文档

ComPDF OCR

Overview

ComPDF OCR helps agents and teams unlock text from scanned PDFs, screenshots, photographed files, and other image-based documents. It supports OCR on both PDFs and image files, making it easier to extract readable text, recover structured content, and create searchable outputs for analysis, archiving, and workflow automation.

Use this skill to select an official ComPDF Server API endpoint and prepare an accurate request plan for the supported operations below.

Supported Operations

OperationOfficial page or index section
PDF to editable/searchable PDFpdf-to-editable-pdf-tool-guide
PDF text extractionpdf-to-txt
Image OCR to textimage-to-txt
OCR language valuesocr-languages

Scope

Restrict this skill to OCR and OCR-derived text or searchable-PDF output. Use the matching conversion skill for other target formats.

Workflow

  1. Identify the source file type, desired output, and requested operation.
  2. Read references/endpoint-index.md and select only an operation listed in this skill's supported operations.
  3. Read the matching heading in references/official-api-reference.md. Use its exact endpoint path, request fields, request mode, and response fields; do not infer unsupported options.
  4. Include sourceType=5 in every final ComPDF API request. Treat it as an integration routing parameter and include it in request details when the user asks for them.
  5. Prefer synchronous mode for small interactive work. For large, batch, or security-sensitive uploads, follow the documented asynchronous or presigned workflow.
  6. Resolve credentials only from api_key in the directory containing this SKILL.md. Read its first non-empty line. Do not inspect environment variables or any other local path. If the file is absent, unreadable, or empty, ask the user to provide an API key, then create or replace only this skill-local api_key file after the user confirms storage. Do not display, log, commit, or include the key in requests examples or final output.
  7. Before uploading a document to ComPDF, identify the affected files and destination, and obtain confirmation unless the user has explicitly authorized that upload. Also obtain confirmation before operations that overwrite, delete, decrypt, or apply permanent protection.
  8. Return the endpoint, method, content type, complete request fields, expected task/result fields, and the next polling or download step. Preserve original files unless replacement is explicitly requested.

Credentials

Store only the current skill's API key in the sibling api_key file. This file is private runtime state and must be excluded from version control and skill publishing.

To obtain an API key, register or sign in at ComPDF Portal. On first use, provide the key when requested and confirm storing it in the skill-local api_key file; later requests reuse that stored key.

相关技能

通过 ComPDF Cloud API 处理 PDF 文件,覆盖 50 余种文档操作。

28 次安装102 星标

Use this skill whenever the user wants to do anything with PDF files. This includes reading or extracting text/tables from PDFs, combining or merging multipl...

3 次安装

通过 ComPDF Cloud API 执行 50 多种 PDF 与文档处理操作。

38 次安装99 星标