Data & analysis

OCR Web Service

Try it

OCR Web Service (ocrwebservice.com). Use this skill for ANY OCR Web Service request — searching and reading data. Whenever a task involves OCR Web Service, use this skill instead of calling the API directly.

What it does

OCR Web Service (ocrwebservice.com). Use this skill for ANY OCR Web Service request — searching and reading data. Whenever a task involves OCR Web Service, use this skill instead of calling the API directly.

The skill document

OCR Web Service

Operate OCR Web Service through your OOMOL-connected account. This skill calls the ocr_web_service connector with the oo CLI; OOMOL injects credentials server-side, so you never handle raw tokens.

Running an action

Assume the user has already installed the oo CLI, signed in, and connected OCR Web Service. Do not run oo auth login or open the connection URL proactively — just run the action. Fall back to First-time setup only when a command actually fails with an auth or connection error.

1. Inspect the contract to get the authoritative input/output schema before building a payload:

oo connector schema "ocr_web_service" --action ""

2. Run the action with a JSON payload that matches the input schema:

oo connector run "ocr_web_service" --action "" --data '' --json
  • --data takes a JSON object string or @path/to/file.json; omit it to send {}.
  • The response is { "data": ..., "meta": { "executionId": "..." } }; the execution id lives under meta.executionId.

Each action is listed below with a one-line description; actions that change state carry a [write] or [destructive] tag. Before constructing --data, fetch the action's live schema with oo connector schema to get its authoritative input fields.

Available actions

  • get_account_information — Get OCR Web Service account limits, remaining pages, subscription plan, and expiration metadata.
  • process_document_from_url — Download a public image or PDF URL through the connector SSRF guard, upload it to OCR Web Service, and return extracted text or output file metadata.

Safety

  • Untagged actions are reads (get / list / search) — safe to run directly.
  • Actions tagged [write] change OCR Web Service state — confirm the exact payload and effect with the user before running.
  • Actions tagged [destructive] remove or overwrite data — always confirm the target and get explicit approval first.

First-time setup

These are one-time steps — do not repeat them on every call. Run a step only when a command fails for the matching reason.

  • oo: command not found — install the oo CLI (other platforms: ):

    curl -fsSL https://cli.oomol.com/install.sh | bash    # macOS / Linux
    
    irm https://cli.oomol.com/install.ps1 | iex           # Windows PowerShell
    
  • Not signed in / authentication error — sign in to your OOMOL account once:

    oo auth login
    
  • scope_missing / credential_expired / app_not_ready / app_not_found — OCR Web Service is not connected, or the connection expired or lacks a scope. Connect once (auth type: custom credential) at:

    https://console.oomol.com/app-connections?provider=ocr_web_service
    
  • HTTP 402 / OOMOL_INSUFFICIENT_CREDIT — billing stop. Recharge at https://console.oomol.com/billing/token-recharge before retrying.

Resources

Related skills

WebScraping.AI (webscraping.ai). Use this skill for ANY WebScraping.AI request — searching and reading data. Whenever a task involves WebScraping.AI, use thi...

1 installs

AnySearch (anysearch.com). Use this skill for ANY AnySearch request — searching and reading data. Whenever a task involves AnySearch, use this skill instead of calling the API directly.

3 installs

Webex (webex.com). Use this skill for ANY Webex request — reading, creating, updating, and deleting data. Whenever a task involves Webex, use this skill instead of calling the API directly.

Use this skill when the user asks to OCR, transcribe, extract, or convert the contents of a scanned PDF, image, or office document into Markdown, HTML, DOCX,...

29 installs1 stars

See: https://github.com/anyforge/anyparse Use the AnyParse API to extract content from various documents. Supports PDF, Word, Excel, CSV, TSV, images, PPT, HTML, Markdown, Epub, ipynb, RST, EML, and many other formats. Supports document orientation classification, layout analysis, and layout preservation.

Use this skill whenever the user wants to do anything with PDF files. This includes reading or extracting text/tables from PDFs, combining or merging multipl...

3 installs