Name: Paddleocr Text Recognition
Author: PaddlePaddle

搜索技能.../

Paddleocr Text Recognition | Skills Pool

Identify the input source:
- User provides URL: Use the --file-url parameter
- User provides local file path: Use the --file-path parameter
- User uploads image: Save it first, then use --file-path
Input type note:
- Supported file types depend on the model and endpoint configuration.
- Follow the official endpoint/API documentation for the exact supported formats.
Execute OCR:
```
python scripts/ocr_caller.py --file-url "URL provided by user" --pretty
```
Or for local files:
```
python scripts/ocr_caller.py --file-path "file path" --pretty
```
Default behavior: save raw JSON to a temp file:
- If --output is omitted, the script saves automatically under the system temp directory
- Default path pattern: <system-temp>/paddleocr/text-recognition/results/result_<timestamp>_<id>.json
- If --output is provided, it overrides the default temp-file destination
- If --stdout is provided, JSON is printed to stdout and no file is saved
- In save mode, the script prints the absolute saved path on stderr: Result saved to: /absolute/path/...
- In default/custom save mode, read and parse the saved JSON file before responding
- Use --stdout only when you explicitly want to skip file persistence
Parse JSON response:
- In default/custom save mode, load JSON from the saved file path shown by the script
- Check the ok field: true means success, false means error
- Extract text: text field contains all recognized text
- If --stdout is used, parse the stdout JSON directly
- Handle errors: If ok is false, display error.message
Present results to user:
- Display extracted text in a readable format
- If the text is empty, the image may contain no text
- In save mode, always tell the user the saved file path and that full raw JSON is available there

I've extracted the text from the image. Here's the complete content:

[Display the entire text here]

I found some text in the image. Here's a preview:
"The quick brown fox..." (truncated)

python scripts/ocr_caller.py --file-url "https://example.com/invoice.jpg" --pretty

python scripts/ocr_caller.py --file-path "./document.pdf" --pretty

python scripts/ocr_caller.py --file-url "https://example.com/input" --file-type 1 --pretty

python scripts/ocr_caller.py --file-url "https://example.com/input" --stdout --pretty

{
  "ok": true,
  "text": "All recognized text here...",
  "result": { ... },
  "error": null
}

CONFIG_ERROR: PADDLEOCR_OCR_API_URL not configured. Get your API at: https://paddleocr.com

Show the exact error message to the user (including the URL).
Guide the user to configure securely:
- Recommend configuring through the host application's standard method (e.g., settings file, environment variable UI) rather than pasting credentials in chat.
- List the required environment variables:
```
- PADDLEOCR_OCR_API_URL
- PADDLEOCR_ACCESS_TOKEN
- Optional: PADDLEOCR_OCR_TIMEOUT
```
If the user provides credentials in chat anyway (accept any reasonable format), for example:
- PADDLEOCR_OCR_API_URL=https://xxx.paddleocr.com/ocr, PADDLEOCR_ACCESS_TOKEN=abc123...
- Here's my API: https://xxx and token: abc123
- Copy-pasted code format
- Any other reasonable format
- Security note: Warn the user that credentials shared in chat may be stored in conversation history. Recommend setting them through the host application's configuration instead when possible.
Then parse and validate the values:
- Extract PADDLEOCR_OCR_API_URL (look for URLs with paddleocr.com or similar)
- Confirm PADDLEOCR_OCR_API_URL is a full endpoint ending with /ocr
- Extract PADDLEOCR_ACCESS_TOKEN (long alphanumeric string, usually 40+ chars)
Ask the user to confirm the environment is configured.
Retry only after confirmation:
- Once the user confirms the environment variables are available, retry the original OCR task

API_ERROR: Authentication failed (403). Check your token.

API_ERROR: API rate limit exceeded (429)

python scripts/smoke_test.py

Paddleocr Text Recognition

PaddleOCR Text Recognition Skill

When to Use This Skill

How to Use This Skill

Paddleocr Text Recognition

PaddleOCR Text Recognition Skill

When to Use This Skill

How to Use This Skill

Basic Workflow

IMPORTANT: Complete Output Display

Usage Examples

Understanding the Output

First-Time Configuration

Error Handling

Tips for Better Results

Reference Documentation

Testing the Skill

Feishu Doc

Summarize

Nano Pdf

Diffs

Customs Trade Compliance

Nutrient Document Processing