
PDF OCR to Text API
Extract searchable text from scanned PDFs
Extract English text from common image formats
Upload an image and extract recognized text lines plus combined text for document intake and automation workflows.
Run OCR whenever your workflow needs it under one annual subscription, without per-call charges or a sales-led setup.
Store recognized text alongside images so receipts, scans, screenshots, and forms become searchable.
Route extracted text into approval workflows, ticketing systems, databases, and downstream AI processing.
Call the endpoint repeatedly from backend jobs when processing incoming image files or historical document queues.
Subscribe to the OCR API independently when the workflow only needs image-to-text output.

HTTP Protocol:HTTPS
HTTP Method:POST
HTTP Endpoint:https://api.gugudata.io/v1/imagerecognition/ocr
Response Type:application/json; charset=utf-8
DEMO Endpoint:https://api.gugudata.io/v1/imagerecognition/ocr/demo
Live Demo:Try Interactive Demo
Full API Docs:developers.gugudata.io
| Name | Type | Is Required | Default Value | Remark |
|---|---|---|---|---|
| appkey | string | true | YOUR_APPKEY | Application key used for request authentication. Supply it as a query parameter. |
| imagefile | file | true | JPEG, PNG, WebP, TIFF, or BMP multipart image upload up to 10 MiB. |
| Name | Type | Remark |
|---|---|---|
| dataStatus.statusCode | integer | Application-level status code. |
| dataStatus.statusDescription | string | Application-level status description. |
| dataStatus.responseDateTime | string | Response timestamp. |
| dataStatus.dataTotalCount | integer | Number of recognized text lines. |
| data.resultText | array<string> | Recognized non-empty text lines in reading order. |
| data.text | string | Recognized lines joined as plain text, or an empty string when no text is recognized. |
| Status Code | Explanation of Status Code | Remarks |
|---|---|---|
| 200 | Request processed successfully. | Some endpoints expose a separate application-level status field in the response body, such as `dataStatus.statusCode`. |
| 400 | Invalid request parameters or request format. | Check required fields, data types, and request body format. |
| 401 | Missing or unknown application key. | Provide a valid `appkey` with the request. |
| 403 | The application key is recognized but access is not allowed. | The key may be expired, inactive, or not permitted for the requested API. |
| 429 | Request rate or trial usage limit exceeded. | Reduce concurrency or retry after the limit window resets. |
| 500 | Internal service error. | Retry later or contact support if the error persists. |
| 503 | Upstream service unavailable. | Retry later; the requested upstream dependency is temporarily unavailable. |
Connect your AI client once, authorize in the browser, and the client can use the GuGuData API tools available to your account. You do not need to paste an appkey into the MCP client.
https://mcp.gugudata.io/mcp{
"mcpServers": {
"gugudata": {
"url": "https://mcp.gugudata.io/mcp",
"transportType": "streamable-http"
}
}
}
Extract searchable text from scanned PDFs

Convert an image into an editable Word document with OCR

Convert PowerPoint slides into PNG images.

Convert PowerPoint slides into PDF documents.