PDF OCR Search API
Start building for free today – no credit card required.
Search for specific text within scanned PDF documents using our intelligent OCR Search API. Locate words, phrases, or patterns in image-based PDFs with advanced text recognition. Ideal for document discovery, compliance checking, and automated content verification across digitized archives.
Claude, ChatGPT, Cursor or your own agent can call
search-pdf-content-with-ocr directly.
Benefits of our PDF OCR Search API
Instantly find specific information within scanned documents using our powerful OCR Search API.
Search image-based PDFs for keywords, phrases, or regex patterns with high accuracy text recognition.
Whether you're conducting compliance audits, locating specific clauses in contracts, or building
document search systems, our API delivers precise results while handling scanned documents that
would otherwise be unsearchable.
Read our API documentation to learn how to search text in scanned PDFs.
Multi-language Code Example
curl -X POST https://apdf.io/api/pdf/ocr/search \
-H "Authorization: Bearer TOKEN" \
-d file="FILE_URL" \
-d text="SEARCH_TEXT"
const data = new FormData();
data.append('file', 'FILE_URL');
data.append('text', 'SEARCH_TEXT');
fetch('https://apdf.io/api/pdf/ocr/search', {
headers: {'Authorization': 'Bearer TOKEN'},
method: 'POST',
body: data
})
.then(response => response.json())
.then(json => console.log(json));
use GuzzleHttp\Client;
$client = new Client();
$response = $client->post(
'https://apdf.io/api/pdf/ocr/search', [
'headers' => [
'Authorization' => 'Bearer TOKEN'
],
'form_params' => [
'file' => 'FILE_URL',
'text' => 'SEARCH_TEXT'
]
]);
$body = $response->getBody();
echo json_encode($body->getContents());
require 'rest-client'
response = RestClient.post(
'https://apdf.io/api/pdf/ocr/search',
{
'file' => 'FILE_URL',
'text' => 'SEARCH_TEXT'
},
{
Authorization: "Bearer TOKEN"
}
)
puts response.body
import requests
response = requests.post(
'https://apdf.io/api/pdf/ocr/search',
headers = {
'Authorization': 'Bearer TOKEN'
},
data = {
'file': 'FILE_URL',
'text': 'SEARCH_TEXT'
}
)
print(response.text)
import (
"fmt"
"github.com/go-resty/resty/v2"
)
func main() {
client := resty.New()
data := map[string]string{
"file": "FILE_URL",
"text": "SEARCH_TEXT"
}
resp, _ := client.R().
SetFormData(data).
SetHeader("Authorization", "Bearer TOKEN").
Post("https://apdf.io/api/pdf/ocr/search")
fmt.Println(resp.String())
}
import okhttp3.*;
class Pdf {
public static void main(String[] args) throws Exception {
OkHttpClient client = new OkHttpClient();
FormBody formBody = new FormBody.Builder()
.add("file", "FILE_URL")
.add("text", "SEARCH_TEXT")
.build();
Request request = new Request.Builder()
.url("https://apdf.io/api/pdf/ocr/search")
.addHeader("Authorization", "Bearer TOKEN")
.post(formBody)
.build();
Response response = client.newCall(request).execute();
System.out.println(response.body().string());
}
}
No-Code OCR PDF Searching
Integrate our PDF API with Zapier Webhooks to automate your OCR search workflows effortlessly. Whether you're scanning scanned invoices for order numbers, searching contracts for specific terms, or verifying compliance across archived documents, you can trigger OCR search from hundreds of apps supported by Zapier. Automatically locate information within image-based PDFs, streamlining document processing and reducing manual review time.
OCR search tutorials
Search Inside Scanned Contracts and Documents with PHP
Find specific text within scanned PDF documents using PHP and the Apdf OCR Search API. Perfect for legal compliance, contract review, and searching through digitized archives where standard PDF search doesn't work.
OCR Scanned PDFs by Asking Claude
That folder of scans — old leases, signed contracts, photographed paperwork — becomes searchable conversationally: Claude checks the text layer, runs the OCR job through the Apdf MCP tools, and answers questions from the recovered text with page-level receipts.
Extract Text from Scanned Invoices for Automated Data Entry
Automate data entry from scanned invoices and receipts using Node.js and the Apdf OCR Read API. Extract vendor names, invoice numbers, and totals from image-based PDFs and feed them directly into your accounting system.
Your code made the PDF.
Then it went dark.
Opened, read, re-read, dropped on page 4 — you never see any of it. Share the PDFs you generate through Apdf recipient links, and every signal becomes something you can act on: ping Slack, update the CRM, let an agent follow up. Same account, same API token, one more call.
API · Webhook · MCP