Questions about the Image to Text API

The things developers ask before wiring up the Image to Text API — what it returns, what it covers, and where the limits are. Straight answers, no fluff.

Everything you asked

Still stuck on something specific? The team answers real integration questions directly.

Contact support
What image formats are supported?
We support all common image formats including PNG, JPG, JPEG, GIF, BMP, TIFF, and WebP.
What languages can be recognized?
Our OCR supports Latin, Chinese, Japanese, Korean, Arabic, Hebrew, and many other scripts. English has the highest accuracy.
How accurate is the text extraction?
Accuracy depends on image quality. Clear, high-resolution images typically achieve 95%+ accuracy. Blurry or low-contrast images may have lower accuracy.
Can it read handwritten text?
The API is optimized for printed text. Handwritten text recognition has limited accuracy and works best with neat, clear handwriting.
Is there a maximum image size?
Images up to 10MB are supported. Larger images should be resized before submission.
Can I extract text from PDFs?
For PDFs, you would first need to convert to images. Our API processes image files, not PDF documents directly.

Answer the last question by building. Free tier, no card — run the Image to Text API in minutes.

Scaling up?

Volume pricing, custom SLAs, and dedicated support for high-traffic teams.

Contact sales