Zerox – Document OCR with GPT-mini
Zerox – Document OCR with GPT-mini
This started out as a weekend hack with gpt-4-mini, using the very basic strategy of "just ask the ai to ocr the document". But this turned out to be better performing than our current implementation of Unstructured/Textract. At pretty much the same cost. I've tested almost every variant of document OCR over the past year, especially trying things like table / chart extraction. I've found the rules based extraction has always been lacking. Documents are meant to be a visual representation after all. With weird layouts, tables, charts, etc. Using a vision model just make sense! In general, I'd categorize this solution as slow, expensive, and non deterministic. But 6 months ago it was impossible. And 6 months from now it'll be fast, cheap, and probably more reliable!
Share cardActual performance
Launch Intel predictions
Analyze your own launch →Correct prediction on native model
Similar products
Zerox v1 – Document OCR with GPT-vision
GPT-V and OCR for Screen Control
Feeding GPT-3 context from any document
Homoglyph Attack Prevention with OCR
Announcing GPT-4.1, GPT-4.1 mini, & GPT-4.1 nano in the API
Document OCR for Node.js with GPT Vision
Superclass – GPT-Powered Document Classification Service
Mini-cluster of RPI3
Caterpyl – Mini C to x86_64 Compiler
Java Mini Profiler