PDF to MD by LLMs – Extract Text/Tables/Image Descriptives by GPT4o
PDF to MD by LLMs – Extract Text/Tables/Image Descriptives by GPT4o
I've developed a Python API service that uses GPT-4o for OCR on PDFs. It features parallel processing and batch handling for improved performance. Not only does it convert PDF to markdown, but it also describes the images within the PDF using captions like `[Image: This picture shows 4 people waving]`. In testing with NASA's Apollo 17 flight documents, it successfully converted complex, multi-oriented pages into well-structured Markdown. The project is open-source and available on GitHub. Feedback is welcome.
Share cardActual performance
Launch Intel predictions
Analyze your own launch →Correct prediction on native model
Similar products
Extract tables from anything.
Extract text from any pdf in the browser
an API to extract text from a PDF
I made a tool in which you can extract CSV tables from PDF files
Extract tables from PDFs — in your browser.
Extract similar tables & CTRL + F words across PDF's & HTML
Extract tables and edit PDF files right in your browser
Extract PDF tables to Excel. Pay per document, not per page.
I built a tool to extract tables from PDFs
Stanchion – Column-oriented tables in SQLite