Convert any screenshot into clean HTML code using GPT Vision (OSS tool)
Convert any screenshot into clean HTML code using GPT Vision (OSS tool)
Hey everyone, I built a simple React/Python app that takes screenshots of websites and converts them to clean HTML/Tailwind code. It uses GPT-4 Vision to generate the code, and DALL-E 3 to create placeholder images. To run it, all you need is an OpenAI key with GPT vision access. I’m quite pleased with how well it works most of the time. Sometimes, the image generations can be hilariously off. See here for a replica of Taylor Swift’s Instagram page: https://streamable.com/70gow1 I initially had a hard time getting it to work on full page screenshots. GPT4 would code up the first couple of sections and then, get lazy and output placeholder comments for the rest of the page. With some prompt engineering, full page screenshots work a whole lot better now. It’s great for landing pages. Lots of ideas of where to go from here! Let me know if you have feedback and you find this useful :)
Share cardActual performance
Launch Intel predictions
Analyze your own launch →Correct prediction on native model
Similar products
Convert any design into HTML code using GPT4 vision
Design2Code – Convert any screenshot to HTML/CSS code using GPT4 Vision
Convert a screenshot to a working Flutter app (OSS tool)
Screenshot2Code – Generate Code from Screenshots Using GPT-Vision
I wrote a simple tool to convert Markdown to HTML
GPT-3 for vision
Screenshot 2 Text GPT (with vision not OCR)
Using Unicode in HTML
Xlrte – OSS Infrastructure from Code Tool
Convert HTML to HAML in Vim