VisionLaya: Jev with Vision capabilities
VisionLaya: Jev with Vision capabilities
This model makes calibrated, typed decisions about an image plus optional text. It answers choice, score and noul (yes/no probability) questions in one forward pass, with no text generation. It adds image input to Laya by replacing Laya's ModernBERT encoder with SmolVLM-256M-Instruct. Laya's predict(state, questions) API, proper-scoring-rule training and temperature calibration are unchanged. check out the live demo at https://huggingface.co/spaces/thaitea/laya-vision-demo and the source - https://huggingface.co/thaitea/laya-vision-smolvlm-256m - https://github.com/r33drichards/laya-vision
Share cardActual performance
Launch Intel predictions
Analyze your own launch →Incorrect prediction on native model
Similar products
My Capabilities
_done – Expanding AI Capabilities
My implemented vision of a PIM
Halo – Vision Headphones
Our vision of Wunderlist
solidworks
ACME Hugger – Make Nginx have native ACME capabilities
RediSQL a redis module with SQL capabilities
Octogen: e-commerce capabilities for agents
A just another Cron alternative but with much more capabilities