Loading...
Loading...
Today we release VisionPsy-Nano: state-of-the-art vision-language models at 460M parameters, built to run on-device. It leads every model in its weight class on 16 of 17 benchmarks with the highest overall normalized score in its class, and it tops all four capability areas: document understanding and OCR, visual perception, reasoning, and instruction following. Open weights, Apache 2.0.
Source:https://x.com/paoloardoino/status/2082456208854692235
Impact Score