Hi! I’m the author of this blog.
I’m evaluating these VLMs to figure out which ones are good enough to auto-annotate my data, so I can fine-tune my detector.
I wrote a bit more about this here: https://x.com/skalskip92/status/2080334344061694429?s=20
Did you evaluate any that could be self-hosted (or at least ow models), if so which one is the best you seen?
Take a look here: https://playground.roboflow.com/evals. We have few ~30B.
Thank you!
It seems Qwen is kicking ass, and Fable made me laugh when I saw it all alone on the far right of the graph :))