Image scan for all models
I think that Qwen 2.5 VL is not smart enough to recognize images or describe them correctly. From my perspective, it would be nice if Llama 3.1 or Deepseek were the ones to recognize images.
I think that Qwen 2.5 VL is not smart enough to recognize images or describe them correctly. From my perspective, it would be nice if Llama 3.1 or Deepseek were the ones to recognize images.
Log in to comment and vote
Comments1
JP
Mar 7, 2025
Qwen 2.5 VL is a vision model and is trained on image recognition. The other referenced models are not vision models and can only do text inference.