Skip to main content

Elevating Venice AI: Add Vision to the Large Model

Currently, Venice Medium is the one equipped with vision functionality, which is useful for basic image analysis tasks. However, in projects requiring a higher level of detail and precision, such as the generation of complex visual content, in-depth analysis of graphic elements, or advanced customization of styles, this model can fall short in terms of processing power and optimal results.

Incorporating vision capabilities into the Venice Large would open up new possibilities for a wide range of applications. For example, it would enable more sophisticated image analysis in design projects, facilitate the creation of visual prompts with greater depth in creative initiatives, and enhance customization capabilities for specific themes and styles. This could be especially useful for users working on multidisciplinary projects, from generating realistic textures to detailed interpretation of visual elements for various purposes.

Thank you for taking the time to consider this proposal. I am confident that this feature could further elevate the impact of Venice AI within the user community.

Status: Backlog3 comments

Log in to comment and vote

Comments3

  • Tan Caramel

    •

    Oct 27, 2025

    Why does this not have more traction!? :sob:

  • Jade Petal

    •

    Aug 8, 2025

    •

    Merged request

    •

    3 votes

    Vision For Venice Large

    Vision is needed for Venice Large. Venice Medium just continues to repeat itself with the same lines over and over again, and while it has a good way of seeing images because it has the vision feature, it keeps on copying the same lines of the previous messages, and it’s annoying. Giving Venice Large a Vision as well would be a great problem fix.

    • An Anonymous User

      •

      Aug 8, 2025

      •

      Merged request

      •

      2 votes

      Image recognition

      Is it possible to add Image recognition for models like Llama 3.1 405B and Venice Large? So far, image recognition is only possible with Venice Medium, and this model isn’t as effective as Llama 3.1 405B and Venice Large. Would love to know if it is possible to add this feature to future models, if not.