Skip to main content

Log in to comment and vote

Comments5

  • Scarlet Highlighter

    •

    Aug 21, 2025

    This model is a must. Please add it ASAP!

  • An Anonymous User

    •

    Aug 8, 2025

    •

    Merged request

    •

    7 votes

    Update to Qwen3 235B A22B Thinking 2507

    https://huggingface.co/Qwen/Qwen3-235B-A22B-Thinking-2507

    We need more this version that is reasoning, not the Instruct version (without reasoning), as a good replacement to the DeepSeek R1 that was removed from Venice, this new version of Qwen3 235B A22B outperforms in the benchmarks the first version of DeepSeek R1 and even the second one, the 0528.

    I think it could even replace DeepSeek R1 in the API.

    • Coffee Mug

      •

      Jul 25, 2025

      Llama 3.1 405B Instruct can be removed to give more hardware to this model that outperforms DeepSeek R1 0528.

    • Black Quasar

      •

      Aug 1, 2025

      This is exactly what I was looking for. The new qwen 235b 2507 is crazy good. I really don’t understand why r1 is also still on this very old version, when we have 0528.

      Is it so difficult to update? I also agree to just let go of llama 3.1 405b to make room for better models. Stuff is quite old here, but especially the qwen 235b should be updated, now that it’s one of the top models

  • Black Quasar

    •

    Aug 1, 2025

    The new qwen is sick. Is it so difficult to change models with the architecture? This should be very high priority