Qwen3 VL support merged into llama.cpp

https://github.com/ggml-org/llama.cpp/pull/16780

Benchmarks look pretty good, even better than some of the text only models, make sure to take them with a grain of salt tho

Benchmarks

::: spoiler Qwen3 VL 30b a3b (No Thinking) :::

::: spoiler Visual benchmarks for Qwen3 VL 235 A22B (Thinking) :::

26 points · 3 comments · view on lemmy.world

3 Comments

mudkip@lemdro.id · 2 pts · 320d (2 replies)

i am usng it with openwebui and wireguard at school i just upload a pic of my paper and it does all the problems for me

Xylight@lemdro.id · 1 pts · 320d

4/6 bait

devxyn@sh.itjust.works · 1 pts · 320d
[ removed ]
petey@aussie.zone · 1 pts · 317d

I’ve been trying the 32b instruct variant at Q4_K_M and it’s been solid for general use, tool use, and image comprehension. Pretty impressive

galacticwaffle@lemmings.world · -1 pts · 311d
[ removed ]