wird geladen
Qwen3.5 MTP in llama.cpp: 35,8% Akzeptanzrate auf CUDA, ~92% auf Vulkan/Radeon · Lumeric