wird geladen
Qwen3.8-Flash-Next IQ3_XSS mit mmap in llama.cpp: 26 t/s auf 16 GB VRAM + 64 GB RAM · Lumeric