wird geladen
Post-Training-Ternarisierung von Qwen3-4B: 1.64 Bit effektiv, halbe Modellgröße · Lumeric