★ Begriff· Modell-Architektur
Context Window
Maximale Token-Anzahl, die ein LLM in einem Request verarbeiten kann (Input + Output). 2026-Stand: GPT-5 ~256k, Claude Sonnet 4.6 ~1M, Gemini 2.5 ~2M Tokens.
Verwandte Tools
Auch bekannt als
kontext-fenster · kontextfenster
Aktivität
16
Mentions in den letzten 7 Tagen
4 Wochen
⚡neu · 16×
Zuletzt erwähnt in
- Qwen3.8-27B INT4 mit 144K Kontext auf RTX 3090 via vLLM AOT2026-09-13
- Lorivo: Serverlose Hosting-Plattform für LoRA-Adapter auf vLLM-Basis2026-09-13
- LocalLLaMA-Community diskutiert Context- und Memory-Plugins für pi.dev mit lokalen Modellen2026-09-12
- TensorSharp: DeepSeek V4.1 Flash auf 8× A40 mit bis zu 40 tok/s2026-09-12
- Community-Debatte: Westliche Open-Source-Modelle ab 120B für Produktion2026-09-12