e08d812598
- Bench-Matrix Qwen3.6 (5 Configs, separater Port): MTP n-max 3 bestaetigt (+26% tg), n-max 4 lohnt nicht, KV Q8_0 gratis (78,6=78,6 t/s) bei halbem KV-Speicher - deploy-Config: -ctk/-ctv q8_0 am hermes-Eintrag (Live-Schaltung: User-Freigabe noetig) - voice.py: chat_first_content-Metrik (echte Hirn-Latenz bis erster Inhalts-Token) - Report um Umsetzungs-Nachtrag ergaenzt Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>