1.4 KiB
1.4 KiB
Profilmatrix
Diese Datei wird aus config/profile-matrix.json erzeugt. Änderungen gehören nur in die JSON-Matrix.
Standardprofil: medium · globales Ausgabelimit: 8192 Token
| Profil | API-Alias | Gesamtkontext | Slots | Modell | GPU-Verteilung | Vision | MTP |
|---|---|---|---|---|---|---|---|
| fast | qwen-fast |
76,800 | 1 | Qwen3.8-27B IQ4 Mix | 5080 only | ja | 2 |
| medium | qwen-medium |
160,000 | 1 | Qwen3.8-27B IQ4 XS Pure | 85:15 | ja | 3 |
| beta1 | qwen-beta-1 |
112,000 | 1 | Qwen3.8-27B GSQ-RCO IQ3_S MTP | 5080 model / 3060 vision | ja | 3 |
| large | qwen-large |
192,000 | 1 | Qwen3.8-27B IQ4 XS Pure | 86:14 | ja | 3 |
| ultra | qwen-ultra |
262,144 | 1 | Qwen3.8-27B IQ4 XS Pure | 80:20 | nein | 2 |
| uncensored | qwen-uncensored |
80,000 | 1 | Qwen3.8-27B Abliterated Q4_K_M | 90:10 | ja | 2 |
Zweck
- fast: Schnelles Profil für kurze Chats und zügige Werkzeugaufgaben.
- medium: Ausgewogenes Standardprofil für Alltag und lange agentische Aufgaben.
- beta1: Beta 1: qualitaetsoptimiertes GSQ-RCO-IQ3_S-Testprofil; 112K Startkontext bis zur erneuten Grenzmessung.
- large: Großes Profil für umfangreiche Dokumente und lange technische Arbeiten.
- ultra: Maximaler Textkontext; bewusst ohne Vision-Projektor.
- uncensored: Weniger restriktives Spezialprofil; Werkzeugrechte bleiben unverändert.