CheckCom KI-Preise · Lokale KI

Lokale & quantisierte KI-Modelle

Hardwarebedarf, Quantisierung, lokale Eignung und Messstatus für Unternehmen. Die Übersicht trennt Metadaten, externe Messwerte und spätere CheckCom-eigene Benchmarks klar voneinander.

Stand: 22.08.2026 11:17

Was diese Übersicht leistet

Lokal statt API?

CheckCom ordnet ein, wann lokale Workstation-, Edge- oder API-Routen für Unternehmen prüfenswert sind.

Quantisierung verständlich

Q4, Q5, Q6, Q8, GGUF und weitere Signale werden als Hardware- und Eignungshinweise ausgewertet.

Messstatus sichtbar

Öffentliche Werte werden als Metadaten, externe Orientierung oder später als CheckCom-eigene Messung gekennzeichnet.

850
lokale/quantisierte Kandidaten
433
Anbieter / Familien
15210
RTX-5090-Kandidaten
23
Ollama-Hinweise
251
lokale News-/Release-Signale
Wichtig zur Einordnung: Quantisierte Modelle und lokale Hardwarewerte sind nur vergleichbar, wenn Modellversion, Quantisierung, Runtime, Kontextlänge, Prompt und Hardwareprofil dokumentiert sind. V3.1 zeigt deshalb zunächst Hardwarefit, Metadaten und externe Benchmarkquellen; echte CheckCom-Tokens/s-Werte folgen erst nach eigenen Messläufen.

Hardwareprofile

HardwareGeeignet fürGrenzenMessstatus
RTX 5090 / 32 GB VRAM7B–35B quantisierte LLMs, Coding, lokale RAG-Tests, schnelle lokale Inferenzsehr große 70B+ Modelle nur mit starker Quantisierung/Offload; mehrere parallele Nutzer prüfenbenchmark_prepared
RTX 4090 / 24 GB VRAMkleinere bis mittlere quantisierte Modelle und lokale Experimente32B+ und lange Kontexte häufig knappexternal_and_future_checkcom
RTX Pro / 96 GB VRAMgroße lokale Modelle, längere Kontexte, professionelle lokale Inferenzhohe Anschaffungskosten; TCO prüfenprofile_prepared
Raspberry Pi 5 + AI HAT+ 2 / Hailo-10HEdge-KI, kleine LLM-/VLM-Kandidaten, Vision, lokale Demo- und Datenschutzszenarienkein Ersatz für große GPU-LLMs; Hailo-kompatible Modelle und Runtime nötigdemo_profile_available
Raspberry Pi 5 CPU-onlykleine CPU-Tests, Klassifikation, Offline-Demosgroße LLMs und lange Kontexte nicht sinnvollinventory_prepared
API / Cloud / EnterpriseSkalierung, Enterprise-Verträge, hohe Parallelität und schnelle Produktivstartslaufende Token-/Seat-Kosten, Datenroute und Vertrag prüfenprice_radar_linked

Lokale und quantisierte Modellkandidaten

Aus dem CheckCom-Modellkatalog abgeleitet. Messstatus „metadata_only“ bedeutet: noch kein CheckCom-Benchmarkwert.

ModellParams BQuantisierungVRAM GBRTX 5090Pi 5 / AI HAT+ 2Score
Qwen2.5-Coder 7B Instruct
Alibaba / Qwen
7.0 Q4
Q4
8.0 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
95/100
metadata_only
Qwen3.6 35B A3B Vocabulary Trimming GGUF
Alibaba / Qwen
35.0 GGUF
Q4_K_S, GGUF, gguf
28.2 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 35B A3B Uncensored xCloud GGUF
Alibaba / Qwen
35.0 GGUF
Q4_K_M, GGUF, gguf
28.2 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 35B A3B TQ2Q GGUF
Alibaba / Qwen
35.0 GGUF
GGUF, gguf
28.2 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 35B A3B StyleTune GGUF
Alibaba / Qwen
35.0 GGUF
IQ4_XS, GGUF, gguf
28.2 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 35B A3B Q5 K M GGUF
Alibaba / Qwen
35.0 GGUF
q5_k_m, Q5, GGUF
32.7 offload
nur mit Offload oder reduzierten Einstellungen realistisch
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 35B A3B Ornith 80 20 Beta GGUF
Alibaba / Qwen
35.0 GGUF
Q4_K_S, GGUF, gguf
28.2 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 35B A3B NVFP4 Gguf
Alibaba / Qwen
35.0 GGUF
Gguf, gguf
28.2 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 35B A3B NVFP4 GGUF
Alibaba / Qwen
35.0 GGUF
q3, GGUF, gguf
22.3 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 35B A3B MTP GGUF Tater NoThink
Alibaba / Qwen
35.0 GGUF
Q4_K_M, GGUF, gguf
28.2 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 35B A3B DS4 GGUF
Alibaba / Qwen
35.0 GGUF
Q4_K_S, GGUF, gguf
28.2 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 35B A3B Claude 4.7 Distill NVFP4 GGUF
Alibaba / Qwen
35.0 GGUF
GGUF, gguf
28.2 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 35B A3B Claude 4.7 Distill MXFP4 MoE GGUF
Alibaba / Qwen
35.0 GGUF
GGUF, gguf
28.2 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.5 35B A3B BitClass3 GGUF
Alibaba / Qwen
35.0 GGUF
Q3_K_S, GGUF, gguf
22.3 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.5 35B A3B APEX GGUF
Alibaba / Qwen
35.0 GGUF
GGUF, gguf
28.2 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Huihui Qwen AgentWorld 35B A3B Abliterated UD Q3 K M GGUF
Alibaba / Qwen
35.0 GGUF
Q3_K_M, Q3, GGUF
22.3 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Endy Qwen3.6 CyberSec 35B A3B GGUF
Alibaba / Qwen
35.0 GGUF
Q2_K, GGUF, gguf
18.7 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3 32B Qwople GGUF
Alibaba / Qwen
32.0 GGUF
Q4_K_M, GGUF, gguf
26.3 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3 32B GGUF
Alibaba / Qwen
32.0 GGUF
IQ3_M, GGUF, gguf
20.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen2.5 Coder 32B Python Specialist I1 GGUF
Alibaba / Qwen
32.0 GGUF
IQ3_M, GGUF, gguf
20.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen2.5 Coder 32B Python Specialist GGUF
Alibaba / Qwen
32.0 GGUF
IQ4_XS, GGUF, gguf
26.3 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen2.5 Coder 32B Instruct GGUF
Alibaba / Qwen
32.0 GGUF
Q2_K, GGUF, gguf
17.7 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen2.5 32B Instruct AdiTurbo GGUF
Alibaba / Qwen
32.0 GGUF
Q3_K_M, GGUF, gguf
20.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
DeepSeek R1 Distill Qwen32B Q4 K M GGUF
Alibaba / Qwen
32.0 GGUF
q4_k_m, Q4, GGUF
26.3 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3 Coder 30B A3B Instruct GGUF
Alibaba / Qwen
30.0 GGUF
Q4_K_M, GGUF, gguf
25.1 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3 30B A3B GGUF
Alibaba / Qwen
30.0 GGUF
IQ3_M, GGUF, gguf
20.0 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Ektome Qwen3 30B A3B PristinelyUncensored GGUF
Alibaba / Qwen
30.0 GGUF
Q2_K, GGUF, gguf
17.0 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.8 27B ZipBrain GGUF
Alibaba / Qwen
27.0 GGUF
IQ4_XS, GGUF, gguf
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.8 27B Uncensored Q4 K M GGUF
Alibaba / Qwen
27.0 GGUF
Q4_K_M, Q4, GGUF
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.8 27B Uncensored Aggressive I1 GGUF
Alibaba / Qwen
27.0 GGUF
IQ3_M, GGUF, gguf
18.6 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.8 27B Uncensored Aggressive GGUF
Alibaba / Qwen
27.0 GGUF
IQ4_XS, GGUF, gguf
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.8 27B ROCmFPX GGUF
Alibaba / Qwen
27.0 GGUF
GGUF, gguf
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.8 27B I1 GGUF
Alibaba / Qwen
27.0 GGUF
IQ3_M, GGUF, gguf
18.6 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.8 27B Fable5 Distill Abliterated Imatrix GGUF
Alibaba / Qwen
27.0 GGUF
GGUF, gguf
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.8 27B Fable Distill GGUF
Alibaba / Qwen
27.0 GGUF
IQ4_XS, GGUF, gguf
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.8 27B Cold Fusion GAIN V1.1 Q4 K M GGUF
Alibaba / Qwen
27.0 GGUF
q4_k_m, Q4, GGUF
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.8 27B Abliterated MTP GGUF
Alibaba / Qwen
27.0 GGUF
IQ2_M, GGUF, gguf
15.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.8 27B 5090 Goldilocks GGUF
Alibaba / Qwen
27.0 GGUF
GGUF, gguf
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 27B Samantha Uncensored GGUF
Alibaba / Qwen
27.0 GGUF
Q4_K_M, GGUF, gguf
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 27B Q4 K M GGUF
Alibaba / Qwen
27.0 GGUF
q4_k_m, Q4, GGUF
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 27B NVFP4 GGUF
Alibaba / Qwen
27.0 GGUF
q3, GGUF, gguf
18.6 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 27B Mtp Gguf
Alibaba / Qwen
27.0 GGUF
Q4_K_X, Gguf, Q4_K_XL
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 27B KR MTP GGUF
Alibaba / Qwen
27.0 GGUF
Q5_K_X, GGUF, Q5_K_XL
26.7 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 27B KR GGUF
Alibaba / Qwen
27.0 GGUF
Q5_K_X, GGUF, Q5_K_XL
26.7 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 27B GGUF
Alibaba / Qwen
27.0 GGUF
Q4_K_M, GGUF, gguf
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 27B Esper4 I1 GGUF
Alibaba / Qwen
27.0 GGUF
IQ3_M, GGUF, gguf
18.6 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 27B Esper4 GGUF
Alibaba / Qwen
27.0 GGUF
IQ4_XS, GGUF, gguf
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 27B Architect Polaris2 Fable B F451 MTP ROCmFPX GGUF
Alibaba / Qwen
27.0 GGUF
Q6, GGUF, gguf
30.8 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 27B AEON RYS Agentic Coder PatchCode GGUF
Alibaba / Qwen
27.0 GGUF
IQ4_NL, GGUF, gguf
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Ostrich 27B Qwen3.8 260815 I1 GGUF
Alibaba / Qwen
27.0 GGUF
IQ3_M, GGUF, gguf
18.6 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Huihui Qwen3.8 27B Abliterated GGUF
Alibaba / Qwen
27.0 GGUF
IQ3_M, GGUF, gguf
18.6 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Elster Vernunft Qwen3.6 27B GGUF
Alibaba / Qwen
27.0 GGUF
IQ4_XS, GGUF, gguf
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Dagger Qwen3.6 27B GGUF MTP
Alibaba / Qwen
27.0 GGUF
Q3_K_M, GGUF, gguf
18.6 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.8 20B Minitron Q4 K M GGUF
Alibaba / Qwen
20.0 GGUF
q4_k_m, Q4, GGUF
18.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 14B A3B VibeForged V2 GGUF
Alibaba / Qwen
14.0 GGUF
F16, GGUF, gguf
20.4 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3 14B PragReST I1 GGUF
Alibaba / Qwen
14.0 GGUF
IQ1_M, GGUF, gguf
20.4 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen2.5 Coder 14B Instruct Jbliterated
Alibaba / Qwen
14.0 GGUF
Q4_K_M, gguf
13.7 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen2.5 Coder 14B Instruct Heretic GGUF
Alibaba / Qwen
14.0 GGUF
IQ4_XS, GGUF, gguf
13.7 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen2.5 Coder 14B Instruct GGUF
Alibaba / Qwen
14.0 GGUF
f16, GGUF, gguf
20.4 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen2.5 Coder 14B Instruct GGUF
Alibaba / Qwen
14.0 GGUF
fp16, GGUF, gguf
34.4 offload
nur mit Offload oder reduzierten Einstellungen realistisch
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen2.5 14B Instruct Heretic
Alibaba / Qwen
14.0 GGUF
Q4_K_M, gguf
13.7 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Lfed Qwen2.5 Coder 14B Sql Gguf
Alibaba / Qwen
14.0 GGUF
Q4_K_M, Gguf, gguf
13.7 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Iol Qwen2.5 14B Instruct AWQ
Alibaba / Qwen
14.0 AWQ
AWQ
13.7 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
H100 Qwen3 14B MonEspaceSante CPT I1 GGUF
Alibaba / Qwen
14.0 GGUF
IQ2_M, GGUF, gguf
9.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
H100 Qwen3 14B MonEspaceSante CPT GGUF
Alibaba / Qwen
14.0 GGUF
IQ4_XS, GGUF, gguf
13.7 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Ektome Qwen2.5 Coder 14B Instruct PristinelyUncensored
Alibaba / Qwen
14.0 GGUF
IQ3_M, gguf
11.3 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Adi Qwen2.5 14B GLM 5.2 General GGUF
Alibaba / Qwen
14.0 GGUF
q4_k_m, GGUF, gguf
13.7 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
QwenPaw Flash 9B Heretic MTP GGUF
Alibaba / Qwen
9.0 GGUF
BF16, GGUF, gguf
23.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
QwenPaw Flash 9B Heretic Imatrix GGUF
Alibaba / Qwen
9.0 GGUF
BF16, GGUF, gguf
23.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
QwenPaw Flash 9B Heretic GGUF
Alibaba / Qwen
9.0 GGUF
BF16, GGUF, gguf
23.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.5 9B Ultra Uncensored Heretic V2 I1 GGUF
Alibaba / Qwen
9.0 GGUF
IQ1_M, GGUF, gguf
14.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.5 9B Ultra Uncensored Heretic V1 I1 GGUF
Alibaba / Qwen
9.0 GGUF
IQ1_M, GGUF, gguf
14.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.5 9B The Defiant Fable Uncensored Heretic NEO IMATRIX MAX MTP GGUF
Alibaba / Qwen
9.0 GGUF
BF16, GGUF, gguf
23.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.5 9B The Defiant Fable Uncensored Heretic NEO IMATRIX MAX GGUF
Alibaba / Qwen
9.0 GGUF
BF16, GGUF, gguf
23.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.5 9B Samantha Uncensored GGUF
Alibaba / Qwen
9.0 GGUF
Q4_K_M, GGUF, gguf
10.6 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.5 9B Heretic Imatrix GGUF
Alibaba / Qwen
9.0 GGUF
BF16, GGUF, gguf
23.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.5 9B Heretic GGUF
Alibaba / Qwen
9.0 GGUF
BF16, GGUF, gguf
23.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.5 9B Haskell Rust Python Q6 K GGUF
Alibaba / Qwen
9.0 GGUF
q6_k, Q6, GGUF
13.1 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.5 9B Haskell Rust Python IQ4 XS GGUF
Alibaba / Qwen
9.0 GGUF
iq4_xs, GGUF, gguf
10.6 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.5 9B GGUF
Alibaba / Qwen
9.0 GGUF
IQ1_KT, GGUF, gguf
14.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.5 9B DeepSeek V4 Flash GGUF
Alibaba / Qwen
9.0 GGUF
Q3_K_M, GGUF, gguf
9.0 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.5 9B DFlash GGUF
Alibaba / Qwen
9.0 GGUF
bf16, DFlash, GGUF
23.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Huihui Qwen3.5 9B Claude 4.6 Opus Abliterated Heretic GGUF
Alibaba / Qwen
9.0 GGUF
f16, GGUF, gguf
14.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
FACET Terminal Qwen3.5 9B GGUF
Alibaba / Qwen
9.0 GGUF
f16, GGUF, gguf
14.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Adi Qwen3.5 9B GLM 5.2 General GGUF
Alibaba / Qwen
9.0 GGUF
q4_k_m, GGUF, gguf
10.6 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
VideoKR Qwen3 VL 8B GGUF
Alibaba / Qwen
8.0 GGUF
f16, GGUF, gguf
13.8 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
SenseNova U1.5 8B GGUFs
Hugging Face / sonstige Owner
8.0 GGUF
Q2_K, gguf
7.8 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
RavenX Conjecture Qwen3 8B GGUF
Alibaba / Qwen
8.0 GGUF
Q8, GGUF, Q8_0
13.8 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3 VL 8B Instruct Heretic GGUF
Alibaba / Qwen
8.0 GGUF
f16, GGUF, gguf
13.8 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3 VL 8B Instruct Abliterated V2 GGUF
Alibaba / Qwen
8.0 GGUF
f16, GGUF, gguf
13.8 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3 Embed 8B GGUF
Alibaba / Qwen
8.0 GGUF
iq4_xs, GGUF, gguf
10.0 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3 8B UnBias Plus SFT Instruct Legacy GGUF
Alibaba / Qwen
8.0 GGUF
f16, GGUF, gguf
13.8 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3 8B PragReST I1 GGUF
Alibaba / Qwen
8.0 GGUF
IQ1_M, GGUF, gguf
13.8 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3 8B PragReST GGUF
Alibaba / Qwen
8.0 GGUF
f16, GGUF, gguf
13.8 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3 8B MedReasonPath GGUF
Alibaba / Qwen
8.0 GGUF
f16, GGUF, gguf
13.8 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3 8B Heretic GGUF
Alibaba / Qwen
8.0 GGUF
f16, GGUF, gguf
13.8 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3 8B Gguf
Alibaba / Qwen
8.0 GGUF
Q8, Gguf, Q8_0
13.8 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3 8B Apostate I1 GGUF
Alibaba / Qwen
8.0 GGUF
IQ1_S, GGUF, gguf
13.8 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3 8B Apostate GGUF
Alibaba / Qwen
8.0 GGUF
f16, GGUF, gguf
13.8 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
PDP Qwen3 8B SFT GGUF
Alibaba / Qwen
8.0 GGUF
f16, GGUF, gguf
13.8 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
LFM2.5 8B A1B GGUF
LFM2.5
8.0 GGUF
BF16, GGUF, gguf
21.8 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
L40S Qwen3 8B MonEspaceSante CPT SFT Anti Hallucination I1 GGUF
Alibaba / Qwen
8.0 GGUF
IQ2_S, GGUF, gguf
7.8 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Huihui Qwen3 VL 8B Instruct Abliterated FP8
Alibaba / Qwen
8.0 FP8
FP8
10.0 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Huihui Qwen3 VL 8B Instruct Abliterated AWQ Int4
Alibaba / Qwen
8.0 AWQ
AWQ, Int4, int4
10.0 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Adi Qwen3 8B GLM 5.2 General GGUF
Alibaba / Qwen
8.0 GGUF
q4_k_m, GGUF, gguf
10.0 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Yoiko Img Prompt Gen Qwen2.5 7B Instruct Q4 K M
Alibaba / Qwen
7.0 GGUF
Q4_K_M, Q4, gguf
9.3 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
VideoKR Qwen2.5 VL 7B GGUF
Alibaba / Qwen
7.0 GGUF
f16, GGUF, gguf
12.7 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Typhoon2 Qwen2.5 7B Instruct GGUF
Alibaba / Qwen
7.0 GGUF
Q4_K_M, GGUF, gguf
9.3 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Stiles Seymens Qwen2.5 7B Generalist Q4KM GGUF
Alibaba / Qwen
7.0 GGUF
Q4_K_M, GGUF, gguf
9.3 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Sofia Qwen2.5 7B I1 GGUF
Alibaba / Qwen
7.0 GGUF
IQ1_M, GGUF, gguf
12.7 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Sofia Qwen2.5 7B GGUF
Alibaba / Qwen
7.0 GGUF
f16, GGUF, gguf
12.7 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen2.5 VL 7B Instruct Int8 Convrot Comfyui
Alibaba / Qwen
7.0 Int8
Int8, int8
9.3 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen2.5 Coder 7B VN Master Polymath FINAL GGUF
Alibaba / Qwen
7.0 GGUF
Q3_K_M, GGUF, gguf
8.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen2.5 Coder 7B Instruct Jbliterated
Alibaba / Qwen
7.0 GGUF
Q4_K_M, gguf
9.3 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen2.5 Coder 7B Instruct Heretic
Alibaba / Qwen
7.0 GGUF
Q4_K_M, gguf
9.3 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen2.5 Coder 7B Instruct GGUF
Alibaba / Qwen
7.0 GGUF
fp16, GGUF, gguf
19.7 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen2.5 Coder 7B Instruct GGUF
Alibaba / Qwen
7.0 GGUF
q3_k_m, GGUF, gguf
8.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen2.5 7B Instruct INT8 Quanto
Alibaba / Qwen
7.0 INT8
INT8
9.3 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen2.5 7B Instruct Heretic
Alibaba / Qwen
7.0 GGUF
Q4_K_M, gguf
9.3 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen2.5 7B Instruct FP8 Dynamic TR171
Alibaba / Qwen
7.0 FP8
FP8
9.3 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only

Lokales Ollama-Inventar

Status: unavailable · Modelle: 0

Falls Ollama auf der Workstation läuft, werden installierte Modelle mit Größe, Digest, Familie und Quantisierung ausgelesen. Geschwindigkeit wird erst durch Benchmarkläufe gemessen.

API, Workstation oder Edge?

CheckCom nutzt diese Daten, um im KI-Auswahlassistenten lokale Alternativen zu API-/Cloud-Routen sichtbar zu machen. Für Unternehmen zählt nicht nur die Modellgröße, sondern auch Betriebsaufwand, Datenschutz, Latenz, Wartung und Auslastung.

Lokale / quantisierte News- und Release-Signale

DatumAnbieterSignalTypQuelle
2026-08-20Moonshot / KimiKimi K3 and GLM-5.3 are now
Frontier Radar #4: Chinas KI-Modelle schließen auf
local_or_quantized_updateQuelle
2026-08-20OpenAIGPT-5.6
AWS bringt OpenAI-GPT-5.6-Modelle nach Indien
new_model_or_versionQuelle
2026-08-19Google / GeminiGemma 4
Lokale KI schneller machen: So zünden Sie den LLM-Turbo
availability_or_api_updateQuelle
2026-08-19OpenAIGPT
ChatGPT verändert Suchverhalten von Studierenden
availability_or_api_updateQuelle
2026-08-18Google / GeminiGemini AI
Googles Gemini AI scannt Workspace-Daten standardmäßig
local_or_quantized_updateQuelle
2026-08-18OpenAIGPT
OpenAI wechselt zum B2B-Fokus
availability_or_api_updateQuelle
2026-08-18OpenAIQwen3.8-27B
Qwen3.8-27B: 3 Millionen Downloads in drei Tagen
local_or_quantized_updateQuelle
2026-08-17xAI / GrokGrok Bot
SpaceXAI stellt Grok Bot für autonome KI-Agenten vor
new_model_or_versionQuelle
2026-08-16OpenAIGPT-5.6
OpenAI lanciert Ultrafast: GPT-5.6 Sol mit 750 Tokens pro Sekunde
new_model_or_versionQuelle
2026-08-16Alibaba / QwenQwen 3.8 27B
Qwen 3.8 27B: Ausgezeichnet, aber zu überlegend
new_model_or_versionQuelle
2026-08-16Alibaba / QwenQwen3.8-27B Alibaba
Alibaba veröffentlicht Qwen3.8-27B für lokale Nutzung
new_model_or_versionQuelle
2026-08-16OpenAIGPT
OpenAI: 8,3-fache Nutzungslücke bei KI-Agenten in Unternehmen
availability_or_api_updateQuelle
2026-08-15Google / GeminiQwen AI models have surpassed 3 billion global downloads in six months
Alibabas KI-Modelle übertrumpfen Meta und Google
local_or_quantized_updateQuelle
2026-08-14Mistral AIMistral erweitert Angebot Die Organisation
KI-Update: Hate Aid kritisiert KI-Brillen, Mistral erweitert Angebot
local_or_quantized_updateQuelle
2026-08-14OpenAIQwen team
Alibaba veröffentlicht Qwen 3.8 mit offenen Modellgewichten
new_model_or_versionQuelle
2026-08-14OpenAIGPT-5.6
OpenAI lanciert Ultrafast-Modus für GPT-5.6 Sol
new_model_or_versionQuelle
2026-08-14Zhipu / GLMGLM-5.3
Zhipu AI veröffentlicht GLM-5.3 als stärkstes Open-Weights-Coding-Modell
new_model_or_versionQuelle
2026-08-14OpenAIGPT-5.6
OpenAI bringt GPT-5.6 Sol mit neuem Ultrafast-Modus
preview_or_experimentalQuelle
2026-08-13OpenAIGPT-5.6
OpenAI: GPT-5.6 Sol mit 14-facher Geschwindigkeit dank Cerebras
new_model_or_versionQuelle
2026-08-13OpenAIGrok 4.6 iguala a GPT-5.6 Sol con 61 puntos y menor coste para startups Grok 4.6 iguala a GPT-5.6 So
Grok 4.6 erreicht GPT-5.6 Sol-Niveau mit geringeren Kosten
new_model_or_versionQuelle
2026-08-13Mistral AIMistral als KI
Mistral baut KI-Infrastruktur in Europa aus
local_or_quantized_updateQuelle
2026-08-13AnthropicGrok 4.6
Grok 4.6: Effizienzvorteil für Startups
local_or_quantized_updateQuelle
2026-08-13OpenAIGPT-5.6
OpenAI stellt GPT-5.6 Sol mit 14-facher Geschwindigkeit vor
new_model_or_versionQuelle
2026-08-13OpenAIGemini vor
KI-Geisterbeschwörung: Traditionelle Wahrsager in Südkorea verlieren an Bedeutung
availability_or_api_updateQuelle
2026-08-13OpenAIKimi K3
Kimi K3: Open-Source-KI-Modell mit 2,8 Billionen Parametern
availability_or_api_updateQuelle
2026-08-12DeepSeekDeepSeek V4 Pro 0813
DeepSeek V4 Pro 0813 über OpenRouter verfügbar
new_model_or_versionQuelle
2026-08-12OpenAIGrok 4.6 de SpaceXAI
Grok 4.6 von SpaceXAI: 61 Punkte in IA bei 60 % geringerem Kosten als GPT-5.6
new_model_or_versionQuelle
2026-08-12xAI / GrokGrok 4.6
Grok 4.6: xAI schickt neues Modell für autonome KI-Agenten ins Rennen
new_model_or_versionQuelle
2026-08-12Google / GeminiGrok Imagine 2.0
Grok Imagine 2.0: Zweitschnellste KI für Bildgenerierung 2026
new_model_or_versionQuelle
2026-08-12OpenAIGrok Bot as AI agent race shifts toward autonomous work SpaceXAI
SpaceXAI startet Grok Bot im AI-Agenten-Rennen
new_model_or_versionQuelle
2026-08-11Mistral AIMistral AI lanza IA soberana europea
Mistral AI lanciert europäische KI-Infrastruktur mit 1 GW Kapazität
new_model_or_versionQuelle
2026-08-11xAI / GrokGrok Bot as 24
xAI startet Grok Bot als 24/7-Mitarbeiter mit eigenem virtuellen Computer
preview_or_experimentalQuelle
2026-08-11xAI / GrokGrok Bot
SpaceXAI startet Grok Bot: KI-Agenten arbeiten autonom
new_model_or_versionQuelle
2026-08-11Google / GeminiGemini Spark
COM360 lanciert universellen KI-Exekutiv-Assistenten
new_model_or_versionQuelle
2026-08-11OpenAIGPT-5.6
OpenAI lanciert GPT-5.6-Cyber zur Abwehr von KI-Angriffen
new_model_or_versionQuelle
2026-08-10DeepSeekQwen3.8 vs Kimi K3 vs DeepSeek V4
Qwen3.8 vs. Kimi K3 vs. DeepSeek V4: Offene Gewichte kosten ab 20 Millionen
local_or_quantized_updateQuelle
2026-08-10DeepSeekDeepSeek V4-Flash Broke 4-Bit Quantization
DeepSeek V4-Flash Broke 4-Bit Quantization: Q4 Is Only 4% Smaller Than Q8
local_or_quantized_updateQuelle
2026-08-10AnthropicClaude Code
Docker Sandboxes: IA-Agenten wie Claude Code sicher ausführen
new_model_or_versionQuelle
2026-08-10Microsoft / AzurePhi Silica
PowerToys 0.101: Lokale KI für Advanced Paste
preview_or_experimentalQuelle
2026-08-09OpenAIGPT Image
xAI stellt Imagine Image 2.0 vor
new_model_or_versionQuelle

Externe Benchmarkquellen

Diese Quellen dienen zunächst als Orientierung. Öffentliche Leistungswerte werden nur mit Messstatus und Vergleichbarkeit angezeigt.

QuelleTypFokusEinordnung
LocalScore / OpenBenchmarkingexternal_reproducibleGeneration speed, TTFT, Prompt speed, hardware comparisonMethodisch interessante externe Benchmarkquelle. Werte müssen nach Modell, Quantisierung, Runtime und Hardwareprofil abgeglichen werden.
QuelLLM.fr Benchmarksexternal_public_resultRTX 5090, RTX 4090, Mac, CPU, llama.cpp, Q4Nützlich für schnelle Hardwareorientierung. Veröffentlichung auf CheckCom nur mit Kennzeichnung als extern gemessen.
Öffentliche RTX-5090-LLM-Benchmarks / GitHubcommunity_benchmarkRTX 5090, VRAM, Power, Tokens/s, LM Studio / llama.cppSehr relevant für RTX-5090-Klassen, aber je nach Repository unterschiedlich gut dokumentiert.
Ollama lokale APIlocal_metadatainstalled models, size, digest, family, parameter size, quantization levelSehr gut für lokale Modellinventarisierung. Geschwindigkeit entsteht erst durch eigene Laufzeitmessung.
Raspberry Pi AI HAT+ 2 Dokumentationofficial_hardware_docsRaspberry Pi 5, AI HAT+ 2, Hailo-10H, LLM/VLM, Edge AIOffizielle Hardwarequelle für Eignung und Grenzen, keine vollständige Modell-Benchmarktabelle.

KI-Modelle verstehen

Neue LLM-Versionen, Quantisierung, lokale KI, Raspberry Pi, Workstation-Hardware und KI-Kosten verständlich erklärt – mit Links zu den wichtigsten CheckCom-Radaren.

Ratgeber öffnen