CheckCom AI Prices · Local AI

Local & quantized AI models

Hardware fit, quantization, local suitability and measurement status for enterprises. The overview separates metadata, external measurements and future CheckCom-owned benchmarks.

Updated: 22.08.2026 11:17

What this overview does

Local instead of API?

CheckCom helps assess when local workstation, edge or API routes are worth considering.

Quantization in context

Q4, Q5, Q6, Q8, GGUF and related signals are interpreted as hardware and fit indicators.

Measurement status shown

Public values are labelled as metadata, external orientation or future CheckCom-owned measurement.

850
local/quantized candidates
433
providers / families
15210
RTX-5090 candidates
23
Ollama hints
251
local news/release signals
Important: Quantized model and local hardware values are comparable only when model version, quantization, runtime, context length, prompt and hardware profile are documented. V3.1 therefore starts with hardware fit, metadata and external benchmark sources; CheckCom-owned tokens/s values require benchmark runs.

Hardware profiles

HardwareBest forLimitsMeasurement status
RTX 5090 / 32 GB VRAM7B–35B quantisierte LLMs, Coding, lokale RAG-Tests, schnelle lokale Inferenzsehr große 70B+ Modelle nur mit starker Quantisierung/Offload; mehrere parallele Nutzer prüfenbenchmark_prepared
RTX 4090 / 24 GB VRAMkleinere bis mittlere quantisierte Modelle und lokale Experimente32B+ und lange Kontexte häufig knappexternal_and_future_checkcom
RTX Pro / 96 GB VRAMgroße lokale Modelle, längere Kontexte, professionelle lokale Inferenzhohe Anschaffungskosten; TCO prüfenprofile_prepared
Raspberry Pi 5 + AI HAT+ 2 / Hailo-10HEdge-KI, kleine LLM-/VLM-Kandidaten, Vision, lokale Demo- und Datenschutzszenarienkein Ersatz für große GPU-LLMs; Hailo-kompatible Modelle und Runtime nötigdemo_profile_available
Raspberry Pi 5 CPU-onlykleine CPU-Tests, Klassifikation, Offline-Demosgroße LLMs und lange Kontexte nicht sinnvollinventory_prepared
API / Cloud / EnterpriseSkalierung, Enterprise-Verträge, hohe Parallelität und schnelle Produktivstartslaufende Token-/Seat-Kosten, Datenroute und Vertrag prüfenprice_radar_linked

Local and quantized model candidates

Derived from the CheckCom model catalog. measurement_status metadata_only means: no CheckCom benchmark yet.

ModelParams BQuantizationVRAM GBRTX 5090Pi 5 / AI HAT+ 2Score
Qwen2.5-Coder 7B Instruct
Alibaba / Qwen
7.0 Q4
Q4
8.0 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
95/100
metadata_only
Qwen3.6 35B A3B Vocabulary Trimming GGUF
Alibaba / Qwen
35.0 GGUF
Q4_K_S, GGUF, gguf
28.2 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 35B A3B Uncensored xCloud GGUF
Alibaba / Qwen
35.0 GGUF
Q4_K_M, GGUF, gguf
28.2 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 35B A3B TQ2Q GGUF
Alibaba / Qwen
35.0 GGUF
GGUF, gguf
28.2 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 35B A3B StyleTune GGUF
Alibaba / Qwen
35.0 GGUF
IQ4_XS, GGUF, gguf
28.2 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 35B A3B Q5 K M GGUF
Alibaba / Qwen
35.0 GGUF
q5_k_m, Q5, GGUF
32.7 offload
nur mit Offload oder reduzierten Einstellungen realistisch
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 35B A3B Ornith 80 20 Beta GGUF
Alibaba / Qwen
35.0 GGUF
Q4_K_S, GGUF, gguf
28.2 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 35B A3B NVFP4 Gguf
Alibaba / Qwen
35.0 GGUF
Gguf, gguf
28.2 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 35B A3B NVFP4 GGUF
Alibaba / Qwen
35.0 GGUF
q3, GGUF, gguf
22.3 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 35B A3B MTP GGUF Tater NoThink
Alibaba / Qwen
35.0 GGUF
Q4_K_M, GGUF, gguf
28.2 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 35B A3B DS4 GGUF
Alibaba / Qwen
35.0 GGUF
Q4_K_S, GGUF, gguf
28.2 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 35B A3B Claude 4.7 Distill NVFP4 GGUF
Alibaba / Qwen
35.0 GGUF
GGUF, gguf
28.2 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 35B A3B Claude 4.7 Distill MXFP4 MoE GGUF
Alibaba / Qwen
35.0 GGUF
GGUF, gguf
28.2 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.5 35B A3B BitClass3 GGUF
Alibaba / Qwen
35.0 GGUF
Q3_K_S, GGUF, gguf
22.3 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.5 35B A3B APEX GGUF
Alibaba / Qwen
35.0 GGUF
GGUF, gguf
28.2 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Huihui Qwen AgentWorld 35B A3B Abliterated UD Q3 K M GGUF
Alibaba / Qwen
35.0 GGUF
Q3_K_M, Q3, GGUF
22.3 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Endy Qwen3.6 CyberSec 35B A3B GGUF
Alibaba / Qwen
35.0 GGUF
Q2_K, GGUF, gguf
18.7 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3 32B Qwople GGUF
Alibaba / Qwen
32.0 GGUF
Q4_K_M, GGUF, gguf
26.3 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3 32B GGUF
Alibaba / Qwen
32.0 GGUF
IQ3_M, GGUF, gguf
20.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen2.5 Coder 32B Python Specialist I1 GGUF
Alibaba / Qwen
32.0 GGUF
IQ3_M, GGUF, gguf
20.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen2.5 Coder 32B Python Specialist GGUF
Alibaba / Qwen
32.0 GGUF
IQ4_XS, GGUF, gguf
26.3 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen2.5 Coder 32B Instruct GGUF
Alibaba / Qwen
32.0 GGUF
Q2_K, GGUF, gguf
17.7 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen2.5 32B Instruct AdiTurbo GGUF
Alibaba / Qwen
32.0 GGUF
Q3_K_M, GGUF, gguf
20.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
DeepSeek R1 Distill Qwen32B Q4 K M GGUF
Alibaba / Qwen
32.0 GGUF
q4_k_m, Q4, GGUF
26.3 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3 Coder 30B A3B Instruct GGUF
Alibaba / Qwen
30.0 GGUF
Q4_K_M, GGUF, gguf
25.1 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3 30B A3B GGUF
Alibaba / Qwen
30.0 GGUF
IQ3_M, GGUF, gguf
20.0 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Ektome Qwen3 30B A3B PristinelyUncensored GGUF
Alibaba / Qwen
30.0 GGUF
Q2_K, GGUF, gguf
17.0 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.8 27B ZipBrain GGUF
Alibaba / Qwen
27.0 GGUF
IQ4_XS, GGUF, gguf
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.8 27B Uncensored Q4 K M GGUF
Alibaba / Qwen
27.0 GGUF
Q4_K_M, Q4, GGUF
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.8 27B Uncensored Aggressive I1 GGUF
Alibaba / Qwen
27.0 GGUF
IQ3_M, GGUF, gguf
18.6 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.8 27B Uncensored Aggressive GGUF
Alibaba / Qwen
27.0 GGUF
IQ4_XS, GGUF, gguf
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.8 27B ROCmFPX GGUF
Alibaba / Qwen
27.0 GGUF
GGUF, gguf
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.8 27B I1 GGUF
Alibaba / Qwen
27.0 GGUF
IQ3_M, GGUF, gguf
18.6 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.8 27B Fable5 Distill Abliterated Imatrix GGUF
Alibaba / Qwen
27.0 GGUF
GGUF, gguf
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.8 27B Fable Distill GGUF
Alibaba / Qwen
27.0 GGUF
IQ4_XS, GGUF, gguf
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.8 27B Cold Fusion GAIN V1.1 Q4 K M GGUF
Alibaba / Qwen
27.0 GGUF
q4_k_m, Q4, GGUF
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.8 27B Abliterated MTP GGUF
Alibaba / Qwen
27.0 GGUF
IQ2_M, GGUF, gguf
15.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.8 27B 5090 Goldilocks GGUF
Alibaba / Qwen
27.0 GGUF
GGUF, gguf
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 27B Samantha Uncensored GGUF
Alibaba / Qwen
27.0 GGUF
Q4_K_M, GGUF, gguf
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 27B Q4 K M GGUF
Alibaba / Qwen
27.0 GGUF
q4_k_m, Q4, GGUF
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 27B NVFP4 GGUF
Alibaba / Qwen
27.0 GGUF
q3, GGUF, gguf
18.6 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 27B Mtp Gguf
Alibaba / Qwen
27.0 GGUF
Q4_K_X, Gguf, Q4_K_XL
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 27B KR MTP GGUF
Alibaba / Qwen
27.0 GGUF
Q5_K_X, GGUF, Q5_K_XL
26.7 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 27B KR GGUF
Alibaba / Qwen
27.0 GGUF
Q5_K_X, GGUF, Q5_K_XL
26.7 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 27B GGUF
Alibaba / Qwen
27.0 GGUF
Q4_K_M, GGUF, gguf
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 27B Esper4 I1 GGUF
Alibaba / Qwen
27.0 GGUF
IQ3_M, GGUF, gguf
18.6 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 27B Esper4 GGUF
Alibaba / Qwen
27.0 GGUF
IQ4_XS, GGUF, gguf
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 27B Architect Polaris2 Fable B F451 MTP ROCmFPX GGUF
Alibaba / Qwen
27.0 GGUF
Q6, GGUF, gguf
30.8 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 27B AEON RYS Agentic Coder PatchCode GGUF
Alibaba / Qwen
27.0 GGUF
IQ4_NL, GGUF, gguf
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Ostrich 27B Qwen3.8 260815 I1 GGUF
Alibaba / Qwen
27.0 GGUF
IQ3_M, GGUF, gguf
18.6 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Huihui Qwen3.8 27B Abliterated GGUF
Alibaba / Qwen
27.0 GGUF
IQ3_M, GGUF, gguf
18.6 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Elster Vernunft Qwen3.6 27B GGUF
Alibaba / Qwen
27.0 GGUF
IQ4_XS, GGUF, gguf
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Dagger Qwen3.6 27B GGUF MTP
Alibaba / Qwen
27.0 GGUF
Q3_K_M, GGUF, gguf
18.6 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.8 20B Minitron Q4 K M GGUF
Alibaba / Qwen
20.0 GGUF
q4_k_m, Q4, GGUF
18.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 14B A3B VibeForged V2 GGUF
Alibaba / Qwen
14.0 GGUF
F16, GGUF, gguf
20.4 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3 14B PragReST I1 GGUF
Alibaba / Qwen
14.0 GGUF
IQ1_M, GGUF, gguf
20.4 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen2.5 Coder 14B Instruct Jbliterated
Alibaba / Qwen
14.0 GGUF
Q4_K_M, gguf
13.7 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen2.5 Coder 14B Instruct Heretic GGUF
Alibaba / Qwen
14.0 GGUF
IQ4_XS, GGUF, gguf
13.7 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen2.5 Coder 14B Instruct GGUF
Alibaba / Qwen
14.0 GGUF
f16, GGUF, gguf
20.4 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen2.5 Coder 14B Instruct GGUF
Alibaba / Qwen
14.0 GGUF
fp16, GGUF, gguf
34.4 offload
nur mit Offload oder reduzierten Einstellungen realistisch
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen2.5 14B Instruct Heretic
Alibaba / Qwen
14.0 GGUF
Q4_K_M, gguf
13.7 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Lfed Qwen2.5 Coder 14B Sql Gguf
Alibaba / Qwen
14.0 GGUF
Q4_K_M, Gguf, gguf
13.7 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Iol Qwen2.5 14B Instruct AWQ
Alibaba / Qwen
14.0 AWQ
AWQ
13.7 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
H100 Qwen3 14B MonEspaceSante CPT I1 GGUF
Alibaba / Qwen
14.0 GGUF
IQ2_M, GGUF, gguf
9.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
H100 Qwen3 14B MonEspaceSante CPT GGUF
Alibaba / Qwen
14.0 GGUF
IQ4_XS, GGUF, gguf
13.7 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Ektome Qwen2.5 Coder 14B Instruct PristinelyUncensored
Alibaba / Qwen
14.0 GGUF
IQ3_M, gguf
11.3 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Adi Qwen2.5 14B GLM 5.2 General GGUF
Alibaba / Qwen
14.0 GGUF
q4_k_m, GGUF, gguf
13.7 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
QwenPaw Flash 9B Heretic MTP GGUF
Alibaba / Qwen
9.0 GGUF
BF16, GGUF, gguf
23.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
QwenPaw Flash 9B Heretic Imatrix GGUF
Alibaba / Qwen
9.0 GGUF
BF16, GGUF, gguf
23.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
QwenPaw Flash 9B Heretic GGUF
Alibaba / Qwen
9.0 GGUF
BF16, GGUF, gguf
23.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.5 9B Ultra Uncensored Heretic V2 I1 GGUF
Alibaba / Qwen
9.0 GGUF
IQ1_M, GGUF, gguf
14.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.5 9B Ultra Uncensored Heretic V1 I1 GGUF
Alibaba / Qwen
9.0 GGUF
IQ1_M, GGUF, gguf
14.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.5 9B The Defiant Fable Uncensored Heretic NEO IMATRIX MAX MTP GGUF
Alibaba / Qwen
9.0 GGUF
BF16, GGUF, gguf
23.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.5 9B The Defiant Fable Uncensored Heretic NEO IMATRIX MAX GGUF
Alibaba / Qwen
9.0 GGUF
BF16, GGUF, gguf
23.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.5 9B Samantha Uncensored GGUF
Alibaba / Qwen
9.0 GGUF
Q4_K_M, GGUF, gguf
10.6 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.5 9B Heretic Imatrix GGUF
Alibaba / Qwen
9.0 GGUF
BF16, GGUF, gguf
23.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.5 9B Heretic GGUF
Alibaba / Qwen
9.0 GGUF
BF16, GGUF, gguf
23.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.5 9B Haskell Rust Python Q6 K GGUF
Alibaba / Qwen
9.0 GGUF
q6_k, Q6, GGUF
13.1 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.5 9B Haskell Rust Python IQ4 XS GGUF
Alibaba / Qwen
9.0 GGUF
iq4_xs, GGUF, gguf
10.6 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.5 9B GGUF
Alibaba / Qwen
9.0 GGUF
IQ1_KT, GGUF, gguf
14.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.5 9B DeepSeek V4 Flash GGUF
Alibaba / Qwen
9.0 GGUF
Q3_K_M, GGUF, gguf
9.0 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.5 9B DFlash GGUF
Alibaba / Qwen
9.0 GGUF
bf16, DFlash, GGUF
23.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Huihui Qwen3.5 9B Claude 4.6 Opus Abliterated Heretic GGUF
Alibaba / Qwen
9.0 GGUF
f16, GGUF, gguf
14.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
FACET Terminal Qwen3.5 9B GGUF
Alibaba / Qwen
9.0 GGUF
f16, GGUF, gguf
14.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Adi Qwen3.5 9B GLM 5.2 General GGUF
Alibaba / Qwen
9.0 GGUF
q4_k_m, GGUF, gguf
10.6 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
VideoKR Qwen3 VL 8B GGUF
Alibaba / Qwen
8.0 GGUF
f16, GGUF, gguf
13.8 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
SenseNova U1.5 8B GGUFs
Hugging Face / sonstige Owner
8.0 GGUF
Q2_K, gguf
7.8 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
RavenX Conjecture Qwen3 8B GGUF
Alibaba / Qwen
8.0 GGUF
Q8, GGUF, Q8_0
13.8 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3 VL 8B Instruct Heretic GGUF
Alibaba / Qwen
8.0 GGUF
f16, GGUF, gguf
13.8 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3 VL 8B Instruct Abliterated V2 GGUF
Alibaba / Qwen
8.0 GGUF
f16, GGUF, gguf
13.8 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3 Embed 8B GGUF
Alibaba / Qwen
8.0 GGUF
iq4_xs, GGUF, gguf
10.0 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3 8B UnBias Plus SFT Instruct Legacy GGUF
Alibaba / Qwen
8.0 GGUF
f16, GGUF, gguf
13.8 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3 8B PragReST I1 GGUF
Alibaba / Qwen
8.0 GGUF
IQ1_M, GGUF, gguf
13.8 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3 8B PragReST GGUF
Alibaba / Qwen
8.0 GGUF
f16, GGUF, gguf
13.8 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3 8B MedReasonPath GGUF
Alibaba / Qwen
8.0 GGUF
f16, GGUF, gguf
13.8 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3 8B Heretic GGUF
Alibaba / Qwen
8.0 GGUF
f16, GGUF, gguf
13.8 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3 8B Gguf
Alibaba / Qwen
8.0 GGUF
Q8, Gguf, Q8_0
13.8 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3 8B Apostate I1 GGUF
Alibaba / Qwen
8.0 GGUF
IQ1_S, GGUF, gguf
13.8 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3 8B Apostate GGUF
Alibaba / Qwen
8.0 GGUF
f16, GGUF, gguf
13.8 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
PDP Qwen3 8B SFT GGUF
Alibaba / Qwen
8.0 GGUF
f16, GGUF, gguf
13.8 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
LFM2.5 8B A1B GGUF
LFM2.5
8.0 GGUF
BF16, GGUF, gguf
21.8 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
L40S Qwen3 8B MonEspaceSante CPT SFT Anti Hallucination I1 GGUF
Alibaba / Qwen
8.0 GGUF
IQ2_S, GGUF, gguf
7.8 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Huihui Qwen3 VL 8B Instruct Abliterated FP8
Alibaba / Qwen
8.0 FP8
FP8
10.0 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Huihui Qwen3 VL 8B Instruct Abliterated AWQ Int4
Alibaba / Qwen
8.0 AWQ
AWQ, Int4, int4
10.0 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Adi Qwen3 8B GLM 5.2 General GGUF
Alibaba / Qwen
8.0 GGUF
q4_k_m, GGUF, gguf
10.0 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Yoiko Img Prompt Gen Qwen2.5 7B Instruct Q4 K M
Alibaba / Qwen
7.0 GGUF
Q4_K_M, Q4, gguf
9.3 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
VideoKR Qwen2.5 VL 7B GGUF
Alibaba / Qwen
7.0 GGUF
f16, GGUF, gguf
12.7 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Typhoon2 Qwen2.5 7B Instruct GGUF
Alibaba / Qwen
7.0 GGUF
Q4_K_M, GGUF, gguf
9.3 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Stiles Seymens Qwen2.5 7B Generalist Q4KM GGUF
Alibaba / Qwen
7.0 GGUF
Q4_K_M, GGUF, gguf
9.3 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Sofia Qwen2.5 7B I1 GGUF
Alibaba / Qwen
7.0 GGUF
IQ1_M, GGUF, gguf
12.7 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Sofia Qwen2.5 7B GGUF
Alibaba / Qwen
7.0 GGUF
f16, GGUF, gguf
12.7 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen2.5 VL 7B Instruct Int8 Convrot Comfyui
Alibaba / Qwen
7.0 Int8
Int8, int8
9.3 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen2.5 Coder 7B VN Master Polymath FINAL GGUF
Alibaba / Qwen
7.0 GGUF
Q3_K_M, GGUF, gguf
8.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen2.5 Coder 7B Instruct Jbliterated
Alibaba / Qwen
7.0 GGUF
Q4_K_M, gguf
9.3 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen2.5 Coder 7B Instruct Heretic
Alibaba / Qwen
7.0 GGUF
Q4_K_M, gguf
9.3 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen2.5 Coder 7B Instruct GGUF
Alibaba / Qwen
7.0 GGUF
fp16, GGUF, gguf
19.7 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen2.5 Coder 7B Instruct GGUF
Alibaba / Qwen
7.0 GGUF
q3_k_m, GGUF, gguf
8.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen2.5 7B Instruct INT8 Quanto
Alibaba / Qwen
7.0 INT8
INT8
9.3 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen2.5 7B Instruct Heretic
Alibaba / Qwen
7.0 GGUF
Q4_K_M, gguf
9.3 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen2.5 7B Instruct FP8 Dynamic TR171
Alibaba / Qwen
7.0 FP8
FP8
9.3 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only

Local Ollama inventory

Status: unavailable · models: 0

If Ollama is running, installed models are inventoried with size, digest, family and quantization. Speed requires benchmark runs.

API, workstation or edge?

CheckCom uses this data to surface local alternatives to API/cloud routes in the AI advisor. For enterprises, model size is only one factor; operations, privacy, latency, maintenance and utilization matter as well.

Local / quantized news and release signals

DateProviderSignalTypeSource
2026-08-20Moonshot / KimiKimi K3 and GLM-5.3 are now
Frontier Radar #4: China's AI models are catching up
local_or_quantized_updateQuelle
2026-08-20OpenAIGPT-5.6
AWS brings OpenAI's GPT-5.6 models to India
new_model_or_versionQuelle
2026-08-19Google / GeminiGemma 4
Boost Local AI Performance with LLM Turbo
availability_or_api_updateQuelle
2026-08-19OpenAIGPT
ChatGPT is reshaping student research habits
availability_or_api_updateQuelle
2026-08-18Google / GeminiGemini AI
Google's Gemini AI scans Workspace data by default
local_or_quantized_updateQuelle
2026-08-18OpenAIGPT
OpenAI shifts to B2B focus
availability_or_api_updateQuelle
2026-08-18OpenAIQwen3.8-27B
Qwen3.8-27B: 3 million downloads in three days
local_or_quantized_updateQuelle
2026-08-17xAI / GrokGrok Bot
SpaceXAI Launches Grok Bot for Autonomous AI Agents
new_model_or_versionQuelle
2026-08-16OpenAIGPT-5.6
OpenAI Launches Ultrafast: GPT-5.6 Sol at 750 Tokens/Second
new_model_or_versionQuelle
2026-08-16Alibaba / QwenQwen 3.8 27B
Qwen 3.8 27B is excellent, but overthinks by default
new_model_or_versionQuelle
2026-08-16Alibaba / QwenQwen3.8-27B Alibaba
Alibaba releases Qwen3.8-27B for local use
new_model_or_versionQuelle
2026-08-16OpenAIGPT
OpenAI: 8.3x gap in enterprise AI agent usage
availability_or_api_updateQuelle
2026-08-15Google / GeminiQwen AI models have surpassed 3 billion global downloads in six months
Alibaba's AI models outpace Meta, Google to hit 3B downloads
local_or_quantized_updateQuelle
2026-08-14Mistral AIMistral erweitert Angebot Die Organisation
AI Update: Hate Aid criticizes AI glasses, Mistral expands offerings
local_or_quantized_updateQuelle
2026-08-14OpenAIQwen team
Alibaba releases Qwen 3.8 with open model weights
new_model_or_versionQuelle
2026-08-14OpenAIGPT-5.6
OpenAI launches Ultrafast mode for GPT-5.6 Sol
new_model_or_versionQuelle
2026-08-14Zhipu / GLMGLM-5.3
Zhipu AI releases GLM-5.3 as strongest open-weights coding model
new_model_or_versionQuelle
2026-08-14OpenAIGPT-5.6
OpenAI introduces Ultrafast mode for GPT-5.6 Sol
preview_or_experimentalQuelle
2026-08-13OpenAIGPT-5.6
OpenAI: GPT-5.6 Sol 14x Faster with Cerebras
new_model_or_versionQuelle
2026-08-13OpenAIGrok 4.6 iguala a GPT-5.6 Sol con 61 puntos y menor coste para startups Grok 4.6 iguala a GPT-5.6 So
Grok 4.6 matches GPT-5.6 Sol at lower cost
new_model_or_versionQuelle
2026-08-13Mistral AIMistral als KI
Mistral expands AI infrastructure in Europe
local_or_quantized_updateQuelle
2026-08-13AnthropicGrok 4.6
Grok 4.6: Efficiency Edge for Startups
local_or_quantized_updateQuelle
2026-08-13OpenAIGPT-5.6
OpenAI unveils GPT-5.6 Sol with 14x faster inference
new_model_or_versionQuelle
2026-08-13OpenAIGemini vor
South Korea’s blind fortune-tellers eclipsed by AI
availability_or_api_updateQuelle
2026-08-13OpenAIKimi K3
Kimi K3: Open-Source AI Model with 2.8 Trillion Parameters
availability_or_api_updateQuelle
2026-08-12DeepSeekDeepSeek V4 Pro 0813
DeepSeek V4 Pro 0813 available via OpenRouter
new_model_or_versionQuelle
2026-08-12OpenAIGrok 4.6 de SpaceXAI
SpaceXAI's Grok 4.6: 61 AI Points at 60% Lower Cost Than GPT-5.6
new_model_or_versionQuelle
2026-08-12xAI / GrokGrok 4.6
Grok 4.6: xAI Launches New Model for Autonomous AI Agents
new_model_or_versionQuelle
2026-08-12Google / GeminiGrok Imagine 2.0
Grok Imagine 2.0: Second-Best AI Image Generator 2026
new_model_or_versionQuelle
2026-08-12OpenAIGrok Bot as AI agent race shifts toward autonomous work SpaceXAI
SpaceXAI launches Grok Bot in AI agent race
new_model_or_versionQuelle
2026-08-11Mistral AIMistral AI lanza IA soberana europea
Mistral AI launches European AI infrastructure with 1 GW capacity
new_model_or_versionQuelle
2026-08-11xAI / GrokGrok Bot as 24
xAI launches Grok Bot as 24/7 coworker with its own virtual computer
preview_or_experimentalQuelle
2026-08-11xAI / GrokGrok Bot
SpaceXAI Launches Grok Bot: AI Agents Work Autonomously
new_model_or_versionQuelle
2026-08-11Google / GeminiGemini Spark
COM360 Launches Universal AI Executive Assistant
new_model_or_versionQuelle
2026-08-11OpenAIGPT-5.6
OpenAI launches GPT-5.6-Cyber to combat AI-driven attacks
new_model_or_versionQuelle
2026-08-10DeepSeekQwen3.8 vs Kimi K3 vs DeepSeek V4
Qwen3.8 vs Kimi K3 vs DeepSeek V4: Open Weights Stopped Being Free at $20 Million
local_or_quantized_updateQuelle
2026-08-10DeepSeekDeepSeek V4-Flash Broke 4-Bit Quantization
DeepSeek V4-Flash Broke 4-Bit Quantization: Q4 Is Only 4% Smaller Than Q8
local_or_quantized_updateQuelle
2026-08-10AnthropicClaude Code
Docker Sandboxes: Run AI agents like Claude Code securely
new_model_or_versionQuelle
2026-08-10Microsoft / AzurePhi Silica
PowerToys 0.101 Preview: Local AI for Advanced Paste
preview_or_experimentalQuelle
2026-08-09OpenAIGPT Image
xAI Launches Imagine Image 2.0
new_model_or_versionQuelle

External benchmark sources

These sources are used as orientation. Public performance values are shown only with measurement status and comparability.

SourceTypeFocusAssessment
LocalScore / OpenBenchmarkingexternal_reproducibleGeneration speed, TTFT, Prompt speed, hardware comparisonMethodically useful external benchmark source. Values must be matched by model, quantization, runtime and hardware profile.
QuelLLM.fr Benchmarksexternal_public_resultRTX 5090, RTX 4090, Mac, CPU, llama.cpp, Q4Useful for hardware orientation. CheckCom displays it only as externally measured orientation.
Öffentliche RTX-5090-LLM-Benchmarks / GitHubcommunity_benchmarkRTX 5090, VRAM, Power, Tokens/s, LM Studio / llama.cppHighly relevant for RTX-5090-class systems, but documentation quality varies by repository.
Ollama lokale APIlocal_metadatainstalled models, size, digest, family, parameter size, quantization levelVery useful for local model inventory. Speed requires a separate local benchmark run.
Raspberry Pi AI HAT+ 2 Dokumentationofficial_hardware_docsRaspberry Pi 5, AI HAT+ 2, Hailo-10H, LLM/VLM, Edge AIOfficial hardware source for suitability and limits, not a complete model benchmark table.

Understanding AI models

New LLM versions, quantization, local AI, Raspberry Pi, workstation hardware and AI costs explained — with links to the key CheckCom radars.

Open guide