CheckCom AI Prices · Local AI

Local & quantized AI models

Hardware fit, quantization, local suitability and measurement status for enterprises. The overview separates metadata, external measurements and future CheckCom-owned benchmarks.

Updated: 07.10.2026 05:10

What this overview does

Local instead of API?

CheckCom helps assess when local workstation, edge or API routes are worth considering.

Quantization in context

Q4, Q5, Q6, Q8, GGUF and related signals are interpreted as hardware and fit indicators.

Measurement status shown

Public values are labelled as metadata, external orientation or future CheckCom-owned measurement.

850
local/quantized candidates
449
providers / families
21991
RTX-5090 candidates
23
Ollama hints
279
local news/release signals
Important: Quantized model and local hardware values are comparable only when model version, quantization, runtime, context length, prompt and hardware profile are documented. V3.1 therefore starts with hardware fit, metadata and external benchmark sources; CheckCom-owned tokens/s values require benchmark runs.

Hardware profiles

HardwareBest forLimitsMeasurement status
RTX 5090 / 32 GB VRAM7B–35B quantisierte LLMs, Coding, lokale RAG-Tests, schnelle lokale Inferenzsehr große 70B+ Modelle nur mit starker Quantisierung/Offload; mehrere parallele Nutzer prüfenbenchmark_prepared
RTX 4090 / 24 GB VRAMkleinere bis mittlere quantisierte Modelle und lokale Experimente32B+ und lange Kontexte häufig knappexternal_and_future_checkcom
RTX Pro / 96 GB VRAMgroße lokale Modelle, längere Kontexte, professionelle lokale Inferenzhohe Anschaffungskosten; TCO prüfenprofile_prepared
Raspberry Pi 5 + AI HAT+ 2 / Hailo-10HEdge-KI, kleine LLM-/VLM-Kandidaten, Vision, lokale Demo- und Datenschutzszenarienkein Ersatz für große GPU-LLMs; Hailo-kompatible Modelle und Runtime nötigdemo_profile_available
Raspberry Pi 5 CPU-onlykleine CPU-Tests, Klassifikation, Offline-Demosgroße LLMs und lange Kontexte nicht sinnvollinventory_prepared
API / Cloud / EnterpriseSkalierung, Enterprise-Verträge, hohe Parallelität und schnelle Produktivstartslaufende Token-/Seat-Kosten, Datenroute und Vertrag prüfenprice_radar_linked

Local and quantized model candidates

Derived from the CheckCom model catalog. measurement_status metadata_only means: no CheckCom benchmark yet.

ModelParams BQuantizationVRAM GBRTX 5090Pi 5 / AI HAT+ 2Score
Qwen2.5-Coder 7B Instruct
Alibaba / Qwen
7.0 Q4
Q4
8.0 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
95/100
metadata_only
Qwen3.8 Next 40B Exp MoE Healed GGUF
Alibaba / Qwen
40.0 GGUF
Q4_K_M, GGUF, gguf
31.3 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen35B A3B SignOfFour Coder GGUF
Alibaba / Qwen
35.0 GGUF
Q2_K, GGUF, gguf
18.7 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 35B A3B Vocabulary Trimming GGUF
Alibaba / Qwen
35.0 GGUF
Q4_K_S, GGUF, gguf
28.2 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 35B A3B Uncensored xCloud GGUF
Alibaba / Qwen
35.0 GGUF
Q4_K_M, GGUF, gguf
28.2 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 35B A3B TQ2Q GGUF
Alibaba / Qwen
35.0 GGUF
GGUF, gguf
28.2 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 35B A3B StyleTune GGUF
Alibaba / Qwen
35.0 GGUF
IQ4_XS, GGUF, gguf
28.2 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 35B A3B Q5 K M GGUF
Alibaba / Qwen
35.0 GGUF
q5_k_m, Q5, GGUF
32.7 offload
nur mit Offload oder reduzierten Einstellungen realistisch
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 35B A3B Ornith 80 20 Beta GGUF
Alibaba / Qwen
35.0 GGUF
Q4_K_S, GGUF, gguf
28.2 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 35B A3B NVFP4 Gguf
Alibaba / Qwen
35.0 GGUF
Gguf, gguf
28.2 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 35B A3B NVFP4 GGUF
Alibaba / Qwen
35.0 GGUF
q3, GGUF, gguf
22.3 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 35B A3B MTP GGUF Tater NoThink
Alibaba / Qwen
35.0 GGUF
Q4_K_M, GGUF, gguf
28.2 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 35B A3B DS4 GGUF
Alibaba / Qwen
35.0 GGUF
Q4_K_S, GGUF, gguf
28.2 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 35B A3B Claude 4.7 Distill NVFP4 GGUF
Alibaba / Qwen
35.0 GGUF
GGUF, gguf
28.2 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 35B A3B Claude 4.7 Distill MXFP4 MoE GGUF
Alibaba / Qwen
35.0 GGUF
GGUF, gguf
28.2 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 35B A3B Abliterated GGUF
Alibaba / Qwen
35.0 GGUF
IQ2_M, GGUF, gguf
18.7 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.5 35B A3B BitClass3 GGUF
Alibaba / Qwen
35.0 GGUF
Q3_K_S, GGUF, gguf
22.3 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.5 35B A3B APEX GGUF
Alibaba / Qwen
35.0 GGUF
GGUF, gguf
28.2 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Huihui Qwen AgentWorld 35B A3B Abliterated UD Q3 K M GGUF
Alibaba / Qwen
35.0 GGUF
Q3_K_M, Q3, GGUF
22.3 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
GraphForge Qwen3.6 35B A3B SFT I1 GGUF
Alibaba / Qwen
35.0 GGUF
IQ3_M, GGUF, gguf
22.3 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
GraphForge Qwen3.6 35B A3B SFT GGUF
Alibaba / Qwen
35.0 GGUF
IQ4_XS, GGUF, gguf
28.2 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Endy Qwen3.6 CyberSec 35B A3B GGUF
Alibaba / Qwen
35.0 GGUF
Q2_K, GGUF, gguf
18.7 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3 32B Qwople GGUF
Alibaba / Qwen
32.0 GGUF
Q4_K_M, GGUF, gguf
26.3 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3 32B GGUF
Alibaba / Qwen
32.0 GGUF
IQ3_M, GGUF, gguf
20.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen2.5 Coder 32B Python Specialist I1 GGUF
Alibaba / Qwen
32.0 GGUF
IQ3_M, GGUF, gguf
20.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen2.5 Coder 32B Python Specialist GGUF
Alibaba / Qwen
32.0 GGUF
IQ4_XS, GGUF, gguf
26.3 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen2.5 Coder 32B Instruct Jbliterated
Alibaba / Qwen
32.0 GGUF
Q4_K_M, gguf
26.3 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen2.5 Coder 32B Instruct GGUF
Alibaba / Qwen
32.0 GGUF
Q2_K, GGUF, gguf
17.7 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen2.5 32B Instruct AdiTurbo GGUF
Alibaba / Qwen
32.0 GGUF
Q3_K_M, GGUF, gguf
20.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
DeepSeek R1 Distill Qwen32B Q4 K M GGUF
Alibaba / Qwen
32.0 GGUF
q4_k_m, Q4, GGUF
26.3 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3 Coder 30B A3B Instruct GGUF
Alibaba / Qwen
30.0 GGUF
Q4_K_M, GGUF, gguf
25.1 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3 30B A3B Vietnamese Instruct GGUF
Alibaba / Qwen
30.0 GGUF
Q4_K_M, GGUF, gguf
25.1 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3 30B A3B GGUF
Alibaba / Qwen
30.0 GGUF
IQ3_M, GGUF, gguf
20.0 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Ektome Qwen3 30B A3B PristinelyUncensored GGUF
Alibaba / Qwen
30.0 GGUF
Q2_K, GGUF, gguf
17.0 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
ThinkingCap Qwen3.8 27B I1 GGUF
Alibaba / Qwen
27.0 GGUF
IQ3_M, GGUF, gguf
18.6 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Swift Qwen3.8 27B Uncensored MTP Terse Coder I1 GGUF
Alibaba / Qwen
27.0 GGUF
IQ3_M, GGUF, gguf
18.6 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Swift 1.5 Qwen3.8 27B I1 GGUF
Alibaba / Qwen
27.0 GGUF
IQ3_M, GGUF, gguf
18.6 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Swift 1.5 Qwen3.8 27B Heretic I1 GGUF
Alibaba / Qwen
27.0 GGUF
IQ3_M, GGUF, gguf
18.6 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Swift 1.5 Qwen3.8 27B Heretic GSQ RCO GGUF
Alibaba / Qwen
27.0 GGUF
IQ2_S, GGUF, gguf
15.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Swift 1.5 Qwen3.8 27B GGUF
Alibaba / Qwen
27.0 GGUF
IQ4_XS, GGUF, gguf
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.8 27B ZipBrain GGUF
Alibaba / Qwen
27.0 GGUF
IQ4_XS, GGUF, gguf
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.8 27B Uncensored Q4 K M GGUF
Alibaba / Qwen
27.0 GGUF
Q4_K_M, Q4, GGUF
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.8 27B Uncensored Q4 K M GGUF
Alibaba / Qwen
27.0 GGUF
Q4_K_M, Q4, GGUF
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.8 27B Uncensored Aggressive I1 GGUF
Alibaba / Qwen
27.0 GGUF
IQ3_M, GGUF, gguf
18.6 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.8 27B Uncensored Aggressive GGUF
Alibaba / Qwen
27.0 GGUF
IQ4_XS, GGUF, gguf
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.8 27B Terse Coder GGUF
Alibaba / Qwen
27.0 GGUF
q4_k_m, GGUF, gguf
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.8 27B TURBO Fable Cold Fusion 735 882 Heretic Uncensored NM DAU NVFP4 GGUF
Alibaba / Qwen
27.0 GGUF
GGUF, gguf
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.8 27B ROCmFPX GGUF
Alibaba / Qwen
27.0 GGUF
GGUF, gguf
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.8 27B MindMeld I1 GGUF
Alibaba / Qwen
27.0 GGUF
IQ3_M, GGUF, gguf
18.6 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.8 27B I1 GGUF
Alibaba / Qwen
27.0 GGUF
IQ3_M, GGUF, gguf
18.6 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.8 27B Human KO Enterprise Boundary I1 GGUF
Alibaba / Qwen
27.0 GGUF
IQ3_M, GGUF, gguf
18.6 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.8 27B Fable5 Distill Abliterated Imatrix GGUF
Alibaba / Qwen
27.0 GGUF
GGUF, gguf
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.8 27B Fable Distill GGUF
Alibaba / Qwen
27.0 GGUF
IQ4_XS, GGUF, gguf
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.8 27B Cold Fusion GAIN V1.1 Q4 K M GGUF
Alibaba / Qwen
27.0 GGUF
q4_k_m, Q4, GGUF
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.8 27B Abliterated SFT I1 GGUF
Alibaba / Qwen
27.0 GGUF
IQ3_M, GGUF, gguf
18.6 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.8 27B Abliterated MTP GGUF
Alibaba / Qwen
27.0 GGUF
IQ2_M, GGUF, gguf
15.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.8 27B 5090 Goldilocks GGUF
Alibaba / Qwen
27.0 GGUF
GGUF, gguf
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 27B Samantha Uncensored GGUF
Alibaba / Qwen
27.0 GGUF
Q4_K_M, GGUF, gguf
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 27B Q4 K M GGUF
Alibaba / Qwen
27.0 GGUF
q4_k_m, Q4, GGUF
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 27B NVFP4 GGUF
Alibaba / Qwen
27.0 GGUF
q3, GGUF, gguf
18.6 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 27B Mtp Gguf
Alibaba / Qwen
27.0 GGUF
Q4_K_X, Gguf, Q4_K_XL
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 27B KR MTP GGUF
Alibaba / Qwen
27.0 GGUF
Q5_K_X, GGUF, Q5_K_XL
26.7 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 27B KR GGUF
Alibaba / Qwen
27.0 GGUF
Q5_K_X, GGUF, Q5_K_XL
26.7 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 27B GGUF
Alibaba / Qwen
27.0 GGUF
Q4_K_M, GGUF, gguf
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 27B Esper4 I1 GGUF
Alibaba / Qwen
27.0 GGUF
IQ3_M, GGUF, gguf
18.6 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 27B Esper4 GGUF
Alibaba / Qwen
27.0 GGUF
IQ4_XS, GGUF, gguf
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 27B Architect Polaris2 Fable B F451 MTP ROCmFPX GGUF
Alibaba / Qwen
27.0 GGUF
Q6, GGUF, gguf
30.8 tight
voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 27B AEON RYS Agentic Coder PatchCode GGUF
Alibaba / Qwen
27.0 GGUF
IQ4_NL, GGUF, gguf
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Ostrich 27B Qwen3.8 260815 I1 GGUF
Alibaba / Qwen
27.0 GGUF
IQ3_M, GGUF, gguf
18.6 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Medgap Qwen3.8 27B GGUF
Alibaba / Qwen
27.0 GGUF
IQ4_XS, GGUF, gguf
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Loes Qwen3.8 27B I1 GGUF
Alibaba / Qwen
27.0 GGUF
IQ3_M, GGUF, gguf
18.6 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Huihui Qwen3.8 27B Abliterated GGUF
Alibaba / Qwen
27.0 GGUF
IQ3_M, GGUF, gguf
18.6 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
GraphForge Qwen3.6 27B SFT GGUF
Alibaba / Qwen
27.0 GGUF
IQ4_XS, GGUF, gguf
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Elster Vernunft Qwen3.6 27B GGUF
Alibaba / Qwen
27.0 GGUF
IQ4_XS, GGUF, gguf
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Dagger Qwen3.6 27B GGUF MTP
Alibaba / Qwen
27.0 GGUF
Q3_K_M, GGUF, gguf
18.6 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
ATX Swift 1.5 Qwen3.8 27B Uncensored MTP GGUF
Alibaba / Qwen
27.0 GGUF
Q4_K_M, GGUF, gguf
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
ATX Swift 1.5 Qwen3.8 27B Uncensored IQ4 XS M GGUF
Alibaba / Qwen
27.0 GGUF
IQ4_XS, GGUF, gguf
23.2 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.8 20B Minitron Q4 K M GGUF
Alibaba / Qwen
20.0 GGUF
q4_k_m, Q4, GGUF
18.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Lumynax Reasoning DeepSeek R1 Qwen15b Gguf
Alibaba / Qwen
15.0 GGUF
Q4_K_M, Gguf, gguf
15.8 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.6 14B A3B VibeForged V2 GGUF
Alibaba / Qwen
14.0 GGUF
F16, GGUF, gguf
20.4 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3 14B Uncensored GGUF
Alibaba / Qwen
14.0 GGUF
IQ4_XS, GGUF, gguf
13.7 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3 14B PragReST I1 GGUF
Alibaba / Qwen
14.0 GGUF
IQ1_M, GGUF, gguf
20.4 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen2.5 Coder 14B Instruct Uncensored Q4 K M GGUF
Alibaba / Qwen
14.0 GGUF
q4_k_m, Q4, GGUF
13.7 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen2.5 Coder 14B Instruct Heretic GGUF
Alibaba / Qwen
14.0 GGUF
IQ4_XS, GGUF, gguf
13.7 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen2.5 Coder 14B Instruct GGUF
Alibaba / Qwen
14.0 GGUF
f16, GGUF, gguf
20.4 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen2.5 Coder 14B Instruct GGUF
Alibaba / Qwen
14.0 GGUF
fp16, GGUF, gguf
34.4 offload
nur mit Offload oder reduzierten Einstellungen realistisch
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen2.5 14B Instruct KanColle Hibiki Verniy GGUF
Alibaba / Qwen
14.0 GGUF
F16, GGUF, gguf
20.4 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen2.5 14B Instruct Heretic
Alibaba / Qwen
14.0 GGUF
F16, gguf, Q4_K_M
13.7 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen Qwen2.5 Coder 14B Instruct GGUF Q8 0
Alibaba / Qwen
14.0 GGUF
Q8, GGUF, Q8_0
20.4 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen Qwen2.5 Coder 14B Instruct GGUF Q6 K
Alibaba / Qwen
14.0 GGUF
Q6_K, GGUF, Q6
17.6 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen Qwen2.5 Coder 14B Instruct GGUF Q4 K M
Alibaba / Qwen
14.0 GGUF
Q4_K_M, GGUF, Q4
13.7 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen Qwen2.5 Coder 14B Instruct GGUF Q3 K M
Alibaba / Qwen
14.0 GGUF
Q3_K_M, GGUF, Q3
11.3 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Lfed Qwen2.5 Coder 14B Sql Gguf
Alibaba / Qwen
14.0 GGUF
Q4_K_M, Gguf, gguf
13.7 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Iol Qwen2.5 14B Instruct AWQ
Alibaba / Qwen
14.0 AWQ
AWQ
13.7 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
H100 Qwen3 14B MonEspaceSante CPT I1 GGUF
Alibaba / Qwen
14.0 GGUF
IQ2_M, GGUF, gguf
9.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
H100 Qwen3 14B MonEspaceSante CPT GGUF
Alibaba / Qwen
14.0 GGUF
IQ4_XS, GGUF, gguf
13.7 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Ektome Qwen2.5 Coder 14B Instruct PristinelyUncensored
Alibaba / Qwen
14.0 GGUF
IQ3_M, gguf
11.3 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Adi Qwen2.5 14B GLM 5.2 General GGUF
Alibaba / Qwen
14.0 GGUF
q4_k_m, GGUF, gguf
13.7 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
SkillGym Qwen3.5 9B GGUF
Alibaba / Qwen
9.0 GGUF
f16, GGUF, gguf
14.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
QwenPaw Flash 9B Heretic MTP GGUF
Alibaba / Qwen
9.0 GGUF
BF16, GGUF, gguf
23.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
QwenPaw Flash 9B Heretic Imatrix GGUF
Alibaba / Qwen
9.0 GGUF
BF16, GGUF, gguf
23.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
QwenPaw Flash 9B Heretic GGUF
Alibaba / Qwen
9.0 GGUF
BF16, GGUF, gguf
23.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.5 9B Ultra Uncensored Heretic V2 I1 GGUF
Alibaba / Qwen
9.0 GGUF
IQ1_M, GGUF, gguf
14.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.5 9B Ultra Uncensored Heretic V1 I1 GGUF
Alibaba / Qwen
9.0 GGUF
IQ1_M, GGUF, gguf
14.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.5 9B The Defiant Fable Uncensored Heretic NEO IMATRIX MAX MTP GGUF
Alibaba / Qwen
9.0 GGUF
BF16, GGUF, gguf
23.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.5 9B The Defiant Fable Uncensored Heretic NEO IMATRIX MAX GGUF
Alibaba / Qwen
9.0 GGUF
BF16, GGUF, gguf
23.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.5 9B TAP DPQ V6 GGUF
Alibaba / Qwen
9.0 GGUF
IQ1_S, GGUF, gguf
14.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.5 9B Samantha Uncensored GGUF
Alibaba / Qwen
9.0 GGUF
Q4_K_M, GGUF, gguf
10.6 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.5 9B Heretic Imatrix GGUF
Alibaba / Qwen
9.0 GGUF
BF16, GGUF, gguf
23.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.5 9B Heretic GGUF
Alibaba / Qwen
9.0 GGUF
BF16, GGUF, gguf
23.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.5 9B Haskell Rust Python Q6 K GGUF
Alibaba / Qwen
9.0 GGUF
q6_k, Q6, GGUF
13.1 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.5 9B Haskell Rust Python IQ4 XS GGUF
Alibaba / Qwen
9.0 GGUF
iq4_xs, GGUF, gguf
10.6 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.5 9B GGUF
Alibaba / Qwen
9.0 GGUF
BF16, GGUF, gguf
23.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.5 9B GGUF
Alibaba / Qwen
9.0 GGUF
IQ1_KT, GGUF, gguf
14.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.5 9B DeepSeek V4 Flash GGUF
Alibaba / Qwen
9.0 GGUF
Q3_K_M, GGUF, gguf
9.0 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.5 9B DFlash GGUF
Alibaba / Qwen
9.0 GGUF
bf16, DFlash, GGUF
23.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.5 9B Brainwaves GGUF
Alibaba / Qwen
9.0 GGUF
F16, GGUF, gguf
14.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
Qwen3.5 9B Base GGUF
Alibaba / Qwen
9.0 GGUF
Q4_K_S, GGUF, gguf
10.6 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
MiMo V2.6 Distill Qwen9B GGUF
Alibaba / Qwen
9.0 GGUF
IQ1_M, GGUF, gguf
14.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only
MiMo V2.6 Distill Qwen9B Ablitrated I1 GGUF
Alibaba / Qwen
9.0 GGUF
IQ1_M, GGUF, gguf
14.9 good
passt voraussichtlich gut in das VRAM-Profil
not_recommended
für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt
85/100
metadata_only

Local Ollama inventory

Status: ok · models: 22

If Ollama is running, installed models are inventoried with size, digest, family and quantization. Speed requires benchmark runs.

API, workstation or edge?

CheckCom uses this data to surface local alternatives to API/cloud routes in the AI advisor. For enterprises, model size is only one factor; operations, privacy, latency, maintenance and utilization matter as well.

Local / quantized news and release signals

DateProviderSignalTypeSource
2026-10-05Zhipu / GLMGLM-5.2
Reflection unveils Beam, a new open-weight AI model
new_model_or_versionQuelle
2026-10-05DeepSeekDeepSeek y Qwen Reflection AI presentó este lunes Beam
Reflection AI launches Beam, the open model challenging China
new_model_or_versionQuelle
2026-10-05AnthropicClaude Code
devpit
local_or_quantized_updateQuelle
2026-10-03DeepSeekDeepSeek V4 vs Qwen3.8 vs Kimi K3 vs GLM-5.3 I skipped the vendor tables and lined up the independen
DeepSeek V4 vs Qwen3.8 vs Kimi K3 vs GLM-5.3: Open-Weight LLMs Compared
local_or_quantized_updateQuelle
2026-10-03AnthropicClaude to
Open-source 'BootLoops' tool supports AI in precise scientific calculations
local_or_quantized_updateQuelle
2026-10-02DeepSeekDeepSeek V4 vs Qwen3.8 vs Kimi K3 vs GLM-5.3 https
Best Open-Weight LLMs Tested: DeepSeek V4 vs Qwen3.8 vs Kimi K3 vs GLM-5.3
local_or_quantized_updateQuelle
2026-10-02OpenAIGPT-6
OpenAI's GPT-6 Astra Ultrafast Achieves 8x Speed Boost on NVIDIA
new_model_or_versionQuelle
2026-10-01Google / GeminiGemini AI
Judge dismisses antitrust lawsuits against Google AI search
local_or_quantized_updateQuelle
2026-10-01OpenAIGPT-6.1
OpenAI unveils GPT-6.1 Sol and Dots
new_model_or_versionQuelle
2026-09-30Google / GeminiGemini 4
Google unveils Gemini 4 Argon with 1M tokens
new_model_or_versionQuelle
2026-09-30Google / GeminiGemini 4
Google unveils Gemini 4 Argon AI model – not available yet
new_model_or_versionQuelle
2026-09-30OpenAIGPT-6.1
OpenAI unveils Dots and GPT-6.1 Sol at DevDay 2026
new_model_or_versionQuelle
2026-09-30OpenAIGPT-6
OpenAI launches Dots: AI agents competing with Meta's Muse
new_model_or_versionQuelle
2026-09-30OpenAIGPT-6.1
OpenAI unveils GPT-6.1 Sol and Dots AI agents
new_model_or_versionQuelle
2026-09-30AnthropicGLM-5.3
Anthropic: Zhipu's GLM-5.3 nearly matches Claude Mythos in exploit building
preview_or_experimentalQuelle
2026-09-30OpenAIGPT-6.1
OpenAI cancels GPT-6.1 Astra over security issues
new_model_or_versionQuelle
2026-09-30OpenAIGemini Spark
OpenAI launches personal AI assistant 'dots'
new_model_or_versionQuelle
2026-09-29OpenAIGPT-6.1
OpenAI launches GPT-6.1 Sol at one-fifth the price
new_model_or_versionQuelle
2026-09-29OpenAIGPT-6.1
OpenAI launches GPT-6.1 Sol with improved performance and lower costs
new_model_or_versionQuelle
2026-09-29OpenAIGPT-6.1
OpenAI cancels release of GPT-6.1 Astra model
new_model_or_versionQuelle
2026-09-29OpenAIGPT-6.1
GPT-6.1 Sol Closes in on Astra at a Fifth of the Cost
availability_or_api_updateQuelle
2026-09-29OpenAIGPT-6
OpenAI halves API credits in Pro plan to push pay-per-use model
availability_or_api_updateQuelle
2026-09-29OpenAIGPT-6
OpenAI releases GPT-6.1 with significantly reduced pricing
new_model_or_versionQuelle
2026-09-29OpenAIGPT 6.1
OpenAI cancels release of new AI model
new_model_or_versionQuelle
2026-09-29OpenAIGPT-6.1
OpenAI unveils GPT-6.1 Sol and Dots AI agents
new_model_or_versionQuelle
2026-09-29OpenAIGPT-6.1
OpenAI halts GPT-6.1 Astra model over safety concerns
new_model_or_versionQuelle
2026-09-29Mistral AIMistral eröffnet Hub in München
Mistral opens AI hub in Munich
local_or_quantized_updateQuelle
2026-09-29OpenAIGPT-6.1
OpenAI postpones GPT-6.1 Astra launch
new_model_or_versionQuelle
2026-09-29OpenAIGPT-6.1
OpenAI halts release of GPT-6.1 Astra over safety concerns
new_model_or_versionQuelle
2026-09-29OpenAIGPT-6.1
OpenAI halts GPT-6.1 Astra due to safety concerns
new_model_or_versionQuelle
2026-09-29OpenAIGPT 6.1
OpenAI halts release of new AI model
new_model_or_versionQuelle
2026-09-29OpenAIGPT 6.1
OpenAI halts new AI model after security concerns
new_model_or_versionQuelle
2026-09-28AnthropicClaude Sonnet 5.5
Anthropic's Claude Sonnet 5.5 nearly matches Opus 5.5 on benchmarks
new_model_or_versionQuelle
2026-09-24OpenAIGPT-6
OpenAI Launches GPT-6 Sol and Luna at Half the Price
new_model_or_versionQuelle
2026-09-22OpenAIGPT-6
OpenAI and Anthropic Launch Cheaper AI Models Amid Rising Competition
new_model_or_versionQuelle
2026-09-22AnthropicClaude Opus 5.5
Claude Opus 5.5 now available on Vercel AI Gateway
availability_or_api_updateQuelle
2026-09-22xAI / GrokGrok 4.7 Grok 4.7
Grok 4.7: xAI unveils most powerful model for coding and knowledge work
new_model_or_versionQuelle
2026-09-22OpenAIGPT-6
OpenAI cuts GPT-6 Sol and Luna prices by 50%
new_model_or_versionQuelle
2026-09-22OpenAIGPT-6
OpenAI cuts API costs for GPT-6 Sol and Luna by 50%
new_model_or_versionQuelle
2026-09-22OpenAIGPT-6
OpenAI launches GPT-6 Sol and Luna with 50% cheaper API
new_model_or_versionQuelle

External benchmark sources

These sources are used as orientation. Public performance values are shown only with measurement status and comparability.

SourceTypeFocusAssessment
LocalScore / OpenBenchmarkingexternal_reproducibleGeneration speed, TTFT, Prompt speed, hardware comparisonMethodically useful external benchmark source. Values must be matched by model, quantization, runtime and hardware profile.
QuelLLM.fr Benchmarksexternal_public_resultRTX 5090, RTX 4090, Mac, CPU, llama.cpp, Q4Useful for hardware orientation. CheckCom displays it only as externally measured orientation.
Öffentliche RTX-5090-LLM-Benchmarks / GitHubcommunity_benchmarkRTX 5090, VRAM, Power, Tokens/s, LM Studio / llama.cppHighly relevant for RTX-5090-class systems, but documentation quality varies by repository.
Ollama lokale APIlocal_metadatainstalled models, size, digest, family, parameter size, quantization levelVery useful for local model inventory. Speed requires a separate local benchmark run.
Raspberry Pi AI HAT+ 2 Dokumentationofficial_hardware_docsRaspberry Pi 5, AI HAT+ 2, Hailo-10H, LLM/VLM, Edge AIOfficial hardware source for suitability and limits, not a complete model benchmark table.