What this overview does
Local instead of API?
CheckCom helps assess when local workstation, edge or API routes are worth considering.
Quantization in context
Q4, Q5, Q6, Q8, GGUF and related signals are interpreted as hardware and fit indicators.
Measurement status shown
Public values are labelled as metadata, external orientation or future CheckCom-owned measurement.
Hardware profiles
| Hardware | Best for | Limits | Measurement status |
|---|---|---|---|
| RTX 5090 / 32 GB VRAM | 7B–35B quantisierte LLMs, Coding, lokale RAG-Tests, schnelle lokale Inferenz | sehr große 70B+ Modelle nur mit starker Quantisierung/Offload; mehrere parallele Nutzer prüfen | benchmark_prepared |
| RTX 4090 / 24 GB VRAM | kleinere bis mittlere quantisierte Modelle und lokale Experimente | 32B+ und lange Kontexte häufig knapp | external_and_future_checkcom |
| RTX Pro / 96 GB VRAM | große lokale Modelle, längere Kontexte, professionelle lokale Inferenz | hohe Anschaffungskosten; TCO prüfen | profile_prepared |
| Raspberry Pi 5 + AI HAT+ 2 / Hailo-10H | Edge-KI, kleine LLM-/VLM-Kandidaten, Vision, lokale Demo- und Datenschutzszenarien | kein Ersatz für große GPU-LLMs; Hailo-kompatible Modelle und Runtime nötig | demo_profile_available |
| Raspberry Pi 5 CPU-only | kleine CPU-Tests, Klassifikation, Offline-Demos | große LLMs und lange Kontexte nicht sinnvoll | inventory_prepared |
| API / Cloud / Enterprise | Skalierung, Enterprise-Verträge, hohe Parallelität und schnelle Produktivstarts | laufende Token-/Seat-Kosten, Datenroute und Vertrag prüfen | price_radar_linked |
Local and quantized model candidates
Derived from the CheckCom model catalog. measurement_status metadata_only means: no CheckCom benchmark yet.
| Model | Params B | Quantization | VRAM GB | RTX 5090 | Pi 5 / AI HAT+ 2 | Score |
|---|---|---|---|---|---|---|
| Qwen2.5-Coder 7B Instruct Alibaba / Qwen |
7.0 | Q4 Q4 |
8.0 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
95/100 metadata_only |
| Qwen3.8 Next 40B Exp MoE Healed GGUF Alibaba / Qwen |
40.0 | GGUF Q4_K_M, GGUF, gguf |
31.3 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen35B A3B SignOfFour Coder GGUF Alibaba / Qwen |
35.0 | GGUF Q2_K, GGUF, gguf |
18.7 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 35B A3B Vocabulary Trimming GGUF Alibaba / Qwen |
35.0 | GGUF Q4_K_S, GGUF, gguf |
28.2 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 35B A3B Uncensored xCloud GGUF Alibaba / Qwen |
35.0 | GGUF Q4_K_M, GGUF, gguf |
28.2 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 35B A3B TQ2Q GGUF Alibaba / Qwen |
35.0 | GGUF GGUF, gguf |
28.2 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 35B A3B StyleTune GGUF Alibaba / Qwen |
35.0 | GGUF IQ4_XS, GGUF, gguf |
28.2 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 35B A3B Q5 K M GGUF Alibaba / Qwen |
35.0 | GGUF q5_k_m, Q5, GGUF |
32.7 | offload nur mit Offload oder reduzierten Einstellungen realistisch |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 35B A3B Ornith 80 20 Beta GGUF Alibaba / Qwen |
35.0 | GGUF Q4_K_S, GGUF, gguf |
28.2 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 35B A3B NVFP4 Gguf Alibaba / Qwen |
35.0 | GGUF Gguf, gguf |
28.2 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 35B A3B NVFP4 GGUF Alibaba / Qwen |
35.0 | GGUF q3, GGUF, gguf |
22.3 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 35B A3B MTP GGUF Tater NoThink Alibaba / Qwen |
35.0 | GGUF Q4_K_M, GGUF, gguf |
28.2 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 35B A3B DS4 GGUF Alibaba / Qwen |
35.0 | GGUF Q4_K_S, GGUF, gguf |
28.2 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 35B A3B Claude 4.7 Distill NVFP4 GGUF Alibaba / Qwen |
35.0 | GGUF GGUF, gguf |
28.2 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 35B A3B Claude 4.7 Distill MXFP4 MoE GGUF Alibaba / Qwen |
35.0 | GGUF GGUF, gguf |
28.2 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 35B A3B Abliterated GGUF Alibaba / Qwen |
35.0 | GGUF IQ2_M, GGUF, gguf |
18.7 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.5 35B A3B BitClass3 GGUF Alibaba / Qwen |
35.0 | GGUF Q3_K_S, GGUF, gguf |
22.3 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.5 35B A3B APEX GGUF Alibaba / Qwen |
35.0 | GGUF GGUF, gguf |
28.2 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Huihui Qwen AgentWorld 35B A3B Abliterated UD Q3 K M GGUF Alibaba / Qwen |
35.0 | GGUF Q3_K_M, Q3, GGUF |
22.3 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| GraphForge Qwen3.6 35B A3B SFT I1 GGUF Alibaba / Qwen |
35.0 | GGUF IQ3_M, GGUF, gguf |
22.3 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| GraphForge Qwen3.6 35B A3B SFT GGUF Alibaba / Qwen |
35.0 | GGUF IQ4_XS, GGUF, gguf |
28.2 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Endy Qwen3.6 CyberSec 35B A3B GGUF Alibaba / Qwen |
35.0 | GGUF Q2_K, GGUF, gguf |
18.7 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3 32B Qwople GGUF Alibaba / Qwen |
32.0 | GGUF Q4_K_M, GGUF, gguf |
26.3 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3 32B GGUF Alibaba / Qwen |
32.0 | GGUF IQ3_M, GGUF, gguf |
20.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen2.5 Coder 32B Python Specialist I1 GGUF Alibaba / Qwen |
32.0 | GGUF IQ3_M, GGUF, gguf |
20.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen2.5 Coder 32B Python Specialist GGUF Alibaba / Qwen |
32.0 | GGUF IQ4_XS, GGUF, gguf |
26.3 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen2.5 Coder 32B Instruct Jbliterated Alibaba / Qwen |
32.0 | GGUF Q4_K_M, gguf |
26.3 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen2.5 Coder 32B Instruct GGUF Alibaba / Qwen |
32.0 | GGUF Q2_K, GGUF, gguf |
17.7 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen2.5 32B Instruct AdiTurbo GGUF Alibaba / Qwen |
32.0 | GGUF Q3_K_M, GGUF, gguf |
20.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| DeepSeek R1 Distill Qwen32B Q4 K M GGUF Alibaba / Qwen |
32.0 | GGUF q4_k_m, Q4, GGUF |
26.3 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3 Coder 30B A3B Instruct GGUF Alibaba / Qwen |
30.0 | GGUF Q4_K_M, GGUF, gguf |
25.1 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3 30B A3B Vietnamese Instruct GGUF Alibaba / Qwen |
30.0 | GGUF Q4_K_M, GGUF, gguf |
25.1 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3 30B A3B GGUF Alibaba / Qwen |
30.0 | GGUF IQ3_M, GGUF, gguf |
20.0 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Ektome Qwen3 30B A3B PristinelyUncensored GGUF Alibaba / Qwen |
30.0 | GGUF Q2_K, GGUF, gguf |
17.0 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| ThinkingCap Qwen3.8 27B I1 GGUF Alibaba / Qwen |
27.0 | GGUF IQ3_M, GGUF, gguf |
18.6 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Swift Qwen3.8 27B Uncensored MTP Terse Coder I1 GGUF Alibaba / Qwen |
27.0 | GGUF IQ3_M, GGUF, gguf |
18.6 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Swift 1.5 Qwen3.8 27B I1 GGUF Alibaba / Qwen |
27.0 | GGUF IQ3_M, GGUF, gguf |
18.6 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Swift 1.5 Qwen3.8 27B Heretic I1 GGUF Alibaba / Qwen |
27.0 | GGUF IQ3_M, GGUF, gguf |
18.6 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Swift 1.5 Qwen3.8 27B Heretic GSQ RCO GGUF Alibaba / Qwen |
27.0 | GGUF IQ2_S, GGUF, gguf |
15.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Swift 1.5 Qwen3.8 27B GGUF Alibaba / Qwen |
27.0 | GGUF IQ4_XS, GGUF, gguf |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.8 27B ZipBrain GGUF Alibaba / Qwen |
27.0 | GGUF IQ4_XS, GGUF, gguf |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.8 27B Uncensored Q4 K M GGUF Alibaba / Qwen |
27.0 | GGUF Q4_K_M, Q4, GGUF |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.8 27B Uncensored Q4 K M GGUF Alibaba / Qwen |
27.0 | GGUF Q4_K_M, Q4, GGUF |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.8 27B Uncensored Aggressive I1 GGUF Alibaba / Qwen |
27.0 | GGUF IQ3_M, GGUF, gguf |
18.6 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.8 27B Uncensored Aggressive GGUF Alibaba / Qwen |
27.0 | GGUF IQ4_XS, GGUF, gguf |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.8 27B Terse Coder GGUF Alibaba / Qwen |
27.0 | GGUF q4_k_m, GGUF, gguf |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.8 27B TURBO Fable Cold Fusion 735 882 Heretic Uncensored NM DAU NVFP4 GGUF Alibaba / Qwen |
27.0 | GGUF GGUF, gguf |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.8 27B ROCmFPX GGUF Alibaba / Qwen |
27.0 | GGUF GGUF, gguf |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.8 27B MindMeld I1 GGUF Alibaba / Qwen |
27.0 | GGUF IQ3_M, GGUF, gguf |
18.6 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.8 27B I1 GGUF Alibaba / Qwen |
27.0 | GGUF IQ3_M, GGUF, gguf |
18.6 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.8 27B Human KO Enterprise Boundary I1 GGUF Alibaba / Qwen |
27.0 | GGUF IQ3_M, GGUF, gguf |
18.6 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.8 27B Fable5 Distill Abliterated Imatrix GGUF Alibaba / Qwen |
27.0 | GGUF GGUF, gguf |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.8 27B Fable Distill GGUF Alibaba / Qwen |
27.0 | GGUF IQ4_XS, GGUF, gguf |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.8 27B Cold Fusion GAIN V1.1 Q4 K M GGUF Alibaba / Qwen |
27.0 | GGUF q4_k_m, Q4, GGUF |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.8 27B Abliterated SFT I1 GGUF Alibaba / Qwen |
27.0 | GGUF IQ3_M, GGUF, gguf |
18.6 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.8 27B Abliterated MTP GGUF Alibaba / Qwen |
27.0 | GGUF IQ2_M, GGUF, gguf |
15.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.8 27B 5090 Goldilocks GGUF Alibaba / Qwen |
27.0 | GGUF GGUF, gguf |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 27B Samantha Uncensored GGUF Alibaba / Qwen |
27.0 | GGUF Q4_K_M, GGUF, gguf |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 27B Q4 K M GGUF Alibaba / Qwen |
27.0 | GGUF q4_k_m, Q4, GGUF |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 27B NVFP4 GGUF Alibaba / Qwen |
27.0 | GGUF q3, GGUF, gguf |
18.6 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 27B Mtp Gguf Alibaba / Qwen |
27.0 | GGUF Q4_K_X, Gguf, Q4_K_XL |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 27B KR MTP GGUF Alibaba / Qwen |
27.0 | GGUF Q5_K_X, GGUF, Q5_K_XL |
26.7 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 27B KR GGUF Alibaba / Qwen |
27.0 | GGUF Q5_K_X, GGUF, Q5_K_XL |
26.7 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 27B GGUF Alibaba / Qwen |
27.0 | GGUF Q4_K_M, GGUF, gguf |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 27B Esper4 I1 GGUF Alibaba / Qwen |
27.0 | GGUF IQ3_M, GGUF, gguf |
18.6 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 27B Esper4 GGUF Alibaba / Qwen |
27.0 | GGUF IQ4_XS, GGUF, gguf |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 27B Architect Polaris2 Fable B F451 MTP ROCmFPX GGUF Alibaba / Qwen |
27.0 | GGUF Q6, GGUF, gguf |
30.8 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 27B AEON RYS Agentic Coder PatchCode GGUF Alibaba / Qwen |
27.0 | GGUF IQ4_NL, GGUF, gguf |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Ostrich 27B Qwen3.8 260815 I1 GGUF Alibaba / Qwen |
27.0 | GGUF IQ3_M, GGUF, gguf |
18.6 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Medgap Qwen3.8 27B GGUF Alibaba / Qwen |
27.0 | GGUF IQ4_XS, GGUF, gguf |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Loes Qwen3.8 27B I1 GGUF Alibaba / Qwen |
27.0 | GGUF IQ3_M, GGUF, gguf |
18.6 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Huihui Qwen3.8 27B Abliterated GGUF Alibaba / Qwen |
27.0 | GGUF IQ3_M, GGUF, gguf |
18.6 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| GraphForge Qwen3.6 27B SFT GGUF Alibaba / Qwen |
27.0 | GGUF IQ4_XS, GGUF, gguf |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Elster Vernunft Qwen3.6 27B GGUF Alibaba / Qwen |
27.0 | GGUF IQ4_XS, GGUF, gguf |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Dagger Qwen3.6 27B GGUF MTP Alibaba / Qwen |
27.0 | GGUF Q3_K_M, GGUF, gguf |
18.6 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| ATX Swift 1.5 Qwen3.8 27B Uncensored MTP GGUF Alibaba / Qwen |
27.0 | GGUF Q4_K_M, GGUF, gguf |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| ATX Swift 1.5 Qwen3.8 27B Uncensored IQ4 XS M GGUF Alibaba / Qwen |
27.0 | GGUF IQ4_XS, GGUF, gguf |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.8 20B Minitron Q4 K M GGUF Alibaba / Qwen |
20.0 | GGUF q4_k_m, Q4, GGUF |
18.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Lumynax Reasoning DeepSeek R1 Qwen15b Gguf Alibaba / Qwen |
15.0 | GGUF Q4_K_M, Gguf, gguf |
15.8 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 14B A3B VibeForged V2 GGUF Alibaba / Qwen |
14.0 | GGUF F16, GGUF, gguf |
20.4 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3 14B Uncensored GGUF Alibaba / Qwen |
14.0 | GGUF IQ4_XS, GGUF, gguf |
13.7 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3 14B PragReST I1 GGUF Alibaba / Qwen |
14.0 | GGUF IQ1_M, GGUF, gguf |
20.4 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen2.5 Coder 14B Instruct Uncensored Q4 K M GGUF Alibaba / Qwen |
14.0 | GGUF q4_k_m, Q4, GGUF |
13.7 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen2.5 Coder 14B Instruct Heretic GGUF Alibaba / Qwen |
14.0 | GGUF IQ4_XS, GGUF, gguf |
13.7 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen2.5 Coder 14B Instruct GGUF Alibaba / Qwen |
14.0 | GGUF f16, GGUF, gguf |
20.4 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen2.5 Coder 14B Instruct GGUF Alibaba / Qwen |
14.0 | GGUF fp16, GGUF, gguf |
34.4 | offload nur mit Offload oder reduzierten Einstellungen realistisch |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen2.5 14B Instruct KanColle Hibiki Verniy GGUF Alibaba / Qwen |
14.0 | GGUF F16, GGUF, gguf |
20.4 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen2.5 14B Instruct Heretic Alibaba / Qwen |
14.0 | GGUF F16, gguf, Q4_K_M |
13.7 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen Qwen2.5 Coder 14B Instruct GGUF Q8 0 Alibaba / Qwen |
14.0 | GGUF Q8, GGUF, Q8_0 |
20.4 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen Qwen2.5 Coder 14B Instruct GGUF Q6 K Alibaba / Qwen |
14.0 | GGUF Q6_K, GGUF, Q6 |
17.6 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen Qwen2.5 Coder 14B Instruct GGUF Q4 K M Alibaba / Qwen |
14.0 | GGUF Q4_K_M, GGUF, Q4 |
13.7 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen Qwen2.5 Coder 14B Instruct GGUF Q3 K M Alibaba / Qwen |
14.0 | GGUF Q3_K_M, GGUF, Q3 |
11.3 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Lfed Qwen2.5 Coder 14B Sql Gguf Alibaba / Qwen |
14.0 | GGUF Q4_K_M, Gguf, gguf |
13.7 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Iol Qwen2.5 14B Instruct AWQ Alibaba / Qwen |
14.0 | AWQ AWQ |
13.7 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| H100 Qwen3 14B MonEspaceSante CPT I1 GGUF Alibaba / Qwen |
14.0 | GGUF IQ2_M, GGUF, gguf |
9.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| H100 Qwen3 14B MonEspaceSante CPT GGUF Alibaba / Qwen |
14.0 | GGUF IQ4_XS, GGUF, gguf |
13.7 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Ektome Qwen2.5 Coder 14B Instruct PristinelyUncensored Alibaba / Qwen |
14.0 | GGUF IQ3_M, gguf |
11.3 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Adi Qwen2.5 14B GLM 5.2 General GGUF Alibaba / Qwen |
14.0 | GGUF q4_k_m, GGUF, gguf |
13.7 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| SkillGym Qwen3.5 9B GGUF Alibaba / Qwen |
9.0 | GGUF f16, GGUF, gguf |
14.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| QwenPaw Flash 9B Heretic MTP GGUF Alibaba / Qwen |
9.0 | GGUF BF16, GGUF, gguf |
23.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| QwenPaw Flash 9B Heretic Imatrix GGUF Alibaba / Qwen |
9.0 | GGUF BF16, GGUF, gguf |
23.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| QwenPaw Flash 9B Heretic GGUF Alibaba / Qwen |
9.0 | GGUF BF16, GGUF, gguf |
23.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.5 9B Ultra Uncensored Heretic V2 I1 GGUF Alibaba / Qwen |
9.0 | GGUF IQ1_M, GGUF, gguf |
14.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.5 9B Ultra Uncensored Heretic V1 I1 GGUF Alibaba / Qwen |
9.0 | GGUF IQ1_M, GGUF, gguf |
14.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.5 9B The Defiant Fable Uncensored Heretic NEO IMATRIX MAX MTP GGUF Alibaba / Qwen |
9.0 | GGUF BF16, GGUF, gguf |
23.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.5 9B The Defiant Fable Uncensored Heretic NEO IMATRIX MAX GGUF Alibaba / Qwen |
9.0 | GGUF BF16, GGUF, gguf |
23.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.5 9B TAP DPQ V6 GGUF Alibaba / Qwen |
9.0 | GGUF IQ1_S, GGUF, gguf |
14.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.5 9B Samantha Uncensored GGUF Alibaba / Qwen |
9.0 | GGUF Q4_K_M, GGUF, gguf |
10.6 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.5 9B Heretic Imatrix GGUF Alibaba / Qwen |
9.0 | GGUF BF16, GGUF, gguf |
23.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.5 9B Heretic GGUF Alibaba / Qwen |
9.0 | GGUF BF16, GGUF, gguf |
23.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.5 9B Haskell Rust Python Q6 K GGUF Alibaba / Qwen |
9.0 | GGUF q6_k, Q6, GGUF |
13.1 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.5 9B Haskell Rust Python IQ4 XS GGUF Alibaba / Qwen |
9.0 | GGUF iq4_xs, GGUF, gguf |
10.6 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.5 9B GGUF Alibaba / Qwen |
9.0 | GGUF BF16, GGUF, gguf |
23.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.5 9B GGUF Alibaba / Qwen |
9.0 | GGUF IQ1_KT, GGUF, gguf |
14.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.5 9B DeepSeek V4 Flash GGUF Alibaba / Qwen |
9.0 | GGUF Q3_K_M, GGUF, gguf |
9.0 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.5 9B DFlash GGUF Alibaba / Qwen |
9.0 | GGUF bf16, DFlash, GGUF |
23.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.5 9B Brainwaves GGUF Alibaba / Qwen |
9.0 | GGUF F16, GGUF, gguf |
14.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.5 9B Base GGUF Alibaba / Qwen |
9.0 | GGUF Q4_K_S, GGUF, gguf |
10.6 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| MiMo V2.6 Distill Qwen9B GGUF Alibaba / Qwen |
9.0 | GGUF IQ1_M, GGUF, gguf |
14.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| MiMo V2.6 Distill Qwen9B Ablitrated I1 GGUF Alibaba / Qwen |
9.0 | GGUF IQ1_M, GGUF, gguf |
14.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
Local Ollama inventory
Status: ok · models: 22
If Ollama is running, installed models are inventoried with size, digest, family and quantization. Speed requires benchmark runs.
API, workstation or edge?
CheckCom uses this data to surface local alternatives to API/cloud routes in the AI advisor. For enterprises, model size is only one factor; operations, privacy, latency, maintenance and utilization matter as well.
Local / quantized news and release signals
| Date | Provider | Signal | Type | Source |
|---|---|---|---|---|
| 2026-10-05 | Zhipu / GLM | GLM-5.2 Reflection unveils Beam, a new open-weight AI model | new_model_or_version | Quelle |
| 2026-10-05 | DeepSeek | DeepSeek y Qwen Reflection AI presentó este lunes Beam Reflection AI launches Beam, the open model challenging China | new_model_or_version | Quelle |
| 2026-10-05 | Anthropic | Claude Code devpit | local_or_quantized_update | Quelle |
| 2026-10-03 | DeepSeek | DeepSeek V4 vs Qwen3.8 vs Kimi K3 vs GLM-5.3 I skipped the vendor tables and lined up the independen DeepSeek V4 vs Qwen3.8 vs Kimi K3 vs GLM-5.3: Open-Weight LLMs Compared | local_or_quantized_update | Quelle |
| 2026-10-03 | Anthropic | Claude to Open-source 'BootLoops' tool supports AI in precise scientific calculations | local_or_quantized_update | Quelle |
| 2026-10-02 | DeepSeek | DeepSeek V4 vs Qwen3.8 vs Kimi K3 vs GLM-5.3 https Best Open-Weight LLMs Tested: DeepSeek V4 vs Qwen3.8 vs Kimi K3 vs GLM-5.3 | local_or_quantized_update | Quelle |
| 2026-10-02 | OpenAI | GPT-6 OpenAI's GPT-6 Astra Ultrafast Achieves 8x Speed Boost on NVIDIA | new_model_or_version | Quelle |
| 2026-10-01 | Google / Gemini | Gemini AI Judge dismisses antitrust lawsuits against Google AI search | local_or_quantized_update | Quelle |
| 2026-10-01 | OpenAI | GPT-6.1 OpenAI unveils GPT-6.1 Sol and Dots | new_model_or_version | Quelle |
| 2026-09-30 | Google / Gemini | Gemini 4 Google unveils Gemini 4 Argon with 1M tokens | new_model_or_version | Quelle |
| 2026-09-30 | Google / Gemini | Gemini 4 Google unveils Gemini 4 Argon AI model – not available yet | new_model_or_version | Quelle |
| 2026-09-30 | OpenAI | GPT-6.1 OpenAI unveils Dots and GPT-6.1 Sol at DevDay 2026 | new_model_or_version | Quelle |
| 2026-09-30 | OpenAI | GPT-6 OpenAI launches Dots: AI agents competing with Meta's Muse | new_model_or_version | Quelle |
| 2026-09-30 | OpenAI | GPT-6.1 OpenAI unveils GPT-6.1 Sol and Dots AI agents | new_model_or_version | Quelle |
| 2026-09-30 | Anthropic | GLM-5.3 Anthropic: Zhipu's GLM-5.3 nearly matches Claude Mythos in exploit building | preview_or_experimental | Quelle |
| 2026-09-30 | OpenAI | GPT-6.1 OpenAI cancels GPT-6.1 Astra over security issues | new_model_or_version | Quelle |
| 2026-09-30 | OpenAI | Gemini Spark OpenAI launches personal AI assistant 'dots' | new_model_or_version | Quelle |
| 2026-09-29 | OpenAI | GPT-6.1 OpenAI launches GPT-6.1 Sol at one-fifth the price | new_model_or_version | Quelle |
| 2026-09-29 | OpenAI | GPT-6.1 OpenAI launches GPT-6.1 Sol with improved performance and lower costs | new_model_or_version | Quelle |
| 2026-09-29 | OpenAI | GPT-6.1 OpenAI cancels release of GPT-6.1 Astra model | new_model_or_version | Quelle |
| 2026-09-29 | OpenAI | GPT-6.1 GPT-6.1 Sol Closes in on Astra at a Fifth of the Cost | availability_or_api_update | Quelle |
| 2026-09-29 | OpenAI | GPT-6 OpenAI halves API credits in Pro plan to push pay-per-use model | availability_or_api_update | Quelle |
| 2026-09-29 | OpenAI | GPT-6 OpenAI releases GPT-6.1 with significantly reduced pricing | new_model_or_version | Quelle |
| 2026-09-29 | OpenAI | GPT 6.1 OpenAI cancels release of new AI model | new_model_or_version | Quelle |
| 2026-09-29 | OpenAI | GPT-6.1 OpenAI unveils GPT-6.1 Sol and Dots AI agents | new_model_or_version | Quelle |
| 2026-09-29 | OpenAI | GPT-6.1 OpenAI halts GPT-6.1 Astra model over safety concerns | new_model_or_version | Quelle |
| 2026-09-29 | Mistral AI | Mistral eröffnet Hub in München Mistral opens AI hub in Munich | local_or_quantized_update | Quelle |
| 2026-09-29 | OpenAI | GPT-6.1 OpenAI postpones GPT-6.1 Astra launch | new_model_or_version | Quelle |
| 2026-09-29 | OpenAI | GPT-6.1 OpenAI halts release of GPT-6.1 Astra over safety concerns | new_model_or_version | Quelle |
| 2026-09-29 | OpenAI | GPT-6.1 OpenAI halts GPT-6.1 Astra due to safety concerns | new_model_or_version | Quelle |
| 2026-09-29 | OpenAI | GPT 6.1 OpenAI halts release of new AI model | new_model_or_version | Quelle |
| 2026-09-29 | OpenAI | GPT 6.1 OpenAI halts new AI model after security concerns | new_model_or_version | Quelle |
| 2026-09-28 | Anthropic | Claude Sonnet 5.5 Anthropic's Claude Sonnet 5.5 nearly matches Opus 5.5 on benchmarks | new_model_or_version | Quelle |
| 2026-09-24 | OpenAI | GPT-6 OpenAI Launches GPT-6 Sol and Luna at Half the Price | new_model_or_version | Quelle |
| 2026-09-22 | OpenAI | GPT-6 OpenAI and Anthropic Launch Cheaper AI Models Amid Rising Competition | new_model_or_version | Quelle |
| 2026-09-22 | Anthropic | Claude Opus 5.5 Claude Opus 5.5 now available on Vercel AI Gateway | availability_or_api_update | Quelle |
| 2026-09-22 | xAI / Grok | Grok 4.7 Grok 4.7 Grok 4.7: xAI unveils most powerful model for coding and knowledge work | new_model_or_version | Quelle |
| 2026-09-22 | OpenAI | GPT-6 OpenAI cuts GPT-6 Sol and Luna prices by 50% | new_model_or_version | Quelle |
| 2026-09-22 | OpenAI | GPT-6 OpenAI cuts API costs for GPT-6 Sol and Luna by 50% | new_model_or_version | Quelle |
| 2026-09-22 | OpenAI | GPT-6 OpenAI launches GPT-6 Sol and Luna with 50% cheaper API | new_model_or_version | Quelle |
External benchmark sources
These sources are used as orientation. Public performance values are shown only with measurement status and comparability.
| Source | Type | Focus | Assessment |
|---|---|---|---|
| LocalScore / OpenBenchmarking | external_reproducible | Generation speed, TTFT, Prompt speed, hardware comparison | Methodically useful external benchmark source. Values must be matched by model, quantization, runtime and hardware profile. |
| QuelLLM.fr Benchmarks | external_public_result | RTX 5090, RTX 4090, Mac, CPU, llama.cpp, Q4 | Useful for hardware orientation. CheckCom displays it only as externally measured orientation. |
| Öffentliche RTX-5090-LLM-Benchmarks / GitHub | community_benchmark | RTX 5090, VRAM, Power, Tokens/s, LM Studio / llama.cpp | Highly relevant for RTX-5090-class systems, but documentation quality varies by repository. |
| Ollama lokale API | local_metadata | installed models, size, digest, family, parameter size, quantization level | Very useful for local model inventory. Speed requires a separate local benchmark run. |
| Raspberry Pi AI HAT+ 2 Dokumentation | official_hardware_docs | Raspberry Pi 5, AI HAT+ 2, Hailo-10H, LLM/VLM, Edge AI | Official hardware source for suitability and limits, not a complete model benchmark table. |