What this overview does
Local instead of API?
CheckCom helps assess when local workstation, edge or API routes are worth considering.
Quantization in context
Q4, Q5, Q6, Q8, GGUF and related signals are interpreted as hardware and fit indicators.
Measurement status shown
Public values are labelled as metadata, external orientation or future CheckCom-owned measurement.
Hardware profiles
| Hardware | Best for | Limits | Measurement status |
|---|---|---|---|
| RTX 5090 / 32 GB VRAM | 7B–35B quantisierte LLMs, Coding, lokale RAG-Tests, schnelle lokale Inferenz | sehr große 70B+ Modelle nur mit starker Quantisierung/Offload; mehrere parallele Nutzer prüfen | benchmark_prepared |
| RTX 4090 / 24 GB VRAM | kleinere bis mittlere quantisierte Modelle und lokale Experimente | 32B+ und lange Kontexte häufig knapp | external_and_future_checkcom |
| RTX Pro / 96 GB VRAM | große lokale Modelle, längere Kontexte, professionelle lokale Inferenz | hohe Anschaffungskosten; TCO prüfen | profile_prepared |
| Raspberry Pi 5 + AI HAT+ 2 / Hailo-10H | Edge-KI, kleine LLM-/VLM-Kandidaten, Vision, lokale Demo- und Datenschutzszenarien | kein Ersatz für große GPU-LLMs; Hailo-kompatible Modelle und Runtime nötig | demo_profile_available |
| Raspberry Pi 5 CPU-only | kleine CPU-Tests, Klassifikation, Offline-Demos | große LLMs und lange Kontexte nicht sinnvoll | inventory_prepared |
| API / Cloud / Enterprise | Skalierung, Enterprise-Verträge, hohe Parallelität und schnelle Produktivstarts | laufende Token-/Seat-Kosten, Datenroute und Vertrag prüfen | price_radar_linked |
Local and quantized model candidates
Derived from the CheckCom model catalog. measurement_status metadata_only means: no CheckCom benchmark yet.
| Model | Params B | Quantization | VRAM GB | RTX 5090 | Pi 5 / AI HAT+ 2 | Score |
|---|---|---|---|---|---|---|
| Qwen2.5-Coder 7B Instruct Alibaba / Qwen |
7.0 | Q4 Q4 |
8.0 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
95/100 metadata_only |
| Qwen3.6 35B A3B Vocabulary Trimming GGUF Alibaba / Qwen |
35.0 | GGUF Q4_K_S, GGUF, gguf |
28.2 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 35B A3B Uncensored xCloud GGUF Alibaba / Qwen |
35.0 | GGUF Q4_K_M, GGUF, gguf |
28.2 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 35B A3B TQ2Q GGUF Alibaba / Qwen |
35.0 | GGUF GGUF, gguf |
28.2 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 35B A3B StyleTune GGUF Alibaba / Qwen |
35.0 | GGUF IQ4_XS, GGUF, gguf |
28.2 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 35B A3B Q5 K M GGUF Alibaba / Qwen |
35.0 | GGUF q5_k_m, Q5, GGUF |
32.7 | offload nur mit Offload oder reduzierten Einstellungen realistisch |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 35B A3B Ornith 80 20 Beta GGUF Alibaba / Qwen |
35.0 | GGUF Q4_K_S, GGUF, gguf |
28.2 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 35B A3B NVFP4 Gguf Alibaba / Qwen |
35.0 | GGUF Gguf, gguf |
28.2 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 35B A3B NVFP4 GGUF Alibaba / Qwen |
35.0 | GGUF q3, GGUF, gguf |
22.3 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 35B A3B MTP GGUF Tater NoThink Alibaba / Qwen |
35.0 | GGUF Q4_K_M, GGUF, gguf |
28.2 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 35B A3B DS4 GGUF Alibaba / Qwen |
35.0 | GGUF Q4_K_S, GGUF, gguf |
28.2 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 35B A3B Claude 4.7 Distill NVFP4 GGUF Alibaba / Qwen |
35.0 | GGUF GGUF, gguf |
28.2 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 35B A3B Claude 4.7 Distill MXFP4 MoE GGUF Alibaba / Qwen |
35.0 | GGUF GGUF, gguf |
28.2 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.5 35B A3B BitClass3 GGUF Alibaba / Qwen |
35.0 | GGUF Q3_K_S, GGUF, gguf |
22.3 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.5 35B A3B APEX GGUF Alibaba / Qwen |
35.0 | GGUF GGUF, gguf |
28.2 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Huihui Qwen AgentWorld 35B A3B Abliterated UD Q3 K M GGUF Alibaba / Qwen |
35.0 | GGUF Q3_K_M, Q3, GGUF |
22.3 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Endy Qwen3.6 CyberSec 35B A3B GGUF Alibaba / Qwen |
35.0 | GGUF Q2_K, GGUF, gguf |
18.7 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3 32B Qwople GGUF Alibaba / Qwen |
32.0 | GGUF Q4_K_M, GGUF, gguf |
26.3 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3 32B GGUF Alibaba / Qwen |
32.0 | GGUF IQ3_M, GGUF, gguf |
20.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen2.5 Coder 32B Python Specialist I1 GGUF Alibaba / Qwen |
32.0 | GGUF IQ3_M, GGUF, gguf |
20.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen2.5 Coder 32B Python Specialist GGUF Alibaba / Qwen |
32.0 | GGUF IQ4_XS, GGUF, gguf |
26.3 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen2.5 Coder 32B Instruct GGUF Alibaba / Qwen |
32.0 | GGUF Q2_K, GGUF, gguf |
17.7 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen2.5 32B Instruct AdiTurbo GGUF Alibaba / Qwen |
32.0 | GGUF Q3_K_M, GGUF, gguf |
20.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| DeepSeek R1 Distill Qwen32B Q4 K M GGUF Alibaba / Qwen |
32.0 | GGUF q4_k_m, Q4, GGUF |
26.3 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3 Coder 30B A3B Instruct GGUF Alibaba / Qwen |
30.0 | GGUF Q4_K_M, GGUF, gguf |
25.1 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3 30B A3B GGUF Alibaba / Qwen |
30.0 | GGUF IQ3_M, GGUF, gguf |
20.0 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Ektome Qwen3 30B A3B PristinelyUncensored GGUF Alibaba / Qwen |
30.0 | GGUF Q2_K, GGUF, gguf |
17.0 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.8 27B ZipBrain GGUF Alibaba / Qwen |
27.0 | GGUF IQ4_XS, GGUF, gguf |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.8 27B Uncensored Q4 K M GGUF Alibaba / Qwen |
27.0 | GGUF Q4_K_M, Q4, GGUF |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.8 27B Uncensored Aggressive I1 GGUF Alibaba / Qwen |
27.0 | GGUF IQ3_M, GGUF, gguf |
18.6 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.8 27B Uncensored Aggressive GGUF Alibaba / Qwen |
27.0 | GGUF IQ4_XS, GGUF, gguf |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.8 27B ROCmFPX GGUF Alibaba / Qwen |
27.0 | GGUF GGUF, gguf |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.8 27B I1 GGUF Alibaba / Qwen |
27.0 | GGUF IQ3_M, GGUF, gguf |
18.6 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.8 27B Fable5 Distill Abliterated Imatrix GGUF Alibaba / Qwen |
27.0 | GGUF GGUF, gguf |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.8 27B Fable Distill GGUF Alibaba / Qwen |
27.0 | GGUF IQ4_XS, GGUF, gguf |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.8 27B Cold Fusion GAIN V1.1 Q4 K M GGUF Alibaba / Qwen |
27.0 | GGUF q4_k_m, Q4, GGUF |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.8 27B Abliterated MTP GGUF Alibaba / Qwen |
27.0 | GGUF IQ2_M, GGUF, gguf |
15.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.8 27B 5090 Goldilocks GGUF Alibaba / Qwen |
27.0 | GGUF GGUF, gguf |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 27B Samantha Uncensored GGUF Alibaba / Qwen |
27.0 | GGUF Q4_K_M, GGUF, gguf |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 27B Q4 K M GGUF Alibaba / Qwen |
27.0 | GGUF q4_k_m, Q4, GGUF |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 27B NVFP4 GGUF Alibaba / Qwen |
27.0 | GGUF q3, GGUF, gguf |
18.6 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 27B Mtp Gguf Alibaba / Qwen |
27.0 | GGUF Q4_K_X, Gguf, Q4_K_XL |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 27B KR MTP GGUF Alibaba / Qwen |
27.0 | GGUF Q5_K_X, GGUF, Q5_K_XL |
26.7 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 27B KR GGUF Alibaba / Qwen |
27.0 | GGUF Q5_K_X, GGUF, Q5_K_XL |
26.7 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 27B GGUF Alibaba / Qwen |
27.0 | GGUF Q4_K_M, GGUF, gguf |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 27B Esper4 I1 GGUF Alibaba / Qwen |
27.0 | GGUF IQ3_M, GGUF, gguf |
18.6 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 27B Esper4 GGUF Alibaba / Qwen |
27.0 | GGUF IQ4_XS, GGUF, gguf |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 27B Architect Polaris2 Fable B F451 MTP ROCmFPX GGUF Alibaba / Qwen |
27.0 | GGUF Q6, GGUF, gguf |
30.8 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 27B AEON RYS Agentic Coder PatchCode GGUF Alibaba / Qwen |
27.0 | GGUF IQ4_NL, GGUF, gguf |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Ostrich 27B Qwen3.8 260815 I1 GGUF Alibaba / Qwen |
27.0 | GGUF IQ3_M, GGUF, gguf |
18.6 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Huihui Qwen3.8 27B Abliterated GGUF Alibaba / Qwen |
27.0 | GGUF IQ3_M, GGUF, gguf |
18.6 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Elster Vernunft Qwen3.6 27B GGUF Alibaba / Qwen |
27.0 | GGUF IQ4_XS, GGUF, gguf |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Dagger Qwen3.6 27B GGUF MTP Alibaba / Qwen |
27.0 | GGUF Q3_K_M, GGUF, gguf |
18.6 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.8 20B Minitron Q4 K M GGUF Alibaba / Qwen |
20.0 | GGUF q4_k_m, Q4, GGUF |
18.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 14B A3B VibeForged V2 GGUF Alibaba / Qwen |
14.0 | GGUF F16, GGUF, gguf |
20.4 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3 14B PragReST I1 GGUF Alibaba / Qwen |
14.0 | GGUF IQ1_M, GGUF, gguf |
20.4 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen2.5 Coder 14B Instruct Jbliterated Alibaba / Qwen |
14.0 | GGUF Q4_K_M, gguf |
13.7 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen2.5 Coder 14B Instruct Heretic GGUF Alibaba / Qwen |
14.0 | GGUF IQ4_XS, GGUF, gguf |
13.7 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen2.5 Coder 14B Instruct GGUF Alibaba / Qwen |
14.0 | GGUF f16, GGUF, gguf |
20.4 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen2.5 Coder 14B Instruct GGUF Alibaba / Qwen |
14.0 | GGUF fp16, GGUF, gguf |
34.4 | offload nur mit Offload oder reduzierten Einstellungen realistisch |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen2.5 14B Instruct Heretic Alibaba / Qwen |
14.0 | GGUF Q4_K_M, gguf |
13.7 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Lfed Qwen2.5 Coder 14B Sql Gguf Alibaba / Qwen |
14.0 | GGUF Q4_K_M, Gguf, gguf |
13.7 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Iol Qwen2.5 14B Instruct AWQ Alibaba / Qwen |
14.0 | AWQ AWQ |
13.7 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| H100 Qwen3 14B MonEspaceSante CPT I1 GGUF Alibaba / Qwen |
14.0 | GGUF IQ2_M, GGUF, gguf |
9.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| H100 Qwen3 14B MonEspaceSante CPT GGUF Alibaba / Qwen |
14.0 | GGUF IQ4_XS, GGUF, gguf |
13.7 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Ektome Qwen2.5 Coder 14B Instruct PristinelyUncensored Alibaba / Qwen |
14.0 | GGUF IQ3_M, gguf |
11.3 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Adi Qwen2.5 14B GLM 5.2 General GGUF Alibaba / Qwen |
14.0 | GGUF q4_k_m, GGUF, gguf |
13.7 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| QwenPaw Flash 9B Heretic MTP GGUF Alibaba / Qwen |
9.0 | GGUF BF16, GGUF, gguf |
23.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| QwenPaw Flash 9B Heretic Imatrix GGUF Alibaba / Qwen |
9.0 | GGUF BF16, GGUF, gguf |
23.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| QwenPaw Flash 9B Heretic GGUF Alibaba / Qwen |
9.0 | GGUF BF16, GGUF, gguf |
23.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.5 9B Ultra Uncensored Heretic V2 I1 GGUF Alibaba / Qwen |
9.0 | GGUF IQ1_M, GGUF, gguf |
14.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.5 9B Ultra Uncensored Heretic V1 I1 GGUF Alibaba / Qwen |
9.0 | GGUF IQ1_M, GGUF, gguf |
14.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.5 9B The Defiant Fable Uncensored Heretic NEO IMATRIX MAX MTP GGUF Alibaba / Qwen |
9.0 | GGUF BF16, GGUF, gguf |
23.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.5 9B The Defiant Fable Uncensored Heretic NEO IMATRIX MAX GGUF Alibaba / Qwen |
9.0 | GGUF BF16, GGUF, gguf |
23.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.5 9B Samantha Uncensored GGUF Alibaba / Qwen |
9.0 | GGUF Q4_K_M, GGUF, gguf |
10.6 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.5 9B Heretic Imatrix GGUF Alibaba / Qwen |
9.0 | GGUF BF16, GGUF, gguf |
23.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.5 9B Heretic GGUF Alibaba / Qwen |
9.0 | GGUF BF16, GGUF, gguf |
23.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.5 9B Haskell Rust Python Q6 K GGUF Alibaba / Qwen |
9.0 | GGUF q6_k, Q6, GGUF |
13.1 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.5 9B Haskell Rust Python IQ4 XS GGUF Alibaba / Qwen |
9.0 | GGUF iq4_xs, GGUF, gguf |
10.6 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.5 9B GGUF Alibaba / Qwen |
9.0 | GGUF IQ1_KT, GGUF, gguf |
14.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.5 9B DeepSeek V4 Flash GGUF Alibaba / Qwen |
9.0 | GGUF Q3_K_M, GGUF, gguf |
9.0 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.5 9B DFlash GGUF Alibaba / Qwen |
9.0 | GGUF bf16, DFlash, GGUF |
23.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Huihui Qwen3.5 9B Claude 4.6 Opus Abliterated Heretic GGUF Alibaba / Qwen |
9.0 | GGUF f16, GGUF, gguf |
14.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| FACET Terminal Qwen3.5 9B GGUF Alibaba / Qwen |
9.0 | GGUF f16, GGUF, gguf |
14.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Adi Qwen3.5 9B GLM 5.2 General GGUF Alibaba / Qwen |
9.0 | GGUF q4_k_m, GGUF, gguf |
10.6 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| VideoKR Qwen3 VL 8B GGUF Alibaba / Qwen |
8.0 | GGUF f16, GGUF, gguf |
13.8 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| SenseNova U1.5 8B GGUFs Hugging Face / sonstige Owner |
8.0 | GGUF Q2_K, gguf |
7.8 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| RavenX Conjecture Qwen3 8B GGUF Alibaba / Qwen |
8.0 | GGUF Q8, GGUF, Q8_0 |
13.8 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3 VL 8B Instruct Heretic GGUF Alibaba / Qwen |
8.0 | GGUF f16, GGUF, gguf |
13.8 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3 VL 8B Instruct Abliterated V2 GGUF Alibaba / Qwen |
8.0 | GGUF f16, GGUF, gguf |
13.8 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3 Embed 8B GGUF Alibaba / Qwen |
8.0 | GGUF iq4_xs, GGUF, gguf |
10.0 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3 8B UnBias Plus SFT Instruct Legacy GGUF Alibaba / Qwen |
8.0 | GGUF f16, GGUF, gguf |
13.8 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3 8B PragReST I1 GGUF Alibaba / Qwen |
8.0 | GGUF IQ1_M, GGUF, gguf |
13.8 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3 8B PragReST GGUF Alibaba / Qwen |
8.0 | GGUF f16, GGUF, gguf |
13.8 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3 8B MedReasonPath GGUF Alibaba / Qwen |
8.0 | GGUF f16, GGUF, gguf |
13.8 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3 8B Heretic GGUF Alibaba / Qwen |
8.0 | GGUF f16, GGUF, gguf |
13.8 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3 8B Gguf Alibaba / Qwen |
8.0 | GGUF Q8, Gguf, Q8_0 |
13.8 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3 8B Apostate I1 GGUF Alibaba / Qwen |
8.0 | GGUF IQ1_S, GGUF, gguf |
13.8 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3 8B Apostate GGUF Alibaba / Qwen |
8.0 | GGUF f16, GGUF, gguf |
13.8 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| PDP Qwen3 8B SFT GGUF Alibaba / Qwen |
8.0 | GGUF f16, GGUF, gguf |
13.8 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| LFM2.5 8B A1B GGUF LFM2.5 |
8.0 | GGUF BF16, GGUF, gguf |
21.8 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| L40S Qwen3 8B MonEspaceSante CPT SFT Anti Hallucination I1 GGUF Alibaba / Qwen |
8.0 | GGUF IQ2_S, GGUF, gguf |
7.8 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Huihui Qwen3 VL 8B Instruct Abliterated FP8 Alibaba / Qwen |
8.0 | FP8 FP8 |
10.0 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Huihui Qwen3 VL 8B Instruct Abliterated AWQ Int4 Alibaba / Qwen |
8.0 | AWQ AWQ, Int4, int4 |
10.0 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Adi Qwen3 8B GLM 5.2 General GGUF Alibaba / Qwen |
8.0 | GGUF q4_k_m, GGUF, gguf |
10.0 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Yoiko Img Prompt Gen Qwen2.5 7B Instruct Q4 K M Alibaba / Qwen |
7.0 | GGUF Q4_K_M, Q4, gguf |
9.3 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| VideoKR Qwen2.5 VL 7B GGUF Alibaba / Qwen |
7.0 | GGUF f16, GGUF, gguf |
12.7 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Typhoon2 Qwen2.5 7B Instruct GGUF Alibaba / Qwen |
7.0 | GGUF Q4_K_M, GGUF, gguf |
9.3 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Stiles Seymens Qwen2.5 7B Generalist Q4KM GGUF Alibaba / Qwen |
7.0 | GGUF Q4_K_M, GGUF, gguf |
9.3 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Sofia Qwen2.5 7B I1 GGUF Alibaba / Qwen |
7.0 | GGUF IQ1_M, GGUF, gguf |
12.7 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Sofia Qwen2.5 7B GGUF Alibaba / Qwen |
7.0 | GGUF f16, GGUF, gguf |
12.7 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen2.5 VL 7B Instruct Int8 Convrot Comfyui Alibaba / Qwen |
7.0 | Int8 Int8, int8 |
9.3 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen2.5 Coder 7B VN Master Polymath FINAL GGUF Alibaba / Qwen |
7.0 | GGUF Q3_K_M, GGUF, gguf |
8.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen2.5 Coder 7B Instruct Jbliterated Alibaba / Qwen |
7.0 | GGUF Q4_K_M, gguf |
9.3 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen2.5 Coder 7B Instruct Heretic Alibaba / Qwen |
7.0 | GGUF Q4_K_M, gguf |
9.3 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen2.5 Coder 7B Instruct GGUF Alibaba / Qwen |
7.0 | GGUF fp16, GGUF, gguf |
19.7 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen2.5 Coder 7B Instruct GGUF Alibaba / Qwen |
7.0 | GGUF q3_k_m, GGUF, gguf |
8.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen2.5 7B Instruct INT8 Quanto Alibaba / Qwen |
7.0 | INT8 INT8 |
9.3 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen2.5 7B Instruct Heretic Alibaba / Qwen |
7.0 | GGUF Q4_K_M, gguf |
9.3 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen2.5 7B Instruct FP8 Dynamic TR171 Alibaba / Qwen |
7.0 | FP8 FP8 |
9.3 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
Local Ollama inventory
Status: unavailable · models: 0
If Ollama is running, installed models are inventoried with size, digest, family and quantization. Speed requires benchmark runs.
API, workstation or edge?
CheckCom uses this data to surface local alternatives to API/cloud routes in the AI advisor. For enterprises, model size is only one factor; operations, privacy, latency, maintenance and utilization matter as well.
Local / quantized news and release signals
| Date | Provider | Signal | Type | Source |
|---|---|---|---|---|
| 2026-08-20 | Moonshot / Kimi | Kimi K3 and GLM-5.3 are now Frontier Radar #4: China's AI models are catching up | local_or_quantized_update | Quelle |
| 2026-08-20 | OpenAI | GPT-5.6 AWS brings OpenAI's GPT-5.6 models to India | new_model_or_version | Quelle |
| 2026-08-19 | Google / Gemini | Gemma 4 Boost Local AI Performance with LLM Turbo | availability_or_api_update | Quelle |
| 2026-08-19 | OpenAI | GPT ChatGPT is reshaping student research habits | availability_or_api_update | Quelle |
| 2026-08-18 | Google / Gemini | Gemini AI Google's Gemini AI scans Workspace data by default | local_or_quantized_update | Quelle |
| 2026-08-18 | OpenAI | GPT OpenAI shifts to B2B focus | availability_or_api_update | Quelle |
| 2026-08-18 | OpenAI | Qwen3.8-27B Qwen3.8-27B: 3 million downloads in three days | local_or_quantized_update | Quelle |
| 2026-08-17 | xAI / Grok | Grok Bot SpaceXAI Launches Grok Bot for Autonomous AI Agents | new_model_or_version | Quelle |
| 2026-08-16 | OpenAI | GPT-5.6 OpenAI Launches Ultrafast: GPT-5.6 Sol at 750 Tokens/Second | new_model_or_version | Quelle |
| 2026-08-16 | Alibaba / Qwen | Qwen 3.8 27B Qwen 3.8 27B is excellent, but overthinks by default | new_model_or_version | Quelle |
| 2026-08-16 | Alibaba / Qwen | Qwen3.8-27B Alibaba Alibaba releases Qwen3.8-27B for local use | new_model_or_version | Quelle |
| 2026-08-16 | OpenAI | GPT OpenAI: 8.3x gap in enterprise AI agent usage | availability_or_api_update | Quelle |
| 2026-08-15 | Google / Gemini | Qwen AI models have surpassed 3 billion global downloads in six months Alibaba's AI models outpace Meta, Google to hit 3B downloads | local_or_quantized_update | Quelle |
| 2026-08-14 | Mistral AI | Mistral erweitert Angebot Die Organisation AI Update: Hate Aid criticizes AI glasses, Mistral expands offerings | local_or_quantized_update | Quelle |
| 2026-08-14 | OpenAI | Qwen team Alibaba releases Qwen 3.8 with open model weights | new_model_or_version | Quelle |
| 2026-08-14 | OpenAI | GPT-5.6 OpenAI launches Ultrafast mode for GPT-5.6 Sol | new_model_or_version | Quelle |
| 2026-08-14 | Zhipu / GLM | GLM-5.3 Zhipu AI releases GLM-5.3 as strongest open-weights coding model | new_model_or_version | Quelle |
| 2026-08-14 | OpenAI | GPT-5.6 OpenAI introduces Ultrafast mode for GPT-5.6 Sol | preview_or_experimental | Quelle |
| 2026-08-13 | OpenAI | GPT-5.6 OpenAI: GPT-5.6 Sol 14x Faster with Cerebras | new_model_or_version | Quelle |
| 2026-08-13 | OpenAI | Grok 4.6 iguala a GPT-5.6 Sol con 61 puntos y menor coste para startups Grok 4.6 iguala a GPT-5.6 So Grok 4.6 matches GPT-5.6 Sol at lower cost | new_model_or_version | Quelle |
| 2026-08-13 | Mistral AI | Mistral als KI Mistral expands AI infrastructure in Europe | local_or_quantized_update | Quelle |
| 2026-08-13 | Anthropic | Grok 4.6 Grok 4.6: Efficiency Edge for Startups | local_or_quantized_update | Quelle |
| 2026-08-13 | OpenAI | GPT-5.6 OpenAI unveils GPT-5.6 Sol with 14x faster inference | new_model_or_version | Quelle |
| 2026-08-13 | OpenAI | Gemini vor South Korea’s blind fortune-tellers eclipsed by AI | availability_or_api_update | Quelle |
| 2026-08-13 | OpenAI | Kimi K3 Kimi K3: Open-Source AI Model with 2.8 Trillion Parameters | availability_or_api_update | Quelle |
| 2026-08-12 | DeepSeek | DeepSeek V4 Pro 0813 DeepSeek V4 Pro 0813 available via OpenRouter | new_model_or_version | Quelle |
| 2026-08-12 | OpenAI | Grok 4.6 de SpaceXAI SpaceXAI's Grok 4.6: 61 AI Points at 60% Lower Cost Than GPT-5.6 | new_model_or_version | Quelle |
| 2026-08-12 | xAI / Grok | Grok 4.6 Grok 4.6: xAI Launches New Model for Autonomous AI Agents | new_model_or_version | Quelle |
| 2026-08-12 | Google / Gemini | Grok Imagine 2.0 Grok Imagine 2.0: Second-Best AI Image Generator 2026 | new_model_or_version | Quelle |
| 2026-08-12 | OpenAI | Grok Bot as AI agent race shifts toward autonomous work SpaceXAI SpaceXAI launches Grok Bot in AI agent race | new_model_or_version | Quelle |
| 2026-08-11 | Mistral AI | Mistral AI lanza IA soberana europea Mistral AI launches European AI infrastructure with 1 GW capacity | new_model_or_version | Quelle |
| 2026-08-11 | xAI / Grok | Grok Bot as 24 xAI launches Grok Bot as 24/7 coworker with its own virtual computer | preview_or_experimental | Quelle |
| 2026-08-11 | xAI / Grok | Grok Bot SpaceXAI Launches Grok Bot: AI Agents Work Autonomously | new_model_or_version | Quelle |
| 2026-08-11 | Google / Gemini | Gemini Spark COM360 Launches Universal AI Executive Assistant | new_model_or_version | Quelle |
| 2026-08-11 | OpenAI | GPT-5.6 OpenAI launches GPT-5.6-Cyber to combat AI-driven attacks | new_model_or_version | Quelle |
| 2026-08-10 | DeepSeek | Qwen3.8 vs Kimi K3 vs DeepSeek V4 Qwen3.8 vs Kimi K3 vs DeepSeek V4: Open Weights Stopped Being Free at $20 Million | local_or_quantized_update | Quelle |
| 2026-08-10 | DeepSeek | DeepSeek V4-Flash Broke 4-Bit Quantization DeepSeek V4-Flash Broke 4-Bit Quantization: Q4 Is Only 4% Smaller Than Q8 | local_or_quantized_update | Quelle |
| 2026-08-10 | Anthropic | Claude Code Docker Sandboxes: Run AI agents like Claude Code securely | new_model_or_version | Quelle |
| 2026-08-10 | Microsoft / Azure | Phi Silica PowerToys 0.101 Preview: Local AI for Advanced Paste | preview_or_experimental | Quelle |
| 2026-08-09 | OpenAI | GPT Image xAI Launches Imagine Image 2.0 | new_model_or_version | Quelle |
External benchmark sources
These sources are used as orientation. Public performance values are shown only with measurement status and comparability.
| Source | Type | Focus | Assessment |
|---|---|---|---|
| LocalScore / OpenBenchmarking | external_reproducible | Generation speed, TTFT, Prompt speed, hardware comparison | Methodically useful external benchmark source. Values must be matched by model, quantization, runtime and hardware profile. |
| QuelLLM.fr Benchmarks | external_public_result | RTX 5090, RTX 4090, Mac, CPU, llama.cpp, Q4 | Useful for hardware orientation. CheckCom displays it only as externally measured orientation. |
| Öffentliche RTX-5090-LLM-Benchmarks / GitHub | community_benchmark | RTX 5090, VRAM, Power, Tokens/s, LM Studio / llama.cpp | Highly relevant for RTX-5090-class systems, but documentation quality varies by repository. |
| Ollama lokale API | local_metadata | installed models, size, digest, family, parameter size, quantization level | Very useful for local model inventory. Speed requires a separate local benchmark run. |
| Raspberry Pi AI HAT+ 2 Dokumentation | official_hardware_docs | Raspberry Pi 5, AI HAT+ 2, Hailo-10H, LLM/VLM, Edge AI | Official hardware source for suitability and limits, not a complete model benchmark table. |