Was diese Übersicht leistet
Lokal statt API?
CheckCom ordnet ein, wann lokale Workstation-, Edge- oder API-Routen für Unternehmen prüfenswert sind.
Quantisierung verständlich
Q4, Q5, Q6, Q8, GGUF und weitere Signale werden als Hardware- und Eignungshinweise ausgewertet.
Messstatus sichtbar
Öffentliche Werte werden als Metadaten, externe Orientierung oder später als CheckCom-eigene Messung gekennzeichnet.
Hardwareprofile
| Hardware | Geeignet für | Grenzen | Messstatus |
|---|---|---|---|
| RTX 5090 / 32 GB VRAM | 7B–35B quantisierte LLMs, Coding, lokale RAG-Tests, schnelle lokale Inferenz | sehr große 70B+ Modelle nur mit starker Quantisierung/Offload; mehrere parallele Nutzer prüfen | benchmark_prepared |
| RTX 4090 / 24 GB VRAM | kleinere bis mittlere quantisierte Modelle und lokale Experimente | 32B+ und lange Kontexte häufig knapp | external_and_future_checkcom |
| RTX Pro / 96 GB VRAM | große lokale Modelle, längere Kontexte, professionelle lokale Inferenz | hohe Anschaffungskosten; TCO prüfen | profile_prepared |
| Raspberry Pi 5 + AI HAT+ 2 / Hailo-10H | Edge-KI, kleine LLM-/VLM-Kandidaten, Vision, lokale Demo- und Datenschutzszenarien | kein Ersatz für große GPU-LLMs; Hailo-kompatible Modelle und Runtime nötig | demo_profile_available |
| Raspberry Pi 5 CPU-only | kleine CPU-Tests, Klassifikation, Offline-Demos | große LLMs und lange Kontexte nicht sinnvoll | inventory_prepared |
| API / Cloud / Enterprise | Skalierung, Enterprise-Verträge, hohe Parallelität und schnelle Produktivstarts | laufende Token-/Seat-Kosten, Datenroute und Vertrag prüfen | price_radar_linked |
Lokale und quantisierte Modellkandidaten
Aus dem CheckCom-Modellkatalog abgeleitet. Messstatus „metadata_only“ bedeutet: noch kein CheckCom-Benchmarkwert.
| Modell | Params B | Quantisierung | VRAM GB | RTX 5090 | Pi 5 / AI HAT+ 2 | Score |
|---|---|---|---|---|---|---|
| Qwen2.5-Coder 7B Instruct Alibaba / Qwen |
7.0 | Q4 Q4 |
8.0 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
95/100 metadata_only |
| Qwen3.6 35B A3B Vocabulary Trimming GGUF Alibaba / Qwen |
35.0 | GGUF Q4_K_S, GGUF, gguf |
28.2 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 35B A3B Uncensored xCloud GGUF Alibaba / Qwen |
35.0 | GGUF Q4_K_M, GGUF, gguf |
28.2 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 35B A3B TQ2Q GGUF Alibaba / Qwen |
35.0 | GGUF GGUF, gguf |
28.2 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 35B A3B StyleTune GGUF Alibaba / Qwen |
35.0 | GGUF IQ4_XS, GGUF, gguf |
28.2 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 35B A3B Q5 K M GGUF Alibaba / Qwen |
35.0 | GGUF q5_k_m, Q5, GGUF |
32.7 | offload nur mit Offload oder reduzierten Einstellungen realistisch |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 35B A3B Ornith 80 20 Beta GGUF Alibaba / Qwen |
35.0 | GGUF Q4_K_S, GGUF, gguf |
28.2 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 35B A3B NVFP4 Gguf Alibaba / Qwen |
35.0 | GGUF Gguf, gguf |
28.2 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 35B A3B NVFP4 GGUF Alibaba / Qwen |
35.0 | GGUF q3, GGUF, gguf |
22.3 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 35B A3B MTP GGUF Tater NoThink Alibaba / Qwen |
35.0 | GGUF Q4_K_M, GGUF, gguf |
28.2 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 35B A3B DS4 GGUF Alibaba / Qwen |
35.0 | GGUF Q4_K_S, GGUF, gguf |
28.2 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 35B A3B Claude 4.7 Distill NVFP4 GGUF Alibaba / Qwen |
35.0 | GGUF GGUF, gguf |
28.2 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 35B A3B Claude 4.7 Distill MXFP4 MoE GGUF Alibaba / Qwen |
35.0 | GGUF GGUF, gguf |
28.2 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.5 35B A3B BitClass3 GGUF Alibaba / Qwen |
35.0 | GGUF Q3_K_S, GGUF, gguf |
22.3 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.5 35B A3B APEX GGUF Alibaba / Qwen |
35.0 | GGUF GGUF, gguf |
28.2 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Huihui Qwen AgentWorld 35B A3B Abliterated UD Q3 K M GGUF Alibaba / Qwen |
35.0 | GGUF Q3_K_M, Q3, GGUF |
22.3 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Endy Qwen3.6 CyberSec 35B A3B GGUF Alibaba / Qwen |
35.0 | GGUF Q2_K, GGUF, gguf |
18.7 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3 32B Qwople GGUF Alibaba / Qwen |
32.0 | GGUF Q4_K_M, GGUF, gguf |
26.3 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3 32B GGUF Alibaba / Qwen |
32.0 | GGUF IQ3_M, GGUF, gguf |
20.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen2.5 Coder 32B Python Specialist I1 GGUF Alibaba / Qwen |
32.0 | GGUF IQ3_M, GGUF, gguf |
20.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen2.5 Coder 32B Python Specialist GGUF Alibaba / Qwen |
32.0 | GGUF IQ4_XS, GGUF, gguf |
26.3 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen2.5 Coder 32B Instruct GGUF Alibaba / Qwen |
32.0 | GGUF Q2_K, GGUF, gguf |
17.7 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen2.5 32B Instruct AdiTurbo GGUF Alibaba / Qwen |
32.0 | GGUF Q3_K_M, GGUF, gguf |
20.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| DeepSeek R1 Distill Qwen32B Q4 K M GGUF Alibaba / Qwen |
32.0 | GGUF q4_k_m, Q4, GGUF |
26.3 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3 Coder 30B A3B Instruct GGUF Alibaba / Qwen |
30.0 | GGUF Q4_K_M, GGUF, gguf |
25.1 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3 30B A3B GGUF Alibaba / Qwen |
30.0 | GGUF IQ3_M, GGUF, gguf |
20.0 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Ektome Qwen3 30B A3B PristinelyUncensored GGUF Alibaba / Qwen |
30.0 | GGUF Q2_K, GGUF, gguf |
17.0 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.8 27B ZipBrain GGUF Alibaba / Qwen |
27.0 | GGUF IQ4_XS, GGUF, gguf |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.8 27B Uncensored Q4 K M GGUF Alibaba / Qwen |
27.0 | GGUF Q4_K_M, Q4, GGUF |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.8 27B Uncensored Aggressive I1 GGUF Alibaba / Qwen |
27.0 | GGUF IQ3_M, GGUF, gguf |
18.6 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.8 27B Uncensored Aggressive GGUF Alibaba / Qwen |
27.0 | GGUF IQ4_XS, GGUF, gguf |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.8 27B ROCmFPX GGUF Alibaba / Qwen |
27.0 | GGUF GGUF, gguf |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.8 27B I1 GGUF Alibaba / Qwen |
27.0 | GGUF IQ3_M, GGUF, gguf |
18.6 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.8 27B Fable5 Distill Abliterated Imatrix GGUF Alibaba / Qwen |
27.0 | GGUF GGUF, gguf |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.8 27B Fable Distill GGUF Alibaba / Qwen |
27.0 | GGUF IQ4_XS, GGUF, gguf |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.8 27B Cold Fusion GAIN V1.1 Q4 K M GGUF Alibaba / Qwen |
27.0 | GGUF q4_k_m, Q4, GGUF |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.8 27B Abliterated MTP GGUF Alibaba / Qwen |
27.0 | GGUF IQ2_M, GGUF, gguf |
15.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.8 27B 5090 Goldilocks GGUF Alibaba / Qwen |
27.0 | GGUF GGUF, gguf |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 27B Samantha Uncensored GGUF Alibaba / Qwen |
27.0 | GGUF Q4_K_M, GGUF, gguf |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 27B Q4 K M GGUF Alibaba / Qwen |
27.0 | GGUF q4_k_m, Q4, GGUF |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 27B NVFP4 GGUF Alibaba / Qwen |
27.0 | GGUF q3, GGUF, gguf |
18.6 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 27B Mtp Gguf Alibaba / Qwen |
27.0 | GGUF Q4_K_X, Gguf, Q4_K_XL |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 27B KR MTP GGUF Alibaba / Qwen |
27.0 | GGUF Q5_K_X, GGUF, Q5_K_XL |
26.7 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 27B KR GGUF Alibaba / Qwen |
27.0 | GGUF Q5_K_X, GGUF, Q5_K_XL |
26.7 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 27B GGUF Alibaba / Qwen |
27.0 | GGUF Q4_K_M, GGUF, gguf |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 27B Esper4 I1 GGUF Alibaba / Qwen |
27.0 | GGUF IQ3_M, GGUF, gguf |
18.6 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 27B Esper4 GGUF Alibaba / Qwen |
27.0 | GGUF IQ4_XS, GGUF, gguf |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 27B Architect Polaris2 Fable B F451 MTP ROCmFPX GGUF Alibaba / Qwen |
27.0 | GGUF Q6, GGUF, gguf |
30.8 | tight voraussichtlich knapp; Kontext/KV-Cache und Runtime prüfen |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 27B AEON RYS Agentic Coder PatchCode GGUF Alibaba / Qwen |
27.0 | GGUF IQ4_NL, GGUF, gguf |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Ostrich 27B Qwen3.8 260815 I1 GGUF Alibaba / Qwen |
27.0 | GGUF IQ3_M, GGUF, gguf |
18.6 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Huihui Qwen3.8 27B Abliterated GGUF Alibaba / Qwen |
27.0 | GGUF IQ3_M, GGUF, gguf |
18.6 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Elster Vernunft Qwen3.6 27B GGUF Alibaba / Qwen |
27.0 | GGUF IQ4_XS, GGUF, gguf |
23.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Dagger Qwen3.6 27B GGUF MTP Alibaba / Qwen |
27.0 | GGUF Q3_K_M, GGUF, gguf |
18.6 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.8 20B Minitron Q4 K M GGUF Alibaba / Qwen |
20.0 | GGUF q4_k_m, Q4, GGUF |
18.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.6 14B A3B VibeForged V2 GGUF Alibaba / Qwen |
14.0 | GGUF F16, GGUF, gguf |
20.4 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3 14B PragReST I1 GGUF Alibaba / Qwen |
14.0 | GGUF IQ1_M, GGUF, gguf |
20.4 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen2.5 Coder 14B Instruct Jbliterated Alibaba / Qwen |
14.0 | GGUF Q4_K_M, gguf |
13.7 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen2.5 Coder 14B Instruct Heretic GGUF Alibaba / Qwen |
14.0 | GGUF IQ4_XS, GGUF, gguf |
13.7 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen2.5 Coder 14B Instruct GGUF Alibaba / Qwen |
14.0 | GGUF f16, GGUF, gguf |
20.4 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen2.5 Coder 14B Instruct GGUF Alibaba / Qwen |
14.0 | GGUF fp16, GGUF, gguf |
34.4 | offload nur mit Offload oder reduzierten Einstellungen realistisch |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen2.5 14B Instruct Heretic Alibaba / Qwen |
14.0 | GGUF Q4_K_M, gguf |
13.7 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Lfed Qwen2.5 Coder 14B Sql Gguf Alibaba / Qwen |
14.0 | GGUF Q4_K_M, Gguf, gguf |
13.7 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Iol Qwen2.5 14B Instruct AWQ Alibaba / Qwen |
14.0 | AWQ AWQ |
13.7 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| H100 Qwen3 14B MonEspaceSante CPT I1 GGUF Alibaba / Qwen |
14.0 | GGUF IQ2_M, GGUF, gguf |
9.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| H100 Qwen3 14B MonEspaceSante CPT GGUF Alibaba / Qwen |
14.0 | GGUF IQ4_XS, GGUF, gguf |
13.7 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Ektome Qwen2.5 Coder 14B Instruct PristinelyUncensored Alibaba / Qwen |
14.0 | GGUF IQ3_M, gguf |
11.3 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Adi Qwen2.5 14B GLM 5.2 General GGUF Alibaba / Qwen |
14.0 | GGUF q4_k_m, GGUF, gguf |
13.7 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| QwenPaw Flash 9B Heretic MTP GGUF Alibaba / Qwen |
9.0 | GGUF BF16, GGUF, gguf |
23.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| QwenPaw Flash 9B Heretic Imatrix GGUF Alibaba / Qwen |
9.0 | GGUF BF16, GGUF, gguf |
23.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| QwenPaw Flash 9B Heretic GGUF Alibaba / Qwen |
9.0 | GGUF BF16, GGUF, gguf |
23.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.5 9B Ultra Uncensored Heretic V2 I1 GGUF Alibaba / Qwen |
9.0 | GGUF IQ1_M, GGUF, gguf |
14.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.5 9B Ultra Uncensored Heretic V1 I1 GGUF Alibaba / Qwen |
9.0 | GGUF IQ1_M, GGUF, gguf |
14.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.5 9B The Defiant Fable Uncensored Heretic NEO IMATRIX MAX MTP GGUF Alibaba / Qwen |
9.0 | GGUF BF16, GGUF, gguf |
23.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.5 9B The Defiant Fable Uncensored Heretic NEO IMATRIX MAX GGUF Alibaba / Qwen |
9.0 | GGUF BF16, GGUF, gguf |
23.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.5 9B Samantha Uncensored GGUF Alibaba / Qwen |
9.0 | GGUF Q4_K_M, GGUF, gguf |
10.6 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.5 9B Heretic Imatrix GGUF Alibaba / Qwen |
9.0 | GGUF BF16, GGUF, gguf |
23.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.5 9B Heretic GGUF Alibaba / Qwen |
9.0 | GGUF BF16, GGUF, gguf |
23.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.5 9B Haskell Rust Python Q6 K GGUF Alibaba / Qwen |
9.0 | GGUF q6_k, Q6, GGUF |
13.1 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.5 9B Haskell Rust Python IQ4 XS GGUF Alibaba / Qwen |
9.0 | GGUF iq4_xs, GGUF, gguf |
10.6 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.5 9B GGUF Alibaba / Qwen |
9.0 | GGUF IQ1_KT, GGUF, gguf |
14.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.5 9B DeepSeek V4 Flash GGUF Alibaba / Qwen |
9.0 | GGUF Q3_K_M, GGUF, gguf |
9.0 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3.5 9B DFlash GGUF Alibaba / Qwen |
9.0 | GGUF bf16, DFlash, GGUF |
23.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Huihui Qwen3.5 9B Claude 4.6 Opus Abliterated Heretic GGUF Alibaba / Qwen |
9.0 | GGUF f16, GGUF, gguf |
14.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| FACET Terminal Qwen3.5 9B GGUF Alibaba / Qwen |
9.0 | GGUF f16, GGUF, gguf |
14.9 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Adi Qwen3.5 9B GLM 5.2 General GGUF Alibaba / Qwen |
9.0 | GGUF q4_k_m, GGUF, gguf |
10.6 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| VideoKR Qwen3 VL 8B GGUF Alibaba / Qwen |
8.0 | GGUF f16, GGUF, gguf |
13.8 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| SenseNova U1.5 8B GGUFs Hugging Face / sonstige Owner |
8.0 | GGUF Q2_K, gguf |
7.8 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| RavenX Conjecture Qwen3 8B GGUF Alibaba / Qwen |
8.0 | GGUF Q8, GGUF, Q8_0 |
13.8 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3 VL 8B Instruct Heretic GGUF Alibaba / Qwen |
8.0 | GGUF f16, GGUF, gguf |
13.8 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3 VL 8B Instruct Abliterated V2 GGUF Alibaba / Qwen |
8.0 | GGUF f16, GGUF, gguf |
13.8 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3 Embed 8B GGUF Alibaba / Qwen |
8.0 | GGUF iq4_xs, GGUF, gguf |
10.0 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3 8B UnBias Plus SFT Instruct Legacy GGUF Alibaba / Qwen |
8.0 | GGUF f16, GGUF, gguf |
13.8 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3 8B PragReST I1 GGUF Alibaba / Qwen |
8.0 | GGUF IQ1_M, GGUF, gguf |
13.8 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3 8B PragReST GGUF Alibaba / Qwen |
8.0 | GGUF f16, GGUF, gguf |
13.8 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3 8B MedReasonPath GGUF Alibaba / Qwen |
8.0 | GGUF f16, GGUF, gguf |
13.8 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3 8B Heretic GGUF Alibaba / Qwen |
8.0 | GGUF f16, GGUF, gguf |
13.8 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3 8B Gguf Alibaba / Qwen |
8.0 | GGUF Q8, Gguf, Q8_0 |
13.8 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3 8B Apostate I1 GGUF Alibaba / Qwen |
8.0 | GGUF IQ1_S, GGUF, gguf |
13.8 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen3 8B Apostate GGUF Alibaba / Qwen |
8.0 | GGUF f16, GGUF, gguf |
13.8 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| PDP Qwen3 8B SFT GGUF Alibaba / Qwen |
8.0 | GGUF f16, GGUF, gguf |
13.8 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| LFM2.5 8B A1B GGUF LFM2.5 |
8.0 | GGUF BF16, GGUF, gguf |
21.8 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| L40S Qwen3 8B MonEspaceSante CPT SFT Anti Hallucination I1 GGUF Alibaba / Qwen |
8.0 | GGUF IQ2_S, GGUF, gguf |
7.8 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Huihui Qwen3 VL 8B Instruct Abliterated FP8 Alibaba / Qwen |
8.0 | FP8 FP8 |
10.0 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Huihui Qwen3 VL 8B Instruct Abliterated AWQ Int4 Alibaba / Qwen |
8.0 | AWQ AWQ, Int4, int4 |
10.0 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Adi Qwen3 8B GLM 5.2 General GGUF Alibaba / Qwen |
8.0 | GGUF q4_k_m, GGUF, gguf |
10.0 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Yoiko Img Prompt Gen Qwen2.5 7B Instruct Q4 K M Alibaba / Qwen |
7.0 | GGUF Q4_K_M, Q4, gguf |
9.3 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| VideoKR Qwen2.5 VL 7B GGUF Alibaba / Qwen |
7.0 | GGUF f16, GGUF, gguf |
12.7 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Typhoon2 Qwen2.5 7B Instruct GGUF Alibaba / Qwen |
7.0 | GGUF Q4_K_M, GGUF, gguf |
9.3 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Stiles Seymens Qwen2.5 7B Generalist Q4KM GGUF Alibaba / Qwen |
7.0 | GGUF Q4_K_M, GGUF, gguf |
9.3 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Sofia Qwen2.5 7B I1 GGUF Alibaba / Qwen |
7.0 | GGUF IQ1_M, GGUF, gguf |
12.7 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Sofia Qwen2.5 7B GGUF Alibaba / Qwen |
7.0 | GGUF f16, GGUF, gguf |
12.7 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen2.5 VL 7B Instruct Int8 Convrot Comfyui Alibaba / Qwen |
7.0 | Int8 Int8, int8 |
9.3 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen2.5 Coder 7B VN Master Polymath FINAL GGUF Alibaba / Qwen |
7.0 | GGUF Q3_K_M, GGUF, gguf |
8.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen2.5 Coder 7B Instruct Jbliterated Alibaba / Qwen |
7.0 | GGUF Q4_K_M, gguf |
9.3 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen2.5 Coder 7B Instruct Heretic Alibaba / Qwen |
7.0 | GGUF Q4_K_M, gguf |
9.3 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen2.5 Coder 7B Instruct GGUF Alibaba / Qwen |
7.0 | GGUF fp16, GGUF, gguf |
19.7 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen2.5 Coder 7B Instruct GGUF Alibaba / Qwen |
7.0 | GGUF q3_k_m, GGUF, gguf |
8.2 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen2.5 7B Instruct INT8 Quanto Alibaba / Qwen |
7.0 | INT8 INT8 |
9.3 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen2.5 7B Instruct Heretic Alibaba / Qwen |
7.0 | GGUF Q4_K_M, gguf |
9.3 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
| Qwen2.5 7B Instruct FP8 Dynamic TR171 Alibaba / Qwen |
7.0 | FP8 FP8 |
9.3 | good passt voraussichtlich gut in das VRAM-Profil |
not_recommended für AI HAT+ 2 eher zu groß oder nicht als Hailo-kompatibel belegt |
85/100 metadata_only |
Lokales Ollama-Inventar
Status: unavailable · Modelle: 0
Falls Ollama auf der Workstation läuft, werden installierte Modelle mit Größe, Digest, Familie und Quantisierung ausgelesen. Geschwindigkeit wird erst durch Benchmarkläufe gemessen.
API, Workstation oder Edge?
CheckCom nutzt diese Daten, um im KI-Auswahlassistenten lokale Alternativen zu API-/Cloud-Routen sichtbar zu machen. Für Unternehmen zählt nicht nur die Modellgröße, sondern auch Betriebsaufwand, Datenschutz, Latenz, Wartung und Auslastung.
Lokale / quantisierte News- und Release-Signale
| Datum | Anbieter | Signal | Typ | Quelle |
|---|---|---|---|---|
| 2026-08-20 | Moonshot / Kimi | Kimi K3 and GLM-5.3 are now Frontier Radar #4: Chinas KI-Modelle schließen auf | local_or_quantized_update | Quelle |
| 2026-08-20 | OpenAI | GPT-5.6 AWS bringt OpenAI-GPT-5.6-Modelle nach Indien | new_model_or_version | Quelle |
| 2026-08-19 | Google / Gemini | Gemma 4 Lokale KI schneller machen: So zünden Sie den LLM-Turbo | availability_or_api_update | Quelle |
| 2026-08-19 | OpenAI | GPT ChatGPT verändert Suchverhalten von Studierenden | availability_or_api_update | Quelle |
| 2026-08-18 | Google / Gemini | Gemini AI Googles Gemini AI scannt Workspace-Daten standardmäßig | local_or_quantized_update | Quelle |
| 2026-08-18 | OpenAI | GPT OpenAI wechselt zum B2B-Fokus | availability_or_api_update | Quelle |
| 2026-08-18 | OpenAI | Qwen3.8-27B Qwen3.8-27B: 3 Millionen Downloads in drei Tagen | local_or_quantized_update | Quelle |
| 2026-08-17 | xAI / Grok | Grok Bot SpaceXAI stellt Grok Bot für autonome KI-Agenten vor | new_model_or_version | Quelle |
| 2026-08-16 | OpenAI | GPT-5.6 OpenAI lanciert Ultrafast: GPT-5.6 Sol mit 750 Tokens pro Sekunde | new_model_or_version | Quelle |
| 2026-08-16 | Alibaba / Qwen | Qwen 3.8 27B Qwen 3.8 27B: Ausgezeichnet, aber zu überlegend | new_model_or_version | Quelle |
| 2026-08-16 | Alibaba / Qwen | Qwen3.8-27B Alibaba Alibaba veröffentlicht Qwen3.8-27B für lokale Nutzung | new_model_or_version | Quelle |
| 2026-08-16 | OpenAI | GPT OpenAI: 8,3-fache Nutzungslücke bei KI-Agenten in Unternehmen | availability_or_api_update | Quelle |
| 2026-08-15 | Google / Gemini | Qwen AI models have surpassed 3 billion global downloads in six months Alibabas KI-Modelle übertrumpfen Meta und Google | local_or_quantized_update | Quelle |
| 2026-08-14 | Mistral AI | Mistral erweitert Angebot Die Organisation KI-Update: Hate Aid kritisiert KI-Brillen, Mistral erweitert Angebot | local_or_quantized_update | Quelle |
| 2026-08-14 | OpenAI | Qwen team Alibaba veröffentlicht Qwen 3.8 mit offenen Modellgewichten | new_model_or_version | Quelle |
| 2026-08-14 | OpenAI | GPT-5.6 OpenAI lanciert Ultrafast-Modus für GPT-5.6 Sol | new_model_or_version | Quelle |
| 2026-08-14 | Zhipu / GLM | GLM-5.3 Zhipu AI veröffentlicht GLM-5.3 als stärkstes Open-Weights-Coding-Modell | new_model_or_version | Quelle |
| 2026-08-14 | OpenAI | GPT-5.6 OpenAI bringt GPT-5.6 Sol mit neuem Ultrafast-Modus | preview_or_experimental | Quelle |
| 2026-08-13 | OpenAI | GPT-5.6 OpenAI: GPT-5.6 Sol mit 14-facher Geschwindigkeit dank Cerebras | new_model_or_version | Quelle |
| 2026-08-13 | OpenAI | Grok 4.6 iguala a GPT-5.6 Sol con 61 puntos y menor coste para startups Grok 4.6 iguala a GPT-5.6 So Grok 4.6 erreicht GPT-5.6 Sol-Niveau mit geringeren Kosten | new_model_or_version | Quelle |
| 2026-08-13 | Mistral AI | Mistral als KI Mistral baut KI-Infrastruktur in Europa aus | local_or_quantized_update | Quelle |
| 2026-08-13 | Anthropic | Grok 4.6 Grok 4.6: Effizienzvorteil für Startups | local_or_quantized_update | Quelle |
| 2026-08-13 | OpenAI | GPT-5.6 OpenAI stellt GPT-5.6 Sol mit 14-facher Geschwindigkeit vor | new_model_or_version | Quelle |
| 2026-08-13 | OpenAI | Gemini vor KI-Geisterbeschwörung: Traditionelle Wahrsager in Südkorea verlieren an Bedeutung | availability_or_api_update | Quelle |
| 2026-08-13 | OpenAI | Kimi K3 Kimi K3: Open-Source-KI-Modell mit 2,8 Billionen Parametern | availability_or_api_update | Quelle |
| 2026-08-12 | DeepSeek | DeepSeek V4 Pro 0813 DeepSeek V4 Pro 0813 über OpenRouter verfügbar | new_model_or_version | Quelle |
| 2026-08-12 | OpenAI | Grok 4.6 de SpaceXAI Grok 4.6 von SpaceXAI: 61 Punkte in IA bei 60 % geringerem Kosten als GPT-5.6 | new_model_or_version | Quelle |
| 2026-08-12 | xAI / Grok | Grok 4.6 Grok 4.6: xAI schickt neues Modell für autonome KI-Agenten ins Rennen | new_model_or_version | Quelle |
| 2026-08-12 | Google / Gemini | Grok Imagine 2.0 Grok Imagine 2.0: Zweitschnellste KI für Bildgenerierung 2026 | new_model_or_version | Quelle |
| 2026-08-12 | OpenAI | Grok Bot as AI agent race shifts toward autonomous work SpaceXAI SpaceXAI startet Grok Bot im AI-Agenten-Rennen | new_model_or_version | Quelle |
| 2026-08-11 | Mistral AI | Mistral AI lanza IA soberana europea Mistral AI lanciert europäische KI-Infrastruktur mit 1 GW Kapazität | new_model_or_version | Quelle |
| 2026-08-11 | xAI / Grok | Grok Bot as 24 xAI startet Grok Bot als 24/7-Mitarbeiter mit eigenem virtuellen Computer | preview_or_experimental | Quelle |
| 2026-08-11 | xAI / Grok | Grok Bot SpaceXAI startet Grok Bot: KI-Agenten arbeiten autonom | new_model_or_version | Quelle |
| 2026-08-11 | Google / Gemini | Gemini Spark COM360 lanciert universellen KI-Exekutiv-Assistenten | new_model_or_version | Quelle |
| 2026-08-11 | OpenAI | GPT-5.6 OpenAI lanciert GPT-5.6-Cyber zur Abwehr von KI-Angriffen | new_model_or_version | Quelle |
| 2026-08-10 | DeepSeek | Qwen3.8 vs Kimi K3 vs DeepSeek V4 Qwen3.8 vs. Kimi K3 vs. DeepSeek V4: Offene Gewichte kosten ab 20 Millionen | local_or_quantized_update | Quelle |
| 2026-08-10 | DeepSeek | DeepSeek V4-Flash Broke 4-Bit Quantization DeepSeek V4-Flash Broke 4-Bit Quantization: Q4 Is Only 4% Smaller Than Q8 | local_or_quantized_update | Quelle |
| 2026-08-10 | Anthropic | Claude Code Docker Sandboxes: IA-Agenten wie Claude Code sicher ausführen | new_model_or_version | Quelle |
| 2026-08-10 | Microsoft / Azure | Phi Silica PowerToys 0.101: Lokale KI für Advanced Paste | preview_or_experimental | Quelle |
| 2026-08-09 | OpenAI | GPT Image xAI stellt Imagine Image 2.0 vor | new_model_or_version | Quelle |
Externe Benchmarkquellen
Diese Quellen dienen zunächst als Orientierung. Öffentliche Leistungswerte werden nur mit Messstatus und Vergleichbarkeit angezeigt.
| Quelle | Typ | Fokus | Einordnung |
|---|---|---|---|
| LocalScore / OpenBenchmarking | external_reproducible | Generation speed, TTFT, Prompt speed, hardware comparison | Methodisch interessante externe Benchmarkquelle. Werte müssen nach Modell, Quantisierung, Runtime und Hardwareprofil abgeglichen werden. |
| QuelLLM.fr Benchmarks | external_public_result | RTX 5090, RTX 4090, Mac, CPU, llama.cpp, Q4 | Nützlich für schnelle Hardwareorientierung. Veröffentlichung auf CheckCom nur mit Kennzeichnung als extern gemessen. |
| Öffentliche RTX-5090-LLM-Benchmarks / GitHub | community_benchmark | RTX 5090, VRAM, Power, Tokens/s, LM Studio / llama.cpp | Sehr relevant für RTX-5090-Klassen, aber je nach Repository unterschiedlich gut dokumentiert. |
| Ollama lokale API | local_metadata | installed models, size, digest, family, parameter size, quantization level | Sehr gut für lokale Modellinventarisierung. Geschwindigkeit entsteht erst durch eigene Laufzeitmessung. |
| Raspberry Pi AI HAT+ 2 Dokumentation | official_hardware_docs | Raspberry Pi 5, AI HAT+ 2, Hailo-10H, LLM/VLM, Edge AI | Offizielle Hardwarequelle für Eignung und Grenzen, keine vollständige Modell-Benchmarktabelle. |