Zum Hauptinhalt springen
tsecurity.de LIVE
Echtzeit-Radar & Feeds
Alle RSS Feeds
👥 Community & Social
Intelligence View
⚡ tsecurity.de Intelligence

Hy4-Preview Just Shrunk From 1.5TB to 200GB

YouTube Video

Want to make money and save time with AI? Join here: https://www.skool.com/ai-profit-lab-7462/about

Video notes + links to the tools 👉 https://www.skool.com/ai-profit-lab-7462/about

Get a FREE AI Course + Community + 1,000 AI Agents 👉 https://www.skool.com/ai-seo-with-julian-goldie-1553/about

Get a FREE AI SEO Strategy Session: https://go.juliangoldie.com/strategy-session?utm=julian

Tencent shrunk its 770B-parameter HY4 Preview from 1.5TB to 213GB — but the number everyone's leaving out of the thumbnail changes the whole story. I break down exactly how the mixed-precision quantization works, what it actually scores on benchmarks, and the one line in the docs that stops most people from running it. If you care about open models and where local AI is really headed, this is the honest version.

00:00 Intro – 1.5TB to 200GB, and the missing number
00:40 The Real Math – Why HY4 Preview is 1.56TB
01:06 213GB – The compressed build
01:14 Quantization Explained – In plain English
01:39 The Smart Trick – Mixed precision, not brute force
02:00 The Photo Analogy – Why some layers get protected
02:09 Where the Cuts Go – Experts vs down projection
02:27 Benchmark Results – 85% smaller, ~1 point lost
03:38 The Catch – You still need 214GB of VRAM
03:57 The Better Option – Why Q4_K_M is the default
04:16 Storage vs VRAM – The mistake everyone makes
04:24 Real Speed – 20 tokens/sec on 8 data center GPUs
04:39 Who It's Actually For – Not your desktop
05:19 Inside the Model – MoE, 49B active per token
05:49 Sparse Attention & 1M Context – The efficiency stack
06:06 It Won't Just Run – Stock llama.cpp fails
06:41 The Bigger Pattern – 4 techniques, one direction
07:20 The Real Takeaway – Spend precision where it matters
Zum Aktualisieren ziehen
ZERO-DAY CVE-2026-61591 | djust provides Phoenix LiveView-style reactive server-side rendering for…
Advisory →
TTS Reader • tsecurity.de Voice
tsecurity.de Icon
tsecurity.de App
Offline-Lesen, Eilmeldungen & 0ms Ladezeit

Installiere tsecurity.de direkt auf deinen Home-Bildschirm für das ultimative Vollbild-Magazinerlebnis ohne Browser-Leisten.

Nächster Beitrag
Themen-Radar & Intelligence Matrix
Echtzeit-Taxonomie nach Angriffsvektoren & Plattformen
Community Radar & Live Chat
Sentinel Bot online • Live-Stream
Dein Cluster: Security Explorer
Match:
lädt…
Verbindung zum Community-Stream wird aufgebaut...
Bearbeitungsmodus — Senden überschreibt deine Nachricht
Community-Puls — was gerade passiert
lädt…
Aktivitäten deiner Analysten
lädt…
Neues Thema oder Eilmeldung einreichen

Reiche interessante Links, Zero-Days oder Debatten ein. Die Community entscheidet per Upvote über die Veröffentlichung.

Heiß diskutierte Einreichungen
🔖 Gespeicherte Artikel
📂 Keine gespeicherten Artikel vorhanden.
Rechts: Artikel Ziehen Links: RSS
Hoch: nächster Artikel Runter: zurück / schließen
News NIS-2 Frühwarnung Tier-1 Intel ⏱️ 3 Min vor 10 Min
Artikeldaten werden geladen...

Rechts: Original Links: RSS-Ansicht
↗ Original-Quelle
Social Reaktionen Deine Reaktion zählt
Einstufung & Relevanz-Poll 0 Stimmen
In sozialen Netzwerken teilen 1-Klick