GLM 5.2
Hugging Face Hacked by AI Agent — Saved by a Local Model (2026)
Hugging Face says an autonomous AI agent breached its internal infra on July 16. The models you download are safe — here's what was and wasn't hit.
Hugging Face got hacked -- Open local AI came to the rescue!
Two trillion-parameter 'open' models dropped in ten days — Qwen 3.8 and Kimi K3 — and you can't run either. The same week, Hugging Face got breached and its own responders, blocked by commercial-model guardrails, ran the forensics on GLM 5.2, an open-weight model on their own hardware. Same lesson from opposite ends — plus what to actually run on a 24GB card today.
Inkling 975B vs Your 3090: The Real Memory Math (2026)
Inkling's smallest 1-bit quant is 270GB. A maxed consumer desktop holds 256GB. The open frontier left consumer hardware behind. Here's the honest math.
China May Restrict Its AI Exports — Your Local Models Don't Care
China's commerce ministry met Alibaba, ByteDance and Z.ai about walling off top AI models. The US did the same in June. The weights on your disk don't care.
Qwen 3.7's open weights are overdue — by the math, not vibes
Qwen's own release cadence says the 3.7 open weights should already be out, and they're not. Plus GLM-5.2 running locally: a frontier open model that takes serious hardware.
How to Run GLM 5.2 Locally: GPU, VRAM & Quant Guide
GLM 5.2 is 753B params and 1.51TB at full precision. Run it locally: the live Unsloth quant ladder, every GPU and RAM path, and the quant to actually target.