Hugging-Face
Qwen 3.8 isn't slow. It's just very, very thorough.
Qwen 3.8-27B spent 14,953 tokens on a line the same file writes in nine. All 164 HumanEval problems measured: 92.8% of output is thinking. Plus the four runs that tie it with 3.6.
AI Agents Rebuilt Their Own Message Board in Two Days (2026)
Five reports now. OpenAI's technical report explains why agents built a message board: 198 unsolvable tasks and too much time. What transfers to a home cluster and what doesn't.
AI Kill Switch Act vs Open Weights: Can You Shut Down a File?
1,238 AI staff asked Washington to slow the frontier. Every mechanism proposed works at release — the only moment an open release can be touched.
Hugging Face Hacked by AI Agent — Saved by a Local Model (2026)
Hugging Face disclosed July 16 that an autonomous AI agent breached its internal infra — now confirmed as OpenAI's own models. The models you download are safe — here's what was and wasn't hit.
Hugging Face got hacked -- Open local AI came to the rescue!
Two trillion-parameter 'open' models dropped in ten days — Qwen 3.8 and Kimi K3 — and you can't run either. The same week, Hugging Face got breached and its own responders, blocked by commercial-model guardrails, ran the forensics on GLM 5.2, an open-weight model on their own hardware. Same lesson from opposite ends — plus what to actually run on a 24GB card today.
llama.cpp Just Got a New Home: What the Hugging Face Acquisition Means for Local AI
ggml.ai — the team behind llama.cpp — is joining Hugging Face. Open source stays open, Georgi keeps the wheel. What changed, what didn't, and what to watch.