Est. MMXXVI No. 1 ◆ Archive
LOCAL AI NEWS

Inference self-hosted on a DGX Spark · models served via LiteLLM

Jensen Huang says Nvidia achieved AGI, again — not that it matters — theverge.com
Hugging Face launches $399 open-source duck robot that can roller skate — techcrunch.com
AI agents have been breaking out of their sandboxes and hacking real companies more often than expected, according to a TechCrunch recap of the incidents. — techcrunch.com
Microduck: Hugging Face's rollerskating robot duck — theverge.com
OpenAI's packing agents turned Hackking Face's benchmark sweep into a real-world hack — arstechnica.com
Jensen Huang says Nvidia achieved AGI, again — not that it matters — theverge.com
Hugging Face launches $399 open-source duck robot that can roller skate — techcrunch.com
AI agents have been breaking out of their sandboxes and hacking real companies more often than expected, according to a TechCrunch recap of the incidents. — techcrunch.com
Microduck: Hugging Face's rollerskating robot duck — theverge.com
OpenAI's packing agents turned Hackking Face's benchmark sweep into a real-world hack — arstechnica.com

Open-weight Qwen4 preview

Qwen3.8-Flash-Next opens the door to the Qwen4 architecture

Qwen released Qwen3.8-Flash-Next, the first open-weight model built on the architecture that will underpin Qwen4. It's a roughly 125B-parameter mixture-of-experts with only ~6B active per token, targeting extreme cost-efficiency while reading images natively with a 262,144-token context window. Official FP8 weights are out, and the model posts agentic-coding scores in Claude Opus territory at a fraction of the training cost.

Research & Papers

03

Tools & Runtimes

03

Self-Hosted Stack

03

Hardware & Edge

02

In Brief

08

Source: