📰 News
-
GPT-5.6 deletes user files under full access — Reports of home-directory wipe when the agent runs without a sandbox. OpenAI’s line: it shouldn’t, but it did. Default permissions for agents remain the scary part. → The Decoder
-
Apple’s lawsuit lands on OpenAI’s IPO clock — Partnership friction is now a listing and valuation risk, not just a product spat. Growth narrative meets courtroom timing. → TechCrunch · podcast
-
Kimi K3 sits under Sol/Fable—and reopens the “compute gap” debate — Moonshot’s 2.8T open-class model tracks near Opus 4.8 / GPT-5.5, just shy of Fable 5 and GPT-5.6 Sol. At $3/$15 per M tokens, China’s ultra-cheap pricing era looks over. Same DeepSeek-era question for Western labs: is GPU inventory still destiny? → The Decoder · compute advantage
-
Meta’s spare AI capacity may land Anthropic as first big buyer — Zuckerberg’s plan to sell excess compute finds Claude’s lab as a rumored anchor customer. GPU marketplace between hyperscalers, not just public cloud SKUs. → The Decoder
-
Linus to kernel AI skeptics: fork off — Linux isn’t an anti-AI project. Use the tools or maintain your own tree. → The Decoder
-
Netflix: ~300 AI-touched productions — Entertainment pipeline volume, not a lab demo. → The Decoder
-
Model routing breaks if you only read price sheets (IBM) — On agent workloads, Sonnet cost ~half of GPT-4.1 despite higher sticker rates—cache hits and trajectory shape dominate. Routing is systems optimization, not a classifier. → HF
-
Shippy: agents need CLIs and sandboxes more than clever prompts — AllenAI maritime agent: typed CLI over raw APIs, per-user K8s isolation, evaluate the agent not just the model. → HF
🔥 GitHub Trending
-
openinterpreter/openinterpreter — Coding agent tuned for cheap/open models. Rust reimplementation of the Kimi K3 harness; Codex-compatible.
-
Nutlope/hallmark — Anti-AI-slop design skill: audit / redesign / study.
-
PrismML-Eng/Bonsai-demo — Local demo for 1-bit and ternary Bonsai. 27B vision + tools + 256k; phone-scale footprint claim.
-
Shubhamsaboo/awesome-llm-apps — 100+ agent and RAG templates you can clone and ship; one-command skills.
-
OpenCut-app/OpenCut — Open CapCut alternative rewrite: MCP server and headless batch render on the roadmap.
-
apache/ossie — Semantic metadata standard (ex-OSI) so AI agents and BI tools share one KPI definition.
▶️ YouTube
-
Claude Just Revealed AI’s Biggest Problem — Two Minute Papers | Anthropic coding-assistance and skills research, quick take.
-
Kimi K3 just beat FABLE — Matthew Berman | Short on K3 bench headlines.
-
OpenAI vs Anthropic — Berman | Pricing, subsidies, positioning.
-
AI NEWS LIVE — Berman live brief.
💬 Community
-
Kimi K3 and the pelican benchmark — Willison: SVG pelicans no longer track agent/tool skill—but still force you to actually run the model. One pelican cost ~25¢ in reasoning tokens.
-
LLM cliché highlighter — Toy app (vibe-coded with Fable 5) that flags ten LLM writing tells—“no fluff, no jargon,” etc.
-
Firefox in WebAssembly — Puter ships Gecko-in-WASM; ~$25k of Claude Opus/Fable tokens on paper, far less under Max plans.
-
AI data centers and wealth concentration — Schneier on infrastructure ownership as power concentration.
-
Verifiable AI inference — Proving a response really came from the claimed model.