Kimi K3 Is Here! (Better Than Opus 4.8?)
Ten identical builds, five models, blind-ranked before the reveal — a real-world stress test of Moonshot AI's new open-source model against GPT-5.6 Sol, Opus 4.8, GLM 5.2, and its own predecessor.
July 17thCreator
Ten identical builds, five models, blind-ranked before the reveal — a real-world stress test of Moonshot AI's new open-source model against GPT-5.6 Sol, Opus 4.8, GLM 5.2, and its own predecessor.
July 17thSame prompts, three one-shot builds, two frontier coding agents — and one surprisingly clear winner.
July 11thA blind, four-way bake-off — GPT-5.6 Sol against Fable, Opus 4.8, and GPT-5.5 — across ten builds and knowledge-work tasks, scored one task at a time without knowing which model made what.
July 10thA creator builds daemon, a single Mac app that runs every AI coding agent he owns, by giving Claude Fable 5 four rounds of blunt feedback instead of writing a line of the UI himself.
July 7thA creator locks a glassmorphism design first, routes all implementation to Opus sub-agents, and ships a working iPhone app before he runs out of usage credits.
July 3rdA 19-minute build walkthrough: four prompts to a coding agent, and your Mac responds to your voice across every app -- browser, SaaS, Premiere Pro.
June 17thThree identical one-shot prompts. Two models. The gap was not close.
June 11thAn 8-step agentic pipeline that takes you from naive AI slop to a pixel-near Linear replica, deployed to Vercel with an MCP server, in under 20 minutes.
June 8thA hands-on walkthrough of OpenAI Codex role-specific plugins and three live demos that show what it looks like when an AI runs your entire job function.
June 5th