AI & Agent
Building intelligent systems with LLMs, autonomous agents, and multi-agent architectures. Covers Strands Agents, Amazon Bedrock, RAG, and real-world AI workflows.
13 articles
0-1. Messi's World Cup Is Over. No Models This Time — Just a Proper Goodbye
Argentina 0-1 Spain: Ferran Torres' 106th-minute goal ended Messi's sixth and final World Cup. Congratulations to Spain — and a farewell to Leo: the tears after the Egypt comeback, the letter a 15-year-old Enzo Fernández wrote him in 2016, six tournaments across 20 years, and why he was always a human being, never a god.
Inside Kimi K3's Two Architectural Pillars: KDA and AttnRes, Explained From Zero
Beginner-friendly, analogy-driven walkthrough: KDA turns attention's 'open-book exam' into 'one page of smart notes' — 75% less KV-cache memory, 6.3x faster decoding at 1M tokens. AttnRes turns the residual stream's 'running ledger' into 'a notebook with an index' — 25% better training efficiency at under 2% overhead. By the end you'll know exactly how 2.8T parameters stay standing.
Kimi K3 Is Here: 2.8 Trillion Parameters, Fully Open Source — This Time It's Different
Moonshot AI releases Kimi K3: 2.8 trillion parameters, the world's first open-source 3T-class model, 1M-token context, and native multimodality. A deep dive into the KDA and AttnRes architecture innovations, its #4 global ranking on Artificial Analysis, pricing that matches Claude Sonnet 5, and the four long-term ways an open frontier model reshapes the industry.
On the Eve of the Final, I Had fable 5 Run Another Million Simulations: Messi Kicked More Than Half the Gap Away
All six predictions from the last article hit: Spain 2-0 France, Argentina 2-1 England, and the 41.8%-probability Spain–Argentina final came true. Before the final I fed the semifinal data back to fable 5 and re-ran a million simulations: Spain's win probability dropped from 54.1% to 51.8%, Argentina's rose from 45.9% to 48.2% — Messi, leading the tournament scoring chart, took back the points the model had docked for his age. Plus the story of a photo taken 19 years ago: Messi and baby Yamal.
I'm a Messi Fan, But I Asked the Strongest Model Ever — fable 5 — to Predict the World Cup. The Answer Hurt.
The 2026 World Cup semifinals are set, so I had Claude Fable 5 build a prediction model — Elo ratings, age-curve and tournament-experience adjustments, a dedicated penalty-shootout submodel — and run 1,000,000 Monte Carlo simulations. The result is the 'Argentina Paradox': Argentina has the highest probability of reaching the final (66.5%), yet Spain is the most likely champion (37.5%). A sensitivity analysis shows Messi's form is literally the dividing line.
Slow Down in the AI Wave: A Reality Check for Employees and Bosses
At least half of today's AI anxiety is manufactured by marketing. First-hand data from PwC, MIT, METR, and DX on the 2026 AI reality: 56% of CEOs report zero ROI, 95% of enterprise GenAI pilots fail, and AI coding delivers ~10% gains, not 10x. Five grounded rules for workers, five for leaders.
When a Pop Star Ships an App: Is Hand-Written Code Becoming a Lost Art?
Chinese pop star Tiger Hu, with zero programming background, vibe-coded a fan app that reached the App Store social chart. Interns now call manual coding 'the ancient craft.' But the 90% of engineering below the waterline — operations, security, reliability — makes experienced engineers more scarce, not less. A confession from a hand-coder of fifteen years.
Anthropic Goes Big: Claude Fable 5 and Mythos 5 Are Here
On June 9, 2026, Anthropic shipped Claude Fable 5 and Mythos 5 simultaneously: the same underlying model, differing only in guardrails. Fable 5 is SOTA on nearly every benchmark, with the lead growing on longer tasks (Stripe migrated a 50-million-line codebase in one day). Priced at $10/$50 — less than half of Mythos Preview — and already live on Amazon Bedrock. Plus: why did the hidden Mythos suddenly come out of the vault?
How Do I Explain "Agent" to My Wife?
From 'the Lark bot can't do math' to 'raising your own AI lobster' — an AI-agent explainer for normal humans. Thirteen burning questions covering LLMs, tokens, Tools, MCP, RAG, Skills, Memory, Multi-Agent systems, and 2026 model prices. The AI isn't dumb — it just hasn't been raised properly yet.
Amazon AI Strategy 2026: Why the Biggest Player Is the Least Visible
Custom chips, global infrastructure, massive investment — yet Amazon is invisible in the AI race. Here's what's really going on.
WeChat x OpenClaw: Platform Strategy in the AI Agent Era
WeChat's native OpenClaw integration signals a major shift. Why the world's largest messaging app opening up to AI agents matters.
How to Build AI Agents for Ad Creative Generation
Automate ad copywriting, image, and video production with Strands Agents and Amazon Bedrock. A practical, code-first guide.
How to Build a RAG System with LangChain and Elasticsearch
A hands-on guide to building Retrieval Augmented Generation — from vector embeddings to context-enhanced LLM answers.