AI & Agent

Building intelligent systems with LLMs, autonomous agents, and multi-agent architectures. Covers Strands Agents, Amazon Bedrock, RAG, and real-world AI workflows.

13 articles

0-1. Messi's World Cup Is Over. No Models This Time — Just a Proper Goodbye
AI

0-1. Messi's World Cup Is Over. No Models This Time — Just a Proper Goodbye

Argentina 0-1 Spain: Ferran Torres' 106th-minute goal ended Messi's sixth and final World Cup. Congratulations to Spain — and a farewell to Leo: the tears after the Egypt comeback, the letter a 15-year-old Enzo Fernández wrote him in 2016, six tournaments across 20 years, and why he was always a human being, never a god.

9 min
Inside Kimi K3's Two Architectural Pillars: KDA and AttnRes, Explained From Zero
AI

Inside Kimi K3's Two Architectural Pillars: KDA and AttnRes, Explained From Zero

Beginner-friendly, analogy-driven walkthrough: KDA turns attention's 'open-book exam' into 'one page of smart notes' — 75% less KV-cache memory, 6.3x faster decoding at 1M tokens. AttnRes turns the residual stream's 'running ledger' into 'a notebook with an index' — 25% better training efficiency at under 2% overhead. By the end you'll know exactly how 2.8T parameters stay standing.

15 min
Kimi K3 Is Here: 2.8 Trillion Parameters, Fully Open Source — This Time It's Different
AI

Kimi K3 Is Here: 2.8 Trillion Parameters, Fully Open Source — This Time It's Different

Moonshot AI releases Kimi K3: 2.8 trillion parameters, the world's first open-source 3T-class model, 1M-token context, and native multimodality. A deep dive into the KDA and AttnRes architecture innovations, its #4 global ranking on Artificial Analysis, pricing that matches Claude Sonnet 5, and the four long-term ways an open frontier model reshapes the industry.

14 min
On the Eve of the Final, I Had fable 5 Run Another Million Simulations: Messi Kicked More Than Half the Gap Away
AI

On the Eve of the Final, I Had fable 5 Run Another Million Simulations: Messi Kicked More Than Half the Gap Away

All six predictions from the last article hit: Spain 2-0 France, Argentina 2-1 England, and the 41.8%-probability Spain–Argentina final came true. Before the final I fed the semifinal data back to fable 5 and re-ran a million simulations: Spain's win probability dropped from 54.1% to 51.8%, Argentina's rose from 45.9% to 48.2% — Messi, leading the tournament scoring chart, took back the points the model had docked for his age. Plus the story of a photo taken 19 years ago: Messi and baby Yamal.

10 min
I'm a Messi Fan, But I Asked the Strongest Model Ever — fable 5 — to Predict the World Cup. The Answer Hurt.
AI

I'm a Messi Fan, But I Asked the Strongest Model Ever — fable 5 — to Predict the World Cup. The Answer Hurt.

The 2026 World Cup semifinals are set, so I had Claude Fable 5 build a prediction model — Elo ratings, age-curve and tournament-experience adjustments, a dedicated penalty-shootout submodel — and run 1,000,000 Monte Carlo simulations. The result is the 'Argentina Paradox': Argentina has the highest probability of reaching the final (66.5%), yet Spain is the most likely champion (37.5%). A sensitivity analysis shows Messi's form is literally the dividing line.

12 min
Slow Down in the AI Wave: A Reality Check for Employees and Bosses
AI

Slow Down in the AI Wave: A Reality Check for Employees and Bosses

At least half of today's AI anxiety is manufactured by marketing. First-hand data from PwC, MIT, METR, and DX on the 2026 AI reality: 56% of CEOs report zero ROI, 95% of enterprise GenAI pilots fail, and AI coding delivers ~10% gains, not 10x. Five grounded rules for workers, five for leaders.

16 min
When a Pop Star Ships an App: Is Hand-Written Code Becoming a Lost Art?
AI

When a Pop Star Ships an App: Is Hand-Written Code Becoming a Lost Art?

Chinese pop star Tiger Hu, with zero programming background, vibe-coded a fan app that reached the App Store social chart. Interns now call manual coding 'the ancient craft.' But the 90% of engineering below the waterline — operations, security, reliability — makes experienced engineers more scarce, not less. A confession from a hand-coder of fifteen years.

9 min
Anthropic Goes Big: Claude Fable 5 and Mythos 5 Are Here
AI

Anthropic Goes Big: Claude Fable 5 and Mythos 5 Are Here

On June 9, 2026, Anthropic shipped Claude Fable 5 and Mythos 5 simultaneously: the same underlying model, differing only in guardrails. Fable 5 is SOTA on nearly every benchmark, with the lead growing on longer tasks (Stripe migrated a 50-million-line codebase in one day). Priced at $10/$50 — less than half of Mythos Preview — and already live on Amazon Bedrock. Plus: why did the hidden Mythos suddenly come out of the vault?

10 min
How Do I Explain "Agent" to My Wife?
AI

How Do I Explain "Agent" to My Wife?

From 'the Lark bot can't do math' to 'raising your own AI lobster' — an AI-agent explainer for normal humans. Thirteen burning questions covering LLMs, tokens, Tools, MCP, RAG, Skills, Memory, Multi-Agent systems, and 2026 model prices. The AI isn't dumb — it just hasn't been raised properly yet.

18 min
Amazon AI Strategy 2026: Why the Biggest Player Is the Least Visible
AI

Amazon AI Strategy 2026: Why the Biggest Player Is the Least Visible

Custom chips, global infrastructure, massive investment — yet Amazon is invisible in the AI race. Here's what's really going on.

WeChat x OpenClaw: Platform Strategy in the AI Agent Era
AI

WeChat x OpenClaw: Platform Strategy in the AI Agent Era

WeChat's native OpenClaw integration signals a major shift. Why the world's largest messaging app opening up to AI agents matters.

8 min
AI

How to Build AI Agents for Ad Creative Generation

Automate ad copywriting, image, and video production with Strands Agents and Amazon Bedrock. A practical, code-first guide.

15 min
AI

How to Build a RAG System with LangChain and Elasticsearch

A hands-on guide to building Retrieval Augmented Generation — from vector embeddings to context-enhanced LLM answers.

18 min