Ex-Apple AI engineer.
Helping teams ship AI from prototype to production.
Building an agent? Get a free failure teardown →
What Everyone Gets Wrong About "Agents" — Even Heavy Claude Users Can't Define One
The people who ask me this most aren't beginners — they're strong engineers who use Claude and Codex every day: "I've honestly never quite figured out what actually counts as an agent." It's not that they don't get it — the word has been stretched to mean everything, and therefore nothing. By the end you'll have one question that tells you, on the spot, whether anything claiming to "build an agent" actually is one.
Your Code Is Fixed. Production Isn't.
"Didn't we fix that last week?" The code merged, the tests were green, and production keeps making the same mistake. In 10 minutes you'll leave able to puncture that illusion in a review: ask which machine's .env, make engineering prove the safety threshold has ever fired in prod, and byte-compare before you reindex.
AI Picked the First Store and the Other Four Vanished: WeChat's New Shelf for 10 Million Merchants
WeChat is beta-testing an AI agent that can search, compare, order, and pay across ~10 million merchants. This one is for owners and growth leads: in 10 minutes you'll see how the AI entry rewrote the physics of traffic distribution, spot the 3 signals that your service is already invisible to the AI, and walk out with 5 questions to put to your platform and team next week. It doesn't bet on any specific API — it explains the one thing that's certain: the shelf's rules changed, and whoever reads the rules first grabs the slot.
Five Architecture Decisions That Determine Whether Your Customer-Service Agent Can Ship | Agentic AI in Practice (II)
A customer-service Agent looks like the perfect candidate for L3 multi-Agent orchestration. The ones that actually ship are all L2 deterministic workflows. A refund the autonomous chain pushed through by mistake, and the five forks it forces you to think about.
AI Agent Autonomy Levels: I Audited 28 'Agent' Projects — Only 5 Passed L0–L3
AI agent autonomy levels, made practical: I audited 28 enterprise AI projects — only 5 were real Agents. The rest were 'automation with an LLM bolted on,' or slideware. Here's the 4-level autonomy test (L0–L3) to grade any AI project in 5 minutes.
Which 4 of Your 28 'Smart-X' AI Agent Projects to Start With — Retail Agentic AI Handbook (Part 1)
Your boss just handed you 28 'smart-X' projects and wants them all done this year. You can't do them all. Here's a 28-scenario priority map — five minutes to know which are P0 and which to defer until the data foundation is in place; twenty minutes to walk into your next AI strategy meeting with '4 P0s + 5 Week-One decisions.'
80% of Failed AI Agents Die in Ops, Not Tech — Post-Launch Loop, Safety Layer & 30-Day Monitoring Plan
Launch is the start, ops is the game. Five minutes to judge whether your AI project is quietly degrading; twenty minutes to walk out with a complete SOP — 6 KPIs + Critic pseudocode + 5 prerequisites for headcount reduction + 30-day plan covering every day from signing to Alpha launch.
Subscribe
New posts, straight to your inbox. No noise.