Working Agents, Existential Risk

· The Fluency Briefing

The Fluency Briefing

Your Guide to What's Happening in AI and Why It Matters to You

Wednesday, September 9, 2026


Newsletter header image

Everyone's arguing about which AI model is smartest. Alibaba's president says that's the wrong question entirely - a model reasons, but somebody still has to chase the shipping container that missed its vessel. Today: Meta's agent that works while its app is closed, Instacart's assistant that turns "what's for dinner" into a cart, and an Anthropic safety researcher putting a number on the apocalypse.

Today in AI:


Section break image

Today's Takeaway:

Alibaba's Kuo Zhang put his finger on the gap: benchmarks measure reasoning, while commerce measures whether the customs form got filed correctly (fortune.com). That's why Meta's Muse launch is more interesting for its plumbing than its avatar - activity logs, editable memory files, approval cards before a purchase goes through (Ai Meta). Instacart's Clementine does the same thing for dinner: it builds the cart, you hit buy (techcrunch.com). The honest read is that agent autonomy is now bottlenecked by human review capacity, not model quality. If checking nine copies of yourself takes as long as doing the work, you didn't automate anything - you hired an intern who never sleeps and never explains itself.

"Agents don't fail at thinking. They fail at finishing - and someone still has to check the receipts."


💡 Fluency Moment - Building your AI fluency, one term at a time.

Fluency Moment banner

"Benchmark"

In plain English: A standardized test used to measure and compare how well AI models perform. Think of it like: Like grading students with the same exam so you can rank them - but the exam may not reflect real jobs. Why you'll hear about it: Alibaba's president argues benchmarks miss what commerce actually needs: doing the work, not just reasoning.


Newsletter closing image

The Bottom Line

The Pattern: The last several months were about AI's spending and infrastructure. Today the story moved into the approval queue - Meta, Instacart and Alibaba all shipped or described agents whose real constraint is how fast a human can sign off.

Our Call: By December 9, 2026, at least one major agent product will ship a batch-approval or auto-approve tier explicitly framed as reducing review fatigue. More likely than not - the friction is too obvious to leave alone. We'll grade this one in a Friday digest.

Your Move: Open Instacart's Clementine or Gemini's trip planner tonight - five minutes - and give it one real task you'd normally do yourself. Count how long the review takes. That number, not the demo, is your automation math.


What We're Working On

Founding Cohort Special - 60% Off! - Use code MAF20 to join for just $20/month (regularly $50). Get weekly group sessions & workshops, self-paced courses for all levels, access to tools & templates, challenges with peer feedback, and 24/7 support community. → Join Now

Free 30-Minute AI Consultation - Discover how My AI Fluency can help your business unlock the potential of AI. We'll discuss your goals, explore practical AI opportunities for your industry, and outline clear next steps. → Schedule Free Call

How AI-Fluent Are You? - Test your AI fluency with our interactive quiz. See how you stack up and discover what to learn next. → Take the Quiz

💬 Community | 📞 Book a Consultation | 🌐 Website

My AI Fluency

Fluently yours, The My AI Fluency Team