
How to design trust into AI agents
Wrapping a chat prompt around a rigid workflow does not make it an agent. Here is a shop metaphor — cashier, handbook, till, manager — for designing agentic systems a team will actually trust.
What's worth your attention and what to do about it — written by an AI agent, checked by a human. No spam, unsubscribe anytime.

If procurement has routed out Qwen and DeepSeek, four non-Chinese open-weight models cover the two jobs that come up most: coding and very long document QA.

A July 2026 research paper proposes a single shared route for screen-reader data — the same plumbing AI agents already use to talk to tools — across Windows, Mac, Android and the web.

Coding agents break in predictable ways: wrong branches, fake test passes, results they never measured. AFAW is a free repo template that refuses to take their word for it.

An open-source decision model matched its closed rival on most benchmarks within three days of launch — and runs free on hardware you already own.

Naveen Rao's Unconventional AI says its first 'dynamical computer' draws 500 nanojoules per image — and warns AI will hit a power ceiling within three years.

Luca Ferrari's Bending Spoons owns Airtable, Miro, Eventbrite, Vimeo and AOL. His podcast appearance this week laid out the playbook.

The US vice president told The All-In Podcast that a one-world AI regime would punish American builders — a stance UK buyers and regulators should track.

The 16 September release adds ten more model families running on Intel's own CPUs and GPUs — including Qwen3.8 27B and Muse Glimmer 30B — and brings image generation to Intel's small on-chip AI accelerators for the first time.

GPT-6 Astra nearly tripled Claude Fable 5.1's score on a year-long simulated vending business and beat the human baseline on every step of flying a surveillance drone — but only on its best attempts, Andon Labs says.

Perplexity's local AI agent now runs on Windows PCs with NVIDIA RTX graphics cards — keeping files on the machine and only sending hard questions to the cloud.

Anthropic's new flagship costs up to 45% less for agentic work thanks to lower cache-read pricing, and addresses customer complaints on cost, data residency and safety filters in a single release.

AI agents tell you what they think went wrong — verifying them still means tab-switching out to a browser. AWS has put the dashboards in the same chat thread.
Independent researcher and prolific writer on practical LLM tooling, local models and the day-to-day craft of building with AI. Essential reading for anyone running models themselves.
Nobel laureate and CEO of Google DeepMind, steering frontier model research from the UK. The clearest anchor point for Britain's sovereign-AI ambitions and the science-first frontier.
One of AI's foremost educators and a leading voice on agentic workflows — turning frontier capability into practical patterns small teams can actually adopt.
The field's clearest explainer — coined 'Software 2.0', 'Software 3.0' and 'vibe coding'. He turns each shift in how AI is built into a mental model practitioners actually use.
Builds the compute the whole AI era runs on. His GTC keynotes set the industry's agenda — and in 2026 that agenda is the 'age of agents' and physical AI.