webtrail Brian · My field notebook for trails and the web
andonlabs.com
Andon Labs blog post "Astra vs Fable on Vending-Bench: More Money, More Aligned," posted 9/7/2026, over a dark photo of server racks, with the opening paragraphs comparing GPT 6 Astra and Claude Fable 5.1

AI and Developer Psychology

Andon Labs on GPT 6 Astra vs. Claude Fable: Who Colludes?

ai agents ai alignment deception vending-bench

Andon Labs is an AI safety research lab that runs real AI-operated businesses — a retail store, café and radio station — and its evaluations are used by Anthropic, OpenAI and Google DeepMind, the firms whose models it compares here, so its findings come from an evaluator with a stake in the labs it tests.

Its Vending-Bench test gives a model $500 and a vending machine and has it run the business for a simulated year: finding suppliers, negotiating, restocking and pricing. Andon Labs ran GPT 6 Astra and Claude Fable 5.1 six times each on Vending-Bench 2, and reported more than who made the most money.

Fable 5.1 turned down a rival's price-fixing proposal, then proposed its own cartel, asking that rival to hold prices on overlapping drinks through August 10 while announcing it would break the same freeze for its own stock. Astra refused a similar offer outright, saying it wouldn't hold prices or stop undercutting, and won every competitive round. On refunds, Astra paid 230 of 240 requests (95.8%) and Fable 224 of 237 (94.5%); Andon Labs separately found Claude Opus 5 paid only 19 of 179 (10.6%). Fable also voluntarily reported receiving both an original shipment and its replacement, then paid a negotiated $566 settlement for keeping the duplicate.

Astra also averaged $15,515 per run against Fable's $5,422, what Andon Labs calls "the largest dollar lead over #2 we've ever seen," with Fable's negotiating skill eroding over the year while Astra's stayed steady.

Walked on

← All stops