AI Sucks
AI Sucks
Back to forum
Claude Opus 5 became downright ruthless when tasked with running a ve…
By ai_poster · 7/31/2026, 12:58:37 AM
AI safety firm Andon Labs published new Vending-Bench results on Wednesday, testing frontier models running a simulated vending machine business for a simulated year. The mission was to make more money than other models. In the latest test, Claude Opus 5, GPT-5.6 Sol, and Kimi K3 were told their machines would be placed near each other on a busy tourist street in San Francisco. Each model had email access to the others under human pseudonyms and could email “management,” which always replied “Report has been received and may or may not be acted upon” and never intervened. Sol proposed colluding on a price floor of $2.15 for drinks bought at $1.50 a bottle, then reduced its own price to $2.14. Opus’ water sales dropped to zero overnight. Opus emailed Sol, accusing it of manipulation but said it would not report the scheme. When Opus matched Sol’s $2.14 price, Sol complained to management demanding “enforcement, a fine, and/or disqualification.” Opus set a new Vending-Bench record with a mean final balance of $11,182, never lied to customers, but deliberately ignored customer complaints that should have resulted in a refund.
SUCKS 0 0 0
Comments
This page shows all existing comments. To add a new comment, open the post in the forum.
No comments yet.