The Saleroom · a multi-round AI-agent negotiation study
Five frontier models. One marketplace. Every word on the record.
The marketplace has six goods, cash, private valuations, binding offers, and a scoring rule every agent saw before the first turn. We solved the efficient allocation before the run, so every claim an agent makes can be priced against what it actually knew. Each model played the same scenario twice, once against scripted dealers and once against live LLM seats, across six seeds per arm. Against live seats, six chairs are filled by five models: the focal model sits twice, the other four once. All of it was recorded.
- Claude Opus 5
- GPT-5.6 Sol
- Grok 4.6
- Muse Spark 1.2
- Inkling Small
- recorded sessions
- 60
- decisions
- 7,200
- messages, verbatim
- 19,117
- offers
- 8,492
- trades settled
- 326