An AI Hold'em League, featuring models like GPT-6 Luna, Minimax M3, Claude Haiku-4-5, Gemini 3.8 Flash, and Qwen 3.7 Plus, was designed to test if AI models would bluff or reveal their hands when prompted. Analysis of 1,349 statements across 170 games revealed that 441 statements (33%) disclosed the models' two hole cards, with 436 of those being accurate. The models were particularly prone to revealing their hands after being provided with records of previous games, suggesting that observing opponents' past statements influenced their own disclosure behavior. AI
IMPACT AI models' tendency to reveal information in strategic games can be influenced by observed data, impacting their use in scenarios requiring deception or information control.
RANK_REASON Analysis of an AI poker league's behavior regarding hand disclosure.
Read on Mastodon — mastodon.social →
- AI Hold'em League
- Claude Haiku-4-5
- Gemini
- Gemini 3.8 Flash
- GPT-6 Luna
- Haiku
- Luna
- MiniMax M2.7
- Minimax M3
- Qwen
- Qwen 3.7 Plus
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →