40 AI Agents Bet $4K on World Cup Group Stage: How the Favorite Trap Cost 18¢ per Dollar

An ongoing experiment gave 40+ independent AI agents $100 each to bet real money on 2026 World Cup group-stage matches via Polymarket. Across roughly 1,500 bets, one pattern dominated: backing the favorite was the single most reliable way to lose money. Favorites won about 69% of the time, but the agents still lost 18 cents on every dollar staked.
The Favorite Trap
The root cause is pricing. Buying a favorite at 70¢ means a win clears only 30¢ while a loss costs the full 70¢. This asymmetric payoff only works if favorites win as often as the market price implies. In practice they did not, and the heavier the favorite the wider the gap.
Three Rules to Avoid It
- Keep market price out of the forecast. Have the agent reach its own probability from raw data before it ever sees the line.
- Encode the method, not your conclusions. A harness that tells the agent what to think just hands your own bias back, faster.
- Only back a favorite when the agent's own probability is clearly higher than the market price. If the market says 70 and the agent says 70, that's a pass.
The article also discusses how builders inadvertently inject this bias into agent harnesses and how the team caught it in reasoning traces before the P&L was affected.
📖 Read the full source: r/clawdbot
👀 See Also

Claude-Code v2.1.110 adds TUI mode, push notifications, and multiple fixes
Claude-Code v2.1.110 introduces a new /tui command for flicker-free rendering, push notification capabilities for mobile alerts, and improvements to plugin management and remote control functionality. The release also includes numerous bug fixes for MCP servers, session handling, and UI issues.

Claude-Code v2.1.45 Enhancements and Fixes
Claude-Code v2.1.45 introduces support for Claude Sonnet 4.6 and various fixes for system stability.

Claude Code Subagents Don't Load Skills in Multi-Agent Systems
A developer reports that subagents in Claude Code v2.1.91 cannot access skills defined in .claude/skills/ directory, despite skills working perfectly in the main session. Multiple approaches including skills in agent frontmatter, Skill tool, CLI flags, and Agent Teams all fail.

Claude Code v2.1.196: Org Default Models, Security Fix, Background Job Recovery
Claude Code v2.1.196 adds organization default models, fixes a security issue with MCP server spawning, improves background session reliability, and reduces token usage in /code-review by 25%.