AI models tested in the Elm Wealth experiment predicted market direction 60% of the time — better than humans — but made the same catastrophic mistake: they didn't calibrate bet sizes to their confidence. Even our supposed future overlords took too much risk when uncertain.