21:22
2026-07-31
lesswrong.com
ai-safety
The temporal lockbox: a hardened observatory for AI misalignment
A new proposal introduces the 'temporal lockbox,' a hardened observatory for AI misalignment that scores AI agents' weather forecasts against future measurements that cannot be influenced by the forecβ¦