What breaks in production AI workflows? A developer building StateGuard reproduced over 30 real AI runtime failures from GitHub issues and found most were runtime contract mismatches between providers, tools, and application code, not model failures. The tool aims to address these failures, and the creator is seeking feedback from builders on runtime failures they face. Over the past few weeks, we've reproduced 30+ real AI runtime failures from GitHub issues instead of just reading about them. Most weren't model failures - they were runtime contract mismatches between providers, tools, and application code. That led us to build StateGuard. We'd genuinely love feedback from builders: what runtime failures are you facing, and is this something you'd use? Incase y'all wanna connect: GitHub: https://github.com/dood1ebyte/stateguard https://github.com/dood1ebyte/stateguard LinkedIn: https://www.linkedin.com/in/adivaishnav https://www.linkedin.com/in/adivaishnav