Are the Financial Reasoning from LLMs Credible? A Real World Test over Long-Horizon Statements
A new benchmark called FinIndices, introduced by researchers in a preprint on arXiv (2607.28661v1), tests large language models on financial reasoning over long-horizon statements up to 32K tokens and…