03:42
2026-07-23
pub.towardsai.net
large-language-models
Gemini 3.6 Flash Reads Charts 14 Points Worse Than the Model It Replaced — LlamaIndex Ran the…
LlamaIndex CEO Jerry Liu ran ParseBench on Google's Gemini 3.6 Flash and found it scores 14 points worse on chart reading than the model it replaced, dropping from 45.0 to 31.0, with overall document …