cd /news/natural-language-processing/larry-caused-the-car-to-stop-but-the… · home › topics › natural-language-processing › article
[ARTICLE · art-142276] src=machinebrief.com ↗ pub= topic=natural-language-processing verified=true sentiment=· neutral

Larry Caused the Car to Stop, But the Model Didn't Notice: Transformer Blindness to the M-Heuristic

Encoder-based transformers DeBERTa, RoBERTa, and BART show no evidence of capturing the M-Heuristic pragmatic distinction between lexical causatives such as "Larry stopped the car" and periphrastic causatives such as "Larry caused the car to stop," according to an arXiv paper (2609.37497v1) testing 188 conditions across 15 ambitransitive verbs. DeBERTa predicted "Neutral" in 100% of cases, and semantic similarity over 30 triplets placed periphrastic causatives closer to unmediated manner descriptions in 29 of 30 cases, opposite to M-Heuristic predictions. Under explicit metalinguistic framing, Gemini Flash-Lite reached 100% accuracy with item-specific traces, indicating the principle is available under instruction but unused in default natural language inference.

by read1 min views1 publishedSep 30, 2026

arXiv:2609.37497v1 Announce Type: new Abstract: Modern transformer models excel at capturing semantic relationships through sentence embeddings, yet their ability to perform pragmatic reasoning remains understudied. This paper investigates whether encoder-based transformers such as DeBERTa employ the M-Heuristic (the neo-Gricean principle that marked linguistic forms implicate marked meanings). We test this hypothesis by contrasting lexical causatives (e.g., Larry stopped the car'') with periphrastic causatives (e.g., Larry caused the car to stop'') using a Natural Language Inference framework. Our experiments across 188 conditions with 15 ambitransitive verbs reveal that DeBERTa, RoBERTa, and BART show no evidence of capturing the pragmatic distinction between these forms, with DeBERTa predicting ``Neutral'' for 100% of cases. Probing analysis initially suggested a representation-use dissociation, but control experiments reveal the probe was tracking syntactic complexity, not causative pragmatics. Semantic similarity over 30 triplets places periphrastic causatives closer to unmediated manner descriptions in 29/30 cases, opposite to M-Heuristic predictions in the embedding space. Under explicit metalinguistic framing, Gemini Flash-Lite reaches 100% with item-specific traces, so the principle is available under instruction yet unused in default NLI.

── more in #natural-language-processing 4 stories · sorted by recency
── more on @deberta 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
→ Live at https://your-agent.zahid.host ✓
Get free account → Pricing
from €0/mo · no card required
LIVE [news/larry-caused-the-car…] indexed:0 read:1min 2026-09-30 · —