cd /news/natural-language-processing/lexlattice-multilingual-extractive-s… · home topics natural-language-processing article
[ARTICLE · art-138830] src=arxiv.org ↗ pub= topic=natural-language-processing verified=true sentiment=↑ positive

LexLattice: Multilingual Extractive Summarization via Neural Cellular Automata on Document Hierarchies

LexLattice, an extractive summarizer that models a legal act's hierarchy as a two-dimensional semantic lattice and consolidates over it with a masked 2D neural cellular automata, attained state-of-the-art ROUGE across all 24 languages of EUR-Lex-Sum in both multilingual and cross-lingual settings, according to the arXiv paper 2609.27032v1. The system surpasses instruction-tuned baselines with billions of parameters while concentrating all trainable capacity in a 1.8M-parameter consolidator over a frozen multilingual encoder. A consolidator trained only on high-resource languages transferred to unseen languages with 0.99 near-lossless retention, which the authors say indicates the model operates on language-agnostic semantic geometry rather than surface form.

by read1 min views1 publishedSep 24, 2026

arXiv:2609.27032v1 Announce Type: new Abstract: Faithfulness is a central concern in legal text summarization, which motivates extractive approaches that select verbatim content traceable to its source. Such methods typically rank paragraphs or other structural units in isolation, yet give little attention to consolidating evidence that is distributed across, and shares salience between, distant parts of a document. We introduce LexLattice, an extractive summarizer that reifies a legal act's hierarchy as a two-dimensional semantic lattice and consolidates over it with a masked 2D neural cellular automata before selection. LexLattice attains state-of-the-art ROUGE across all 24 languages of EUR-Lex-Sum in both multilingual and cross-lingual settings, surpassing instruction-tuned baselines with billions of parameters, despite concentrating all trainable capacity in a 1.8M parameter consolidator over a frozen multilingual encoder. A consolidator trained only on high-resource languages further transfers to unseen languages with near-lossless retention (0.99), indicating that the model operates on language-agnostic semantic geometry rather than surface form. Our results position explicit consolidation over document structure as a compact and traceable alternative to scale for multilingual legal summarization.

── more in #natural-language-processing 4 stories · sorted by recency
── more on @lexlattice 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/lexlattice-multiling…] indexed:0 read:1min 2026-09-24 ·