cd /news/natural-language-processing/do-general-nlp-embeddings-capture-on… · home topics natural-language-processing article
[ARTICLE · art-118571] src=arxiv.org ↗ pub= topic=natural-language-processing verified=true sentiment=· neutral

Do General NLP Embeddings Capture Ontological Reasoning?

A new study introducing the AVA framework finds that general-purpose NLP embedding models struggle to capture ontological reasoning, with the best model achieving only 0.739 triplet accuracy and 0.135 hard negative accuracy across 171,007 contrastive triplets from 163 ontologies. The authors, who posted the paper on arXiv (2609.00177v1), show that fine-tuning improves discrimination but transfers poorly to downstream Semantic Web tasks, challenging assumptions about NLP benchmark performance.

read1 min views1 publishedSep 2, 2026

arXiv:2609.00177v1 Announce Type: new Abstract: General-purpose NLP embedding models perform well on linguistic tasks, but their ability to capture symbolic ontological structure remains unclear. We introduce AVA, a systematic framework for evaluating whether embeddings distinguish logic-sensitive relational semantics in ontologies and knowledge graphs. AVA comprises 171,007 contrastive triplets derived from 163 heterogeneous ontologies using hierarchy inversion, relation substitution, and disjointness injection. Each triplet contains an ontology statement, a semantically equivalent paraphrase, and a logic-sensitive hard negative with contradictory relational meaning. We evaluate more than 25 state-of-the-art embedding models and find substantial limitations: the best model achieves only 0.739 triplet accuracy, while hard negative accuracy falls to 0.135. Fine-tuning improves discrimination by a large margin but transfers poorly to downstream Semantic Web tasks, including taxonomy discovery and ontology alignment. Further analysis suggests that improvements stem partly from perturbation-specific pattern recognition rather than robust ontological understanding. These findings reveal a persistent gap between linguistic representation learning and ontology-level discrimination, challenging the assumption that strong NLP benchmark performance translates to Semantic Web competence.

── more in #natural-language-processing 4 stories · sorted by recency
── more on @ava 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/do-general-nlp-embed…] indexed:0 read:1min 2026-09-02 ·