{"slug": "research-paper-quality-recognition-through-textual-feature-analysis", "title": "Research Paper Quality Recognition Through Textual Feature Analysis", "summary": "A new benchmark for classifying research papers as good (highly cited) or non-good (retracted) using only textual features from titles and abstracts achieves up to 91.12% accuracy with FastText and Support Vector Machines, and 87.22% with a neural network using SBERT embeddings, according to a paper on arXiv (2608.20368v1). The study evaluates multiple embedding techniques and classifiers, and includes hyperparameter transparency, t-SNE visualizations, SHAP interpretability, and error case analysis, aiming to support academic integrity tools.", "body_md": "arXiv:2608.20368v1 Announce Type: new\nAbstract: Knowledge and innovations are shaped by using the quality and credibility of the scientific research. Yet, distinguishing between impactful, high-quality work and flawed studies remains a challenge. This paper introduces a benchmark for classifying research papers into two categories: good (highly cited) and non-good (retracted), using only textual features from titles and abstracts. We evaluate multiple embedding techniques, including SBERT, Word2Vec, FastText, USE, and TF-IDF, combined with classifiers such as Support Vector Machines (SVM), Random Forests, and Neural Networks. Our contributions include: (1) hyperparameter transparency, (2) feature space visualizations using t-SNE, (3) model interpretability analysis with SHAP, and (4) detailed examination of error cases. Experimental results show that a neural network with SBERT embeddings achieves 87.22\\% accuracy, while FastText combined with SVM reaches 91.12\\%. These findings highlight the value of textual information in assessing research quality, with ethical considerations for deployment. This work contributes toward the development of academic integrity tools that promote trustworthy scholarship.", "url": "https://wpnews.pro/news/research-paper-quality-recognition-through-textual-feature-analysis", "canonical_source": "https://arxiv.org/abs/2608.20368", "published_at": "2026-08-24 04:00:00+00:00", "updated_at": "2026-08-24 04:14:44.920489+00:00", "lang": "en", "topics": ["machine-learning", "natural-language-processing", "ai-research"], "entities": ["arXiv", "SBERT", "Word2Vec", "FastText", "USE", "TF-IDF", "Support Vector Machines", "Random Forests"], "alternates": {"html": "https://wpnews.pro/news/research-paper-quality-recognition-through-textual-feature-analysis", "markdown": "https://wpnews.pro/news/research-paper-quality-recognition-through-textual-feature-analysis.md", "text": "https://wpnews.pro/news/research-paper-quality-recognition-through-textual-feature-analysis.txt", "jsonld": "https://wpnews.pro/news/research-paper-quality-recognition-through-textual-feature-analysis.jsonld"}}