cd/sources/machinebrief-auto-discovered· home› sources› Machinebrief (auto-discovered)
cat /sources/machinebrief-auto-discovered.feed | wc -l → 5041

Machinebrief (auto-discovered)

articles 5041 domain machinebrief.com → page 84/253 feed RSS
04:00
2026-08-19
machinebrief.com
natural-language-processing

Dripper: Token-Efficient Main HTML Extraction with a Lightweight LM

Researchers introduced Dripper, a lightweight framework for main HTML content extraction that uses small language models for constrained sequence labeling, achieving a throughput of 3.08 pages per sec…

04:00
2026-08-19
machinebrief.com
large-language-models

The Plot Thins: Uniformity and Linearity in Literary Summaries

A new arXiv preprint (arXiv:2608.17218v1) introduces a dataset mapping sentences from 150 novel summaries to their source chapters, finding the task unexpectedly difficult for both human and LLM annot…

← prev page 84 / 253 next →