grep -r "Community" /news · homesearch
grep -rli "Community" /news

Community

Full-text search across 9702 articles. Combine with topic and date filters; results sorted by relevance.

results 9702 page 141/486
14:58
2026-06-06
vettedconsumer.com
large-language-models

GGUF vs. GPTQ vs. AWQ: The Plain-English Guide to LLM Quantization

GGUF, GPTQ, and AWQ are the three dominant formats for running quantized large language models locally, each optimized for different hardware and use cases. GGUF, the format used by llama.cpp and its derivatives, offers …

14:00
2026-06-06
arxiv.org
large-language-models

Benchmarks in Leipzig

A group of 49 mathematicians compiled a dataset of 100 research-level mathematics questions with known answers during a workshop at the Max Planck Institute for Mathematics in the Sciences in Leipzig, Germany, between Ap…

12:37
2026-06-06
github.com
large-language-models

Free LLM inference handbook: 100 engineers cloned it in week 1

More than 100 engineers cloned a free open-source handbook on large language model inference within its first week of release. The guide consolidates years of production experience and research to address the unique chal…

01:20
2026-06-06
arxiv.org
machine-learning

Unlocking Non-Uniform KV Cache for Efficient Multi-Turn LLM Serving

Researchers introduced Tangram, a serving system that enables non-uniform Key-Value cache compression for multi-turn large language model inference. The system uses deterministic budget allocation, head group page cluste…

23:39
2026-06-05
arxiv.org
machine-learning

Discrete Tilt Matching

Researchers have developed Discrete Tilt Matching (DTM), a likelihood-free method for fine-tuning masked diffusion large language models using reinforcement learning. The approach recasts fine-tuning as state-level match…

23:08
2026-06-05
blog.ppb1701.com
artificial-intelligence

The Box on the Wall

NVIDIA, startup Span, and homebuilder PulteGroup partnered in May to deploy XFRA nodes — liquid-cooled, fanless data center units containing 16 NVIDIA GPUs — mounted on residential exterior walls. The units tap unused el…

22:10
2026-06-05
arxiv.org
large-language-models

A Framework for Confident Model Migration in Production Systems

Researchers have developed a Bayesian statistical framework for migrating production systems reliant on large language models when the underlying model reaches end-of-life. The approach calibrates automated evaluation me…

16:27
2026-06-05
lesswrong.com
ai-safety

Learnings from starting an AI safety research team

A new AI safety research team within Arcadia Impact in London has formed over the past four months, growing to eight members who collaborate with the UK AISI alignment team. The team, led by research lead Andrew Draganov…

15:55
2026-06-05
letsdatascience.com
ai-policy

Seattle enacts one-year moratorium on AI data centers

Seattle city council committees unanimously approved a one-year moratorium on new large-scale data centers, with a full council vote scheduled for June 9, 2026. The pause, paired with a resolution to draft regulations on…

04:00
2026-06-05
arxiv.org
natural-language-processing

Generic Triple-Latent Compression with Gated Associative Retrieval

Researchers introduced generic triple-latent sequence models that maintain a running token state and compressed pair-memory pathway to capture higher-order token interactions without benchmark-specific parsing. The tripl…

02:44
2026-06-05
arxiv.org
machine-learning

OPRD: On-Policy Representation Distillation

Researchers have introduced On-Policy Representation Distillation (OPRD), a method that aligns student and teacher model representations across selected layers during training, bypassing the language model head to elimin…

← prev page 141 / 486 next →