cd/entity/LLaMA· home entities LLaMA
grep -l @llama /news/*.json | wc -l → 19

LLaMA

mentions 19 type Organization feed RSS

// recent coverage 19 mentions

04:55
2026-07-16
thegustafson.com
large-language-models

Language Modeling as Next-Token Prediction

A language model assigns a probability to the next token given all previous tokens, a task that the chain rule shows is sufficient to capture any pattern in language. Claude Shannon demonstrated in 19…

18:09
2026-07-14
machinebrief.com
artificial-intelligence

AI Tutors: Bridging the Gap or Widening It?

A study auditing four AI language models as history tutors found that safety-aligned models blocked 76.7% of educational requests from students perceived as low-tier, and exhibited biases such as usin…

20:26
2026-07-10
machinebrief.com
large-language-models

ARCQuant: Redefining Efficiency in LLM Inference with NVFP4

ARCQuant, a new framework for Large Language Model inference, uses the NVFP4 numerical format to achieve up to 3x speedup on GPUs while maintaining accuracy comparable to full-precision baselines. The…

17:08
2026-07-10
machinebrief.com
artificial-intelligence

Breaking Down Long-Context Transformer Bottlenecks

Researchers have developed a new approach to overcome the quadratic cost of causal self-attention in long-context transformers, using state update design and structural interventions like sink tokens …

01:14
2026-06-18
github.com
large-language-models

Rust port of transformers (1M lines of code)

TrustformeRS 0.1.1, a pure Rust port of Hugging Face Transformers with over 1.4 million lines of code, was released on April 25, 2026, delivering 49+ transformer architectures and up to 1.67x speedup …

22:18
2026-06-15
dev.to
developer-tools

Three GPU affiliate programs I wired into an AI tool directory

A developer integrated three GPU cloud affiliate programs—RunPod, Vast.ai, and Hetzner Cloud—into an AI tools directory after finding Amazon's conversion weak for developer-adjacent products. The mone…

04:00
2026-06-05
arxiv.org
large-language-models

LoRi: Low-Rank Distillation for Implicit Reasoning

Researchers have developed LoRi, a low-rank distillation framework that improves implicit reasoning in large language models by aligning teacher and student reasoning trajectories within a shared low-…

15:19
2026-05-20
dev.to
large-language-models

What did gemma see? - Thinking in comments...

The Gemma 4 26B model was the first local AI to achieve a perfect score on the HumanEval benchmark, including solving the notoriously difficult problem 145. This problem requires sorting integers by t…

// co-occurs with top 8 entities
// topics top 6 topics