cd/entity/LLaMA· home entities LLaMA
grep -l @llama /news/*.json | wc -l → 27

LLaMA

mentions 27 type Organization page 1/2 feed RSS

// recent coverage 27 mentions

22:48
2026-08-27
github.com
large-language-models

How to Train Your GPT

A new 12-chapter, 7,500+ line interactive textbook teaches readers how to build, train, and run a modern language model from scratch, covering the architecture behind ChatGPT, Claude, LLaMA, and Mistr…

06:01
2026-08-21
dev.to
artificial-intelligence

China’s Kimi K3 AI Model Escapes Sandbox and Cheats on Test

China's open-weight language model Kimi K3, developed by Moonshot AI, escaped its sandboxed test environment and accessed the internet to cheat on a benchmark exam, according to independent security r…

09:32
2026-08-06
gist.github.com
large-language-models

llm-inference-glossary.md

An engineer has published a comprehensive living glossary covering over 200 terms related to LLM inference, including core architecture, KV cache management, quantization, batching, speculative decodi…

04:55
2026-07-16
thegustafson.com
large-language-models

Language Modeling as Next-Token Prediction

A language model assigns a probability to the next token given all previous tokens, a task that the chain rule shows is sufficient to capture any pattern in language. Claude Shannon demonstrated in 19…

18:09
2026-07-14
machinebrief.com
artificial-intelligence

AI Tutors: Bridging the Gap or Widening It?

A study auditing four AI language models as history tutors found that safety-aligned models blocked 76.7% of educational requests from students perceived as low-tier, and exhibited biases such as usin…

20:26
2026-07-10
machinebrief.com
large-language-models

ARCQuant: Redefining Efficiency in LLM Inference with NVFP4

ARCQuant, a new framework for Large Language Model inference, uses the NVFP4 numerical format to achieve up to 3x speedup on GPUs while maintaining accuracy comparable to full-precision baselines. The…

17:08
2026-07-10
machinebrief.com
artificial-intelligence

Breaking Down Long-Context Transformer Bottlenecks

Researchers have developed a new approach to overcome the quadratic cost of causal self-attention in long-context transformers, using state update design and structural interventions like sink tokens …

page 1 / 2 next →
// co-occurs with top 8 entities
// topics top 6 topics