cd/entity/Llama· home entities Llama
grep -l @llama /news/*.json | wc -l → 261

Llama

mentions 261 type Organization page 6/14 feed RSS
sameAs · en.wikipedia.org · www.wikidata.org

// recent coverage 261 mentions

11:52
2026-07-21
lesswrong.com
artificial-intelligence

I ran the standard AI litmus tests on my two toddlers (yep)

Engineer Carlo Valenti built his own transformer engine from scratch in C over 18 months to understand AI claims of sentience, then ran the same litmus tests on his two toddlers, finding that his daug…

09:38
2026-07-21
blog.devgenius.io
large-language-models

Ollama Was Fun for About Two Weeks. Then Reality Showed Up.

Ollama, a tool for running large language models locally, initially impresses with ease of use but quickly reveals critical limitations for production use, according to a user account. The tool hides …

04:00
2026-07-20
arxiv.org
large-language-models

An MLIR-Based Compilation Method for Large Language Models

A new MLIR-based compilation method for large language models, developed by researchers and implemented in the TPU-MLIR compiler and LLM-TPU deployment project, addresses challenges in importing train…

21:26
2026-07-19
techpowerup.com
artificial-intelligence

Sale: Get ChatGPT, Gemini, Claude, and More for Life for $80

A lifetime subscription to 1min.AI, which provides access to GPT-5.5 Pro, Claude Opus and Sonnet, Gemini 3.1 Pro, Llama, and Mistral, is on sale for $79.97 (reg. $540). The platform also offers SEO re…

03:48
2026-07-19
thomasdullien.github.io
large-language-models

RL economics, morally charged terms, and "distillation"

Reinforcement learning (RL) is the primary driver of recent advances in coding and mathematics for large language models (LLMs), according to an analysis of the economics of model improvement after hu…

12:17
2026-07-18
dev.to
artificial-intelligence

The Real Moat in Legal AI Isn't the Model—It's the Data

A developer investigating EvenUp's success in legal AI found that the real competitive advantage is not a proprietary model but years of structured data from hundreds of thousands of personal injury c…

19:12
2026-07-17
insideai.news
artificial-intelligence

Meta in Talks for $10 Billion Anthropic Compute Deal in the US

Meta Platforms is in advanced talks to lease computing power to Anthropic in a deal that could be worth $10 billion, according to a New York Times report. The agreement would make Meta a critical infr…

07:55
2026-07-17
dev.to
large-language-models

Attention Sinks: Why Streaming LLMs Break When You Evict Token 0

A developer explains that attention sinks—tokens at position 0 that absorb excess attention weight—cause streaming LLMs to fail when evicted from the KV cache. The softmax normalization forces the mod…

04:56
2026-07-16
thegustafson.com
natural-language-processing

WordPiece, Unigram, and SentencePiece

WordPiece, Unigram, and SentencePiece are three subword tokenization algorithms that differ in how they split text, with WordPiece using a likelihood-maximizing merge criterion and a ## continuation m…

00:00
2026-07-16
machinelearning.apple.com
large-language-models

Embarrassingly Simple Self-Distillation Improves Code Generation

Simple self-distillation (SSD) improves LLM code generation by fine-tuning models on their own sampled outputs, boosting Qwen3-30B-Instruct from 42.4% to 55.3% pass@1 on LiveCodeBench v6, with gains c…

10:20
2026-07-15
dev.to
ai-agents

I’m sick of AI “Thinkslop” in my PRs

A developer known as 'Dessinateur Projeteur' built a local-first engine called NexaVerify that runs code through 8 different AI models to catch hallucinations and logic errors in pull requests. The to…

← prev page 6 / 14 next →
// co-occurs with top 8 entities
// topics top 6 topics