cd/entity/Llama· home› entities› Llama
grep -l @llama /news/*.json | wc -l → 330

Llama

mentions 330 type Organization page 10/17 feed RSS
sameAs · en.wikipedia.org · wikidata.org

// recent coverage 330 mentions

12:17
2026-07-18
dev.to
artificial-intelligence

The Real Moat in Legal AI Isn't the Model—It's the Data

A developer investigating EvenUp's success in legal AI found that the real competitive advantage is not a proprietary model but years of structured data from hundreds of thousands of personal injury c…

19:12
2026-07-17
insideai.news
artificial-intelligence

Meta in Talks for $10 Billion Anthropic Compute Deal in the US

Meta Platforms is in advanced talks to lease computing power to Anthropic in a deal that could be worth $10 billion, according to a New York Times report. The agreement would make Meta a critical infr…

07:55
2026-07-17
dev.to
large-language-models

Attention Sinks: Why Streaming LLMs Break When You Evict Token 0

A developer explains that attention sinks—tokens at position 0 that absorb excess attention weight—cause streaming LLMs to fail when evicted from the KV cache. The softmax normalization forces the mod…

04:56
2026-07-16
thegustafson.com
natural-language-processing

WordPiece, Unigram, and SentencePiece

WordPiece, Unigram, and SentencePiece are three subword tokenization algorithms that differ in how they split text, with WordPiece using a likelihood-maximizing merge criterion and a ## continuation m…

00:00
2026-07-16
machinelearning.apple.com
large-language-models

Embarrassingly Simple Self-Distillation Improves Code Generation

Simple self-distillation (SSD) improves LLM code generation by fine-tuning models on their own sampled outputs, boosting Qwen3-30B-Instruct from 42.4% to 55.3% pass@1 on LiveCodeBench v6, with gains c…

10:20
2026-07-15
dev.to
ai-agents

I’m sick of AI “Thinkslop” in my PRs

A developer known as 'Dessinateur Projeteur' built a local-first engine called NexaVerify that runs code through 8 different AI models to catch hallucinations and logic errors in pull requests. The to…

08:55
2026-07-15
machinebrief.com
artificial-intelligence

AI Confidence: Why Precision Doesn't Always Equal Awareness

A study applying Signal Detection Theory to four AI models—Llama-3-8B-Instruct, Mistral-7B-Instruct-v0.3, Llama-3-8B-Base, and Gemma-2-9B-Instruct—across 224,000 factual QA trials found that models va…

14:13
2026-07-13
byteiota.com
artificial-intelligence

Together AI Raises $800M: Open-Source Inference Just Got Serious

Together AI closed an $800 million Series C at an $8.3 billion valuation, reporting $1.15 billion in annual bookings and an inference engine that hits 500 tokens per second on DeepSeek-V3.1. The compa…

← prev page 10 / 17 next →
// co-occurs with top 8 entities
// topics top 6 topics