cd/entity/Llama· home› entities› Llama
grep -l @llama /news/*.json | wc -l → 330

Llama

mentions 330 type Organization page 11/17 feed RSS
sameAs · en.wikipedia.org · wikidata.org

// recent coverage 330 mentions

06:37
2026-07-13
machinebrief.com
artificial-intelligence

Memory-Managed Attention: Redefining AI's Long-Term Memory

A new study on memory-managed long-context attention achieves a perfect 1.000 score on Track A, a controlled retrieval task, compared to a 0.333 baseline, and outperforms dense retrieval methods by up…

17:12
2026-07-12
byteiota.com
artificial-intelligence

Mesh LLM Lets You Run 235B AI Models Without the Cloud

Mesh LLM, a new open-source tool, launched yesterday and hit the Hacker News front page, enabling users to pool GPUs across machines and run AI models up to 235B parameters via an OpenAI-compatible AP…

18:59
2026-07-11
dev.to
large-language-models

Model Kombat: The LLM Fighting Game!

A developer built Model Kombat, a 2D fighting game that visualizes Large Language Model architectures, parameter scales, and hardware constraints as playable mechanics. The game features fighters repr…

08:08
2026-07-10
machinebrief.com
artificial-intelligence

AI Job Search: The New Norm or Just Hype?

Software developer Tarun Gupta launched Autopilot-Jobhunt, a free AI tool that searches for job postings and generates tailored resumes and cover letters but does not submit applications. The tool use…

07:12
2026-07-10
dev.to
large-language-models

Large Language Models Demystified: A Visual and Practical Guide

A developer published a visual and practical guide to large language models, explaining that they are computer programs trained on massive text datasets to predict the next word. The guide covers how …

00:00
2026-07-10
neuronpedia.org
artificial-intelligence

Welcome to the J-Space 🌌

Neuronpedia announced the J-space, a global workspace in AI models revealed by Jacobian lens, and added support for 11 more models. The feature allows users to see hidden reasoning in AI, with pre-fit…

15:30
2026-07-09
research.ibm.com
artificial-intelligence

CoFrGeNets replace the ‘bones’ of transformer-based models

IBM Research introduced CoFrGeNets (Continued Fraction Generative Networks), a new model architecture that replaces transformer components with structures derived from continued fractions, enabling co…

14:18
2026-07-09
baremetalrt.ai
artificial-intelligence

BareMetalRT – TensorRT-LLM running natively on Windows (no WSL)

BareMetalRT launches TensorRT-LLM natively on Windows without WSL, enabling heterogeneous tensor parallelism across consumer GPUs over standard networking. Users can run HuggingFace models locally via…

14:10
2026-07-09
tokenstead.ai
ai-tools

Show HN: Tokenstead, find AI models for your hardware

Tokenstead, a new web tool, lets users select their hardware to find compatible open AI models with speed estimates and cloud-pricing comparisons, aiming to enable local AI model usage independent of …

12:15
2026-07-09
dev.to
developer-tools

Run Amazon Bedrock locally, with real completions from Ollama

MiniStack 1.4.0 introduces four new services that emulate Amazon Bedrock end-to-end, providing deterministic mock responses for Converse and InvokeModel APIs. A new environment variable allows the moc…

17:08
2026-07-08
dev.to
large-language-models

What Is LLM Orchestration? Patterns, Tools & When You Need One

LLM orchestration coordinates models, providers, and steps to make AI features production-grade, handling routing, failover, caching, guardrails, and observability. A dedicated orchestrator sits betwe…

13:26
2026-07-08
discuss.privacyguides.net
large-language-models

Best privacy-friendly cloud-based LLMs

A user reports that Ask Brave, which combines Llama and Qwen models with Brave's search indexing, is the best privacy-friendly cloud-based LLM option, outperforming Lumo and duck.ai's models like gpt-…

13:15
2026-07-08
dev.to
large-language-models

Building Production AI Systems(Part 2)

A developer details how OpenRouter serves as an abstraction layer for integrating multiple LLM providers into production AI applications. By simply changing the base URL and model name, developers can…

← prev page 11 / 17 next →
// co-occurs with top 8 entities
// topics top 6 topics