cd/entity/Llama· home› entities› Llama
grep -l @llama /news/*.json | wc -l → 330

Llama

mentions 330 type Organization page 15/17 feed RSS
sameAs · en.wikipedia.org · wikidata.org

// recent coverage 330 mentions

07:43
2026-06-19
xelerate.tech
ai-tools

One Model Won't Save You: How We Built Our AI Stack

Pedro Rocha, CTO of xelerate.tech, argues that a single AI platform cannot serve all business needs and describes his company's multi-model stack: GitHub Copilot Pro+ for development, Gemini for offic…

23:26
2026-06-17
discuss.huggingface.co
large-language-models

Introducing KerasFormers: "Transformers" for Keras 3!

KerasFormers, an open-source library built entirely in Keras 3, launches with over 100 transformer models spanning vision, language, multimodal, and speech, supporting seamless execution on TensorFlow…

21:45
2026-06-17
databricks.com
ai-agents

What is an AI agent harness?

An AI agent harness is the software infrastructure that wraps around a large language model to enable it to act on tasks, not just respond to prompts. The harness connects the model to tools, memory, …

06:55
2026-06-17
github.com
large-language-models

Native Inference Engine for macOS 14 or newer

Embershard, a macOS chat app with its own LLM inference engine, has been released in beta v0.1.1 for Apple Silicon devices running macOS 14 or newer. The app bypasses llama.cpp for inference, instead …

00:00
2026-06-17
runagentrun.co.uk
ai-infrastructure

OpenRouter fans prompts to match Claude Fable 5

OpenRouter launched Fusion, a routing layer that sends a single prompt to multiple AI models in parallel and synthesizes their outputs, achieving performance comparable to Anthropic's Claude Fable 5 a…

00:21
2026-06-16
byteiota.com
artificial-intelligence

Meta AI Mode Launches: Muse Spark API Coming This Month

Meta launched AI Mode on Facebook on June 15, replacing link-based search with AI-synthesized answers from public posts, Groups, and Reels, powered by its new Muse Spark model. The Muse Spark API is i…

09:27
2026-06-15
news.ycombinator.com
large-language-models

What are you looking for when reviewing LLM generated code?

A developer reviewing large pull requests with thousands of lines of AI-generated Rust code struggles to comprehend changes at an architectural level, relying on CI tooling for linting and compilation…

21:24
2026-06-12
cryptobriefing.com
artificial-intelligence

Meta’s AI unit faces chaos as executives struggle with strategy

Meta's AI unit is in chaos as executives and employees disagree on strategy, including whether to keep AI models open-source or shift to proprietary systems. The company is moving 7,000 employees into…

05:56
2026-06-12
github.com
large-language-models

LLM for the ESP32-S3

Two ESP32-S3 microcontrollers running a Llama-architecture language model have achieved the first multi-chip pipelined LLM inference on ESP32-class hardware, splitting layers across two boards connect…

04:00
2026-06-12
arxiv.org
large-language-models

Localizing Anchoring Pathways in Language Models

Researchers have identified specific neural pathways in large language models that carry anchoring bias signals, where irrelevant numbers in prompts skew numerical reasoning. Using attribution-based c…

00:00
2026-06-11
signoz.io
artificial-intelligence

Amazon Bedrock Monitoring and Observability with OpenTelemetry

Amazon Bedrock users can now monitor model performance, latency, error rates, and usage trends by integrating OpenTelemetry with SigNoz. The open-source observability framework exports logs, traces, a…

00:00
2026-06-09
iankduncan.com
large-language-models

What's in the Box? A Field Guide to AI Models

A software engineer explores the practicalities of running large language models locally, demystifying technical jargon like parameters, quantization, and MoE to help users choose and deploy models on…

00:00
2026-06-08
fergusfinn.com
artificial-intelligence

The economics of speculative decoding

Speculative decoding, a lossless inference optimisation that predicts future tokens to reduce latency, faces new economic constraints as modern mixture-of-experts (MoE) architectures replace dense tra…

← prev page 15 / 17 next →
// co-occurs with top 8 entities
// topics top 6 topics