cd/sources/hugging-face-blog· home› sources› Hugging Face Blog
cat /sources/hugging-face-blog.feed | wc -l → 1007

Hugging Face Blog

articles 1007 domain huggingface.co → page 16/51 feed RSS
20:23
2026-08-25
huggingface.co
artificial-intelligence

Building Local: My 2026 Headless AI Server Journey

A developer reports that running Qwen 3.8 27B at Q5_K_M quantization on a dual AMD Radeon RX 7900 XT and 7800 XT setup achieves 20 tokens per second with a 256k context window, enabling autonomous mul…

15:14
2026-08-25
huggingface.co
large-language-models

Granite 4.2 LLMs: How They're Built

IBM's Granite Team released Granite 4.2, a family of dense, decoder-only reasoning LLMs in 3B, 8B, and 30B sizes, pre-trained from scratch on roughly 15 trillion tokens with a five-phase strategy exte…

08:32
2026-08-25
huggingface.co
ai-agents

Who Fills In the Form — We Only Sign What the Model Drafted

A developer exploring the execution-state-preflight reference skeleton on GitHub concludes that the model should not own the authority to decide that an execution input is settled, proposing that 'unk…

03:21
2026-08-25
huggingface.co
artificial-intelligence

Funasr using hugging face hub with paraformer-zh errors

A bug in FunASR's paraformer-zh model on Hugging Face Hub causes missing timestamps and a KeyError: 'text' error, according to a developer's investigation. The issue can be worked around by setting pr…

03:19
2026-08-25
huggingface.co
artificial-intelligence

Agentic curious questions

A user with an 8GB M2 MacBook Air is seeking advice on running an agentic AI for programming assistance, considering local versus cloud options and the future feasibility of running powerful models on…

00:00
2026-08-25
huggingface.co
ai-tools

Wire It, Run It, Deploy It: AI Workflows in Gradio

Hugging Face released gr.Workflow, a feature built into Gradio that turns AI pipelines into interactive drag-and-drop canvases where each node is runnable and every intermediate result is visible, whi…

22:26
2026-08-24
huggingface.co
artificial-intelligence

How do you design memory systems for long-running AI agents?

Michael, in a discussion on designing memory systems for long-running AI agents, advises that the runtime, not the model, should serve as the memory system, with applications making final decisions on…

22:25
2026-08-24
huggingface.co
artificial-intelligence

Why LLM agents keep failing (and it’s not the prompt)

A new paper on SSRN proposes ORCA, a cognitive runtime for LLM agents that structures reasoning as reusable components instead of embedding logic in prompts, addressing common failure patterns in agen…

← prev page 16 / 51 next →