cd/sources/hugging-face-blog· home› sources› Hugging Face Blog
cat /sources/hugging-face-blog.feed | wc -l → 1008

Hugging Face Blog

articles 1008 domain huggingface.co → page 25/51 feed RSS
23:37
2026-08-04
huggingface.co
ai-safety

Hostile system?

A proposal for a Real-time Telemetry Channel for AI Safety Filters aims to add transparency to automated content filtering by logging every rule applied, match found, and action taken, allowing real-t…

18:27
2026-08-04
huggingface.co
artificial-intelligence

MiniMax-H3 Quantisation RAM/VRAM

A user attempting to load the MiniMax-H3 text-to-video model from Hugging Face on a machine with 96GB RAM and an R9700 32GB GPU reports that even with 8-bit quantization of the transformer and text en…

13:58
2026-08-04
huggingface.co
artificial-intelligence

Deploy local agents everywhere with LFM2.5-2.6B

Liquid AI released LFM2.5-2.6B, a 2.6-billion-parameter on-device agentic model that outperforms models up to 4x larger on tool use and instruction following, achieving 220 tokens per second on an App…

10:58
2026-08-04
huggingface.co
artificial-intelligence

Decay-Gated O(N) Causal Linear Attention with Fused Triton Kernel

A developer has open-sourced a Decay-Gated O(N) Causal Linear Attention architecture with fused Triton/CUDA kernels, aiming to bypass quadratic multi-head attention bottlenecks. The project includes a…

10:35
2026-08-04
huggingface.co
artificial-intelligence

Chunking strategy for governments and internal org docs

A developer seeking advice on building a retrieval-augmented generation (RAG) system for government and internal organizational documents asks about optimal chunk sizes for BM25 and semantic search (c…

10:24
2026-08-04
huggingface.co
ai-infrastructure

Is there an response length limit for the inference API?

Hugging Face's Inference API returns only 2-3 sentences per response regardless of the model used, according to a user report. The issue may be resolved by setting the `max_new_tokens` parameter highe…

06:45
2026-08-04
huggingface.co
artificial-intelligence

Should travel apps use RAG instead of fine-tuning?

For travel apps, retrieval-augmented generation (RAG) is recommended over fine-tuning for frequently changing facts, with fine-tuning reserved for stable behavior, according to a technical guide refer…

03:34
2026-08-04
huggingface.co
ai-research

Open research catalog: 168 TTS and 106 speech-to-text systems

D3velop LLC's Open Gauntlet leaderboard lists 168 text-to-speech and 106 speech-to-text systems, but the ASR page's hero count of 111 is outdated, and the Fireworks and Yandex entries need updates. Th…

00:48
2026-08-04
huggingface.co
ai-products

Billing/Subscription

A user reported that their $9 subscription to Hugging Face's Hugging Chat Omni ran out of credits before one month, and Hugging Face staff clarified that the service uses Inference Providers on a pay-…

08:51
2026-08-03
huggingface.co
developer-tools

Command-Line Alternative to Postman for AI API Development?

A developer describes evaluating Apidog CLI as a command-line alternative to Postman for testing AI APIs, citing benefits for automation and CI/CD workflows. The developer notes that reusing API test …

← prev page 25 / 51 next →