cd/entity/Transformers· home entities Transformers
grep -l @transformers /news/*.json | wc -l → 36

Transformers

mentions 36 type Organization page 1/2 feed RSS

// recent coverage 36 mentions

23:21
2026-08-18
baseten.co
ai-infrastructure

Inference Engineering by Philip Kiely – Digital Download

Philip Kiely's new book, 'Inference Engineering,' is now available as a digital download, offering a comprehensive guide to the technologies and techniques powering AI inference across runtime, infras…

04:05
2026-08-15
dev.to
large-language-models

Exploring Qwen 3.8 27B: A Powerful AI Model for Developers

Qwen 3.8 27B, a 27-billion-parameter language model, has been released on the Hugging Face platform, offering improved performance for natural language processing tasks such as text generation, transl…

17:00
2026-08-13
promptcube3.com
large-language-models

Why is the DeepSeek-V4-Pro-0813 repo acting so strange on

DeepSeek's V4-Pro-0813 model repository is causing runtime errors during deployment due to a mismatch between the config.json and the .safetensors weights, as reported by a user attempting to load the…

11:49
2026-08-10
dev.to
artificial-intelligence

Only Two AI Updates Cleared My 36-Hour Cutoff

A developer's review of recent AI releases found only two updates met a strict 36-hour cutoff: Meta's Muse Glimmer, a 30B multimodal model under Apache 2.0, and Hugging Face's Transformers 5.15.0. The…

17:23
2026-07-27
cryptobriefing.com
artificial-intelligence

Moonshot completes Kimi K3 rollout with full model weight release

Moonshot AI released the full Kimi K3 model weights and technical report on July 27, 2026, making its 2.8 trillion parameter mixture-of-experts model available to developers and researchers. The model…

15:13
2026-07-27
huggingface.co
artificial-intelligence

Kimi K3 License

Moonshot AI has released Kimi-K3, an image-text-to-text model under a permissive license that allows use, modification, and commercial distribution, with support for Transformers, vLLM, SGLang, and Do…

10:56
2026-07-17
alanbonnici.com
generative-ai

ComfyUI and "Package Hell"

ComfyUI's shared Python environment creates dependency conflicts, as exemplified by Qwen3-TTS requiring transformers==4.57.3 while newer models need Transformers 5.x, causing workflows to break. The a…

02:50
2026-07-12
huggingbay.xyz
artificial-intelligence

mixedbread-ai/mxbai-rerank-xsmall-v1

Mixedbread AI released the mxbai-rerank-xsmall-v1 model on Hugging Face under an Apache-2.0 license, a text-ranking model for the Transformers library. The model has garnered 553,930 downloads and 57 …

08:19
2026-07-11
huggingbay.xyz
artificial-intelligence

black-forest-labs/FLUX.2-klein-4B

Black Forest Labs released FLUX.2-klein-4B, an image-to-image model under the Apache-2.0 license, now available on Hugging Face with over 470,000 downloads. The model is designed for use with diffuser…

23:50
2026-07-08
letsdatascience.com
artificial-intelligence

Hugging Face Speeds Transformers Inference in vLLM

Hugging Face announced that its Transformers modeling backend for vLLM now matches or exceeds native vLLM speed for compatible architectures, reducing the need for separate hand-optimized serving port…

00:00
2026-07-07
huggingbay.xyz
artificial-intelligence

answerdotai/ModernBERT-base

Answerdotai released ModernBERT-base, a 1.2 GB encoder model under Apache-2.0 license, designed for fill-mask tasks with easy local fit on consumer hardware. The model is available via Hugging Bay wit…

00:00
2026-07-07
huggingbay.xyz
large-language-models

Qwen/Qwen3-4B-Instruct-2507

Qwen released Qwen3-4B-Instruct-2507, a 4-billion-parameter instruct model under the Apache-2.0 license, designed for local deployment with quantized files requiring 8-16 GB of RAM/VRAM. The model is …

00:00
2026-07-07
huggingbay.xyz
large-language-models

Qwen/Qwen2.5-1.5B-Instruct

Qwen released Qwen2.5-1.5B-Instruct, an Apache-2.0 licensed 1.5B parameter instruct model for local chat and agent tasks, requiring 8-16 GB RAM/VRAM. The model is available on Hugging Bay with externa…

00:00
2026-07-07
huggingbay.xyz
large-language-models

Qwen/Qwen3-0.6B

Qwen released Qwen3-0.6B, a small Apache-2.0 licensed language model designed for local experiments, lightweight agents, and edge testing. The model is available on Hugging Bay with external metadata …

00:00
2026-07-07
huggingbay.xyz
artificial-intelligence

distilbert/distilbert-base-uncased

Hugging Bay lists distilbert/distilbert-base-uncased, a 260 MB DistilBERT encoder under Apache-2.0, as a compact resilience fallback for high-demand AI artifacts. The model has 500 upstream downloads …

page 1 / 2 next →
// co-occurs with top 8 entities
// topics top 6 topics