cd/entity/Transformers· home› entities› Transformers
grep -l @transformers /news/*.json | wc -l → 52

Transformers

mentions 52 type Organization page 1/3 feed RSS

// recent coverage 52 mentions

20:05
2026-09-13
arxiv.org
machine-learning

Solve the Loop: Attractor Models for Language and Reasoning

Researchers Jacob Fein-Ashley and colleagues posted a paper to arXiv on 12 May 2026 introducing Attractor Models, a recurrent architecture in which a backbone module proposes output embeddings and an …

12:09
2026-09-13
sourcefeed.dev
ai-research

Transformers v5 turned a library into a standard

Hugging Face released Transformers v5.17.0 on September 9, adding seven model architectures in a single minor release, including Tencent's 780-billion-parameter mixture-of-experts model and Moonshot A…

09:50
2026-09-10
discuss.huggingface.co
artificial-intelligence

How to Choose the Right Approach for AI Development Projects?

A community discussion post asks AI developers to share the factors they weigh when choosing a model, framework, or deployment strategy for real-world AI projects, citing tools including Transformers,…

02:46
2026-09-10
arxiv.org
machine-learning

Looped Transformers as Programmable Computers (2023)

Researchers posted a framework on arXiv on 30 January 2023 for using transformer networks as universal computers by programming them with specific weights and placing them in a loop. The paper demonst…

10:56
2026-09-09
github.com
artificial-intelligence

Low-code framework for building custom AI

Ludwig, a declarative deep learning framework hosted by the Linux Foundation AI & Data, announced support for Python 3.12, PyTorch 2.7+, Pydantic 2, Transformers 5, and Ray 2.54, enabling users to tra…

05:41
2026-09-09
huggingface.co
large-language-models

Introduction · Hugging Face

Hugging Face offers a free, ad-free course on large language models (LLMs) and natural language processing (NLP) using its ecosystem, including Transformers, Datasets, Tokenizers, and Accelerate, as w…

02:11
2026-09-09
promptcube3.com
machine-learning

Why India needs more ML infrastructure builders and fewer API

At a recent event, speakers including Aritra Roy Gosthipaty from Hugging Face emphasized the need for India to focus on machine learning infrastructure builders rather than API-level developers, highl…

17:25
2026-09-03
gilesthomas.com
artificial-intelligence

Putting my JAX-trained models on the Hugging Face Hub

Developer gpjt has uploaded PyTorch-compatible versions of all his JAX-trained GPT-2 models to the Hugging Face Hub, including models from his blog series on writing an LLM from scratch and Chinchilla…

00:00
2026-08-31
neuronpedia.org
ai-research

Introducing: interp-engine 🚀🔎

Decode Research has open-sourced interp-engine, a high-performance interpretability engine built from scratch to run production Neuronpedia workloads, including Jacobian Lens, NLAs, circuit tracing, a…

23:21
2026-08-18
baseten.co
ai-infrastructure

Inference Engineering by Philip Kiely – Digital Download

Philip Kiely's new book, 'Inference Engineering,' is now available as a digital download, offering a comprehensive guide to the technologies and techniques powering AI inference across runtime, infras…

04:05
2026-08-15
dev.to
large-language-models

Exploring Qwen 3.8 27B: A Powerful AI Model for Developers

Qwen 3.8 27B, a 27-billion-parameter language model, has been released on the Hugging Face platform, offering improved performance for natural language processing tasks such as text generation, transl…

17:00
2026-08-13
promptcube3.com
large-language-models

Why is the DeepSeek-V4-Pro-0813 repo acting so strange on

DeepSeek's V4-Pro-0813 model repository is causing runtime errors during deployment due to a mismatch between the config.json and the .safetensors weights, as reported by a user attempting to load the…

page 1 / 3 next →
// co-occurs with top 8 entities
// topics top 6 topics