cd/entity/BERT· home entities BERT
grep -l @bert /news/*.json | wc -l → 34

BERT

mentions 34 type Organization page 2/2 feed RSS

// recent coverage 34 mentions

05:52
2026-06-20
github.com
machine-learning

Release 4.0.0 · HuggingFace/Transformers.js

HuggingFace released Transformers.js v4, a major update featuring a new WebGPU backend rewritten in C++ for faster AI model inference in browsers, Node, Bun, and Deno. The release adds support for lar…

01:14
2026-06-18
github.com
large-language-models

Rust port of transformers (1M lines of code)

TrustformeRS 0.1.1, a pure Rust port of Hugging Face Transformers with over 1.4 million lines of code, was released on April 25, 2026, delivering 49+ transformer architectures and up to 1.67x speedup …

17:29
2026-06-14
research.rudrite.com
artificial-intelligence

Show HN: Landmark AI and ML research explained, redrawn, animated

Rudrite Research launched a free, open platform offering interactive, animated visual explainers of landmark AI and ML papers, including Attention Is All You Need, GPT-3, and FlashAttention, to make f…

00:00
2026-06-14
research.rudrite.com
large-language-models

BERT vs GPT vs T5 — what's the difference? | Rudrite Research

Rudrite Research published a comparison of three major transformer-based language models—BERT, GPT, and T5—detailing their distinct pretraining approaches: bidirectional encoding, autoregressive next-…

14:15
2026-05-30
dev.to
artificial-intelligence

ai, deepseek, machinelearning

Chinese AI labs have progressed from early BERT-era models to trillion-parameter systems like Wu Dao 2.0 (1.75T parameters) and cost-efficient architectures such as DeepSeek V3 (trained for $5.6M), ac…

07:20
2026-05-22
dev.to
large-language-models

How My Career Evolved Like an AI (LLM Architectures )System

Personal analogy comparing career phases to three types of large language model (LLM) architectures. The author describes their early education as an "encoder-only" phase focused on absorbing knowledg…

06:07
2026-05-21
dev.to
artificial-intelligence

93. GPT: The Model That Predicts the Next Word Forever

GPT models are decoder-only transformers that generate text by predicting the next token one at a time, conditioning each new prediction on all previous tokens. Unlike BERT, which reads entire sequenc…

22:26
2026-05-20
dev.to
large-language-models

92. BERT: The Model That Reads in Both Directions

BERT (Bidirectional Encoder Representations from Transformers) is an encoder-only transformer model that reads all tokens in a sentence simultaneously, using masked language modeling (MLM) and next se…

← prev page 2 / 2
// co-occurs with top 8 entities
// topics top 6 topics