cd/entity/Transformer· home› entities› Transformer
grep -l @transformer /news/*.json | wc -l → 124

Transformer

mentions 124 type Organization page 6/7 feed RSS

// recent coverage 124 mentions

16:36
2026-06-15
recursive.com
artificial-intelligence

First Steps Toward Automated AI Research

Recursive's automated AI research system achieved state-of-the-art results on three benchmarks: fixed-budget language model training, small-model training speed, and GPU kernel optimization. The syste…

05:00
2026-06-06
the-information-bottleneck.com
artificial-intelligence

Jürgen Schmidhuber: World Models, RL and Year That Changed AI

Jürgen Schmidhuber, a pioneer in artificial intelligence whose lab developed foundational ideas like LSTM, world models, and artificial curiosity decades before they became mainstream, argued in a new…

12:00
2026-06-03
kdnuggets.com
large-language-models

5 Fun Papers That Explain LLMs Clearly

Five foundational research papers explain how large language models work, covering the Transformer architecture, in-context learning, scaling laws, and instruction tuning with human feedback. The pape…

04:00
2026-06-03
arxiv.org
machine-learning

Graph Mamba Survival Analysis Based on Topology-Aware ordering

Researchers have developed TopoMamSurv, a Graph Mamba survival analysis framework that uses topology-aware ordering to address computational bottlenecks in Whole Slide Image analysis. The framework in…

04:00
2026-06-03
arxiv.org
machine-learning

Geometry-Aware Tabular Diffusion

Researchers introduced Geometry-Aware Tabular Diffusion (GATD), a method that improves tabular data synthesis by feeding pairwise angles and lengths from column value differences into diffusion denois…

16:16
2026-05-28
blog.kog.ai
large-language-models

Delayed Tensor Parallelism for Faster Transformer Inference

Kog Team researchers introduced Delayed Tensor Parallelism (DTP), a Transformer architecture that hides communication overhead behind computation and weight streaming to accelerate batch-size-one infe…

15:30
2026-05-27
dev.to
large-language-models

LLM Prompt Caching: The Complete 2026 Guide

Prompt caching can reduce LLM input costs by 50–90% and improve time-to-first-token by 3–10× without quality loss, according to a 2026 developer guide. The optimization stems directly from Transformer…

← prev page 6 / 7 next →
// co-occurs with top 8 entities
// topics top 6 topics