cd/entity/Transformer· home entities Transformer
grep -l @transformer /news/*.json | wc -l → 92

Transformer

mentions 92 type Organization page 5/5 feed RSS

// recent coverage 92 mentions

16:16
2026-05-28
blog.kog.ai
large-language-models

Delayed Tensor Parallelism for Faster Transformer Inference

Kog Team researchers introduced Delayed Tensor Parallelism (DTP), a Transformer architecture that hides communication overhead behind computation and weight streaming to accelerate batch-size-one infe…

15:30
2026-05-27
dev.to
large-language-models

LLM Prompt Caching: The Complete 2026 Guide

Prompt caching can reduce LLM input costs by 50–90% and improve time-to-first-token by 3–10× without quality loss, according to a 2026 developer guide. The optimization stems directly from Transformer…

20:57
2026-05-25
transformernews.ai
artificial-intelligence

Against the METR Graph

AI researcher Nathan Witkin has challenged the validity of METR's widely-cited Long Tasks benchmark, arguing its methodology is fundamentally flawed despite its status as a leading indicator of AI cap…

04:54
2026-05-22
arxiv.org
machine-learning

CODA: Rewriting Transformer Blocks as GEMM-Epilogue Programs

CODA, a GPU kernel abstraction that reparameterizes memory-bound Transformer operations like normalization and activations to execute as GEMM-plus-epilogue programs, keeping data on-chip to reduce glo…

← prev page 5 / 5
// co-occurs with top 8 entities
// topics top 6 topics