cd/entity/RoPE· home entities RoPE
grep -l @rope /news/*.json | wc -l → 19

RoPE

mentions 19 type Organization feed RSS

// recent coverage 19 mentions

22:01
2026-08-22
pub.towardsai.net
artificial-intelligence

One Formula to Map the Positional Encoding Landscape

A new survey of positional encoding methods in Transformers argues that the field is best understood not as a chronological progression but as answers to a single question: where position information …

04:00
2026-07-23
arxiv.org
artificial-intelligence

AdaRoPE: Not All Attention Heads Should Rotate and Scale Equally

A new study from arXiv introduces AdaRoPE, a method that assigns learnable rotation frequencies and attention scaling factors to each attention head in Transformers, outperforming standard Rotary Posi…

23:13
2026-07-22
idlemachines.co.uk
artificial-intelligence

Thinking Machines dropped RoPE, and it's a good idea

Thinking Machines released Inkling, a 975B parameter mixture-of-experts model with 41B active parameters and a 1M token context length, which does not use Rotary Position Embeddings (RoPE). Instead, I…

19:13
2026-07-22
dev.to
large-language-models

RoPE: How 2D Rotations Solved Transformer Long-Context

Rotary Position Embedding (RoPE), introduced by Su et al. in 2021, solves the Transformer long-context problem by rotating Query and Key vectors in 2D sub-planes, making attention depend only on relat…

09:29
2026-06-24
blog.chuanxilu.net
large-language-models

An analysis on why LLMs perform bad on long loop tasks

A technical analysis reveals that large language models fail at long loop tasks due to attention dilution, EOS bias, and stateless architecture, causing protocol drift where constraints are rewritten …

20:00
2026-06-20
discuss.huggingface.co
large-language-models

DNA, LLM and Wick-Leger Correspondance (2nd Rosetta Stone)

A new appendix maps structural parallels between DNA and large language models, identifying a shared developmental-ledger pattern of possibility, gate, commitment, ledger, inheritance, development, an…

09:21
2026-06-18
discuss.huggingface.co
artificial-intelligence

Shannon Prime Lattice

Researchers developed XBAR, an auditable latent crossbar memory fabric that enables model-to-model communication by writing directly into a frozen transformer's KV cache. The system achieves O(1) VRAM…

18:57
2026-06-16
injuly.in
large-language-models

Inference cost at scale with napkin math

A technical analysis calculates the dollar cost per user for serving large language models at scale using napkin math, breaking down GPU resources, matrix multiplication costs, and attention mechanism…

// co-occurs with top 8 entities
// topics top 6 topics