# SMELT: Scaling Laws for Compute-Matched MoE Looped Transformers

> Source: <https://aiflash.com/news/112582/>
> Published: 2026-09-02 03:30:00+00:00

Looped Transformers increase effective depth by iterating a shared block of layers, but most evaluations compare at fixed model size, conflating architectural advantage with extra FLOPs. We study looping on Mixture-of-Experts Transformers while closely matching per-token FLOPs, total non-embedding p
