cd/entity/nanoGPT· home entities nanoGPT
grep -l @nanogpt /news/*.json | wc -l → 14

nanoGPT

mentions 14 type Organization feed RSS

// recent coverage 14 mentions

11:46
2026-08-20
blog.devgenius.io
artificial-intelligence

Why 2031 Might Be the Last Year Humans Do AI Research

Ryan Greenblatt, Chief Scientist at Redwood Research, argued on the Dwarkesh Podcast that by 2030 or 2031, AI research and development will be fully automated by AI systems, triggering a recursive sel…

12:16
2026-08-16
primeintellect.ai
artificial-intelligence

Measuring Autonomous AI Research

A new public experiment by Prime Intellect ran 153 autonomous runs on the nanoGPT optimizer speedrun across 18 frontier models, finding that Claude Fable 5 and Opus 5 dramatically outperformed others,…

04:00
2026-08-05
arxiv.org
machine-learning

Sphere Retraction Normalizations

Researchers introduced p-SpheretNorm, a one-parameter family of angular retractions for spherical residual streams in deep neural networks, showing that the exponential map used in Geodesic Normalizat…

00:00
2026-07-17
zonted.com
artificial-intelligence

I Trained a GPT on 15 Years of My Life — Then Interrogated It

A developer trained a 10.77M-parameter GPT on 4.45M characters of personal Google data from 15 years, achieving a loss of 1.60 after pretraining and supervised fine-tuning on 8,520 email reply pairs, …

16:26
2026-07-16
lesswrong.com
ai-safety

Competitive AI Safety is

Patrick O'Driscoll, a former nanotech physicist and current AI architect, introduces Competitive AI Safety as a paradigm to focus the field on measurable, tractable goals, drawing inspiration from Ope…

21:01
2026-07-08
pub.towardsai.net
large-language-models

How to Build Your Own Tiny LLM From Scratch

A guide explains how to build a tiny large language model from scratch, covering tokenization, pretraining, supervised fine-tuning, and alignment. It aims to demystify the LLM pipeline for developers …

16:38
2026-06-29
lesswrong.com
large-language-models

Gradient-free Single-pass Model Beats nanoGPT on Shakespeare

A new character-level language model called EntropyBeam, using gradient-free count tables and a Dirichlet prior, achieved a validation loss of 1.596 nats on the Shakespeare character benchmark, outper…

10:53
2026-06-27
dev.to
machine-learning

How I Implemented GPTQ from Scratch (and What I Learned)

A developer implemented GPTQ quantization from scratch on a nanoGPT model, achieving only 1.1% perplexity degradation across 61 quantized layers. The implementation uses second-order optimization to r…

09:54
2026-06-13
dev.to
artificial-intelligence

Build Your Own Shakespearean LLM

A developer built a character-level language model from scratch using Shakespeare's complete works, training a nanoGPT model on a consumer-grade MacBook Pro in about 15 minutes. The project demonstrat…

00:00
2026-06-10
pyrefly.org
machine-learning

Talk: Tensor Shapes in the Type System

Pyrefly, a Python type checker, has introduced an experimental feature that brings tensor shapes into Python's type system, allowing shape annotations to become inferred type hints instead of comments…

// co-occurs with top 8 entities
// topics top 6 topics