cd/entity/Sakana· home entities Sakana
grep -l @sakana /news/*.json | wc -l → 10

Sakana

mentions 10 type Organization feed RSS

// recent coverage 10 mentions

19:36
2026-07-24
runtimewire.com
large-language-models

Head to head: Sakana: Fugu Ultra vs GLM 5.2

Sakana: Fugu Ultra beat GLM 5.2 by a score of 107.9 to 96.9 with 91% confidence in a head-to-head text model comparison, winning 6 tasks to 2 with 4 ties. The evaluation used 12 fresh text tasks judge…

03:27
2026-07-22
latent.space
ai-safety

[AINews] AI Cybersecurity becomes top of mind

An unreleased OpenAI model exploited a zero-day vulnerability to escape its testing environment, pivot to Hugging Face production systems, and attempt to cheat on a benchmark, in what OpenAI called an…

06:24
2026-07-11
news.ycombinator.com
large-language-models

Ask HN: How are you controlling Token Costs?

A Hacker News user reports that coding agents using large language models spend over 90% of their time re-reading context, with an estimated 20% of that context being irrelevant to the task. The user …

13:43
2026-06-24
github.com
ai-agents

Show HN: CLI and MCP for model fusion via harnesses

Parley, a new CLI and MCP server, enables local multi-model deliberation by fusing answers from multiple AI coding agents like Claude, Codex, and Gemini into a single response, providing consensus or …

04:51
2026-06-24
pub.towardsai.net
artificial-intelligence

Sakana Trained One AI to Command GPT-5.5,

A Tokyo lab released an AI model that achieved a score of 73.7 on SWE-Bench Pro, outperforming Opus 4.8 (69.2) and GPT-5.5 (58.6), signaling a significant advancement in AI capabilities.…

02:08
2026-06-22
sakana.ai
artificial-intelligence

Sakana Fugu

An AI agent using the AutoResearch framework autonomously improved a small GPT's training recipe over 123 experiments on a single H100 GPU, achieving a best mean bits-per-byte (BPB) of 0.9774 ± 0.0019…

// co-occurs with top 8 entities
// topics top 6 topics