cd/entity/Opus· home entities Opus
grep -l @opus /news/*.json | wc -l → 160

Opus

mentions 160 type Organization page 3/8 feed RSS

// recent coverage 160 mentions

17:39
2026-07-16
lesswrong.com
artificial-intelligence

The getting is good (optimizing unattended runs)

A user reports that AI models differ drastically in how long they can run unattended on a host with sudo without causing system issues: Opus breaks hosts within 2 agent-hours, Gemini 3 Pro within 12, …

11:42
2026-07-16
blog.getcassis.com
large-language-models

Do you put rules or examples in your LLM context?

A study by Cassis on LLM context optimization for data analytics found that adding a full example library raised Haiku's accuracy from 70% to 82% and Sonnet's from 81% to 93%, while removing rules and…

06:49
2026-07-16
gist.github.com
ai-agents

Fable Orchestrator

A developer introduced Fable Orchestrator, a multi-agent system that coordinates specialized AI executors—including models like Sonnet, Opus, and Codex gpt-5.6-sol—for software development tasks. The …

12:30
2026-07-15
andywidjaja.com
artificial-intelligence

The $110/month self-improving pipeline

A developer has built an open-source system called autoloop that autonomously triages, implements, and tests GitHub issues using Claude AI, achieving 27 autonomous merges in two weeks with a 97% succe…

08:45
2026-07-15
news.ycombinator.com
large-language-models

Ask HN: How are you productive with GPT 5.6 Sol?

A user on Hacker News reports that GPT 5.6 Sol is significantly less productive than Opus/Fable, requiring constant steering and producing overreached conclusions and useless defensive code. The user …

05:08
2026-07-15
stencil.so
artificial-intelligence

You only need the frontier model for one single edit

A new analysis by AI researcher and open-source harness developer finds that the common practice of using a frontier model to plan and a cheaper model to execute—dubbed the '/plan' pattern—actually in…

04:25
2026-07-15
machinebrief.com
artificial-intelligence

EG-VAR: Setting a New Standard in AI Reasoning

EG-VAR, a Lean 4-based architecture for AI reasoning, achieved a flawless 120 out of 120 on TableBench numerical reasoning tasks and maintained 100% source fidelity during counterfactual stress tests,…

00:00
2026-07-15
hacktron.ai
ai-tools

Breaking Into PostHog Prod Database

A security researcher using Hacktron's pentest tool discovered that PostHog launched Playwright Chromium with `--no-sandbox` for heatmap screenshots, then used Anthropic's Claude to write an exploit f…

20:46
2026-07-13
thezvi.wordpress.com
artificial-intelligence

Better Call Sol The Workhorse

OpenAI released GPT-5.6-Sol, alongside cheaper models Terra and Luna, pricing Sol at $5/$30, Terra at $2.50/$15, and Luna at $1/$6. CEO Sam Altman said the models address enterprise cost concerns and …

13:00
2026-07-13
dev.to
artificial-intelligence

Building an Autonomous Agent on an M1 Mac, by Choice

A developer has been running an autonomous agent on a 16GB M1 Mac for three months, using small models in the 9B/E4B class by choice rather than economic necessity. The developer argues that small mod…

11:00
2026-07-13
dev.to
large-language-models

Prompt Caching, Batches API, and Model Routing to Cut LLM Costs

Anthropic's prompt caching, Batches API, and model routing can significantly reduce LLM costs without switching providers. Prompt caching reuses prefixes at 0.1× the base input price, the Batches API …

22:10
2026-07-12
gist.github.com
developer-tools

Setup 2025 - https://www.youtube.com/watch?v=6M7LgYkxS4g

A developer shared a starter configuration for Claude Code, a tool that uses AI to assist with coding tasks. The configuration includes a curated allowlist of safe commands, integration with the open-…

← prev page 3 / 8 next →
// co-occurs with top 8 entities
// topics top 6 topics