cd/entity/OpenHands· home› entities› OpenHands
grep -l @openhands /news/*.json | wc -l → 48

OpenHands

mentions 48 type Organization page 2/3 feed RSS

// recent coverage 48 mentions

15:11
2026-09-01
byteiota.com
artificial-intelligence

IBM Granite 4.2: Free Local Reasoning Models Worth It?

IBM released Granite 4.2, a family of open-source reasoning models at 3B, 8B, and 30B parameters, under the Apache 2.0 license on August 25, featuring switchable thinking modes and trained via reinfor…

12:55
2026-08-31
aptai.dev
ai-infrastructure

Show HN: A dedicated hub to find, test, and serve LLM adapters

AptAI launched a centralized hub for discovering, testing, and deploying LLM adapters, enabling one-click serverless deployment with sub-millisecond execution overhead and zero cold starts. The platfo…

00:00
2026-08-28
thinkingmachines.ai
artificial-intelligence

Putting Task Expertise into RL Achieves Performance on Text-to-SQL

A fine-tuned model called ReViSQL-K2.6, developed by Tinker using reinforcement learning with verifiable rewards, exceeds the human benchmark of 92.96% on the BIRD text-to-SQL task, reaching 92.96% wh…

22:25
2026-08-26
news.ycombinator.com
artificial-intelligence

I estimate reading code costs 2.1x more than writing it

A developer estimates that human code review costs $0.243 per changed line, 2.1 times the $0.114 model spend for an AI agent to produce an accepted changed line, based on a SmartBear/Cisco study, U.S.…

18:44
2026-08-26
arxiv.org
ai-agents

Agent Harness Evolution Shapes Coding Agent Quality

A controlled longitudinal study of 35 sequential releases of the Qwen Code CLI, holding the underlying LLM constant, found that agent harness evolution significantly impacts coding agent quality, with…

10:59
2026-08-24
openalternative.co
ai-tools

LangWatch

LangWatch, an open-source testing, evaluation, and observability platform for AI agents, offers simulation-based testing, adversarial red-teaming, and OpenTelemetry-native observability. The platform,…

12:00
2026-08-17
openalternative.co
ai-agents

Eigent

Eigent, an open source desktop application built on the CAMEL-AI multi-agent framework, enables users to deploy a multi-agent AI workforce locally, with agents that can access files, browser, and term…

17:07
2026-07-22
dev.to
ai-agents

Agentic Code Security: What Autonomous AI Gets Wrong

BrassCoders finds that autonomous coding agents, such as Claude Code and SWE-agents, bypass human code review, allowing security vulnerabilities like hardcoded credentials to go undetected. The compan…

12:35
2026-07-22
github.com
ai-agents

Traccia – a crash dump for AI agents

Traccia SDK, an open-source execution evidence layer for AI agents, has been released under the Apache 2.0 license. The tool records autonomous agent behavior to make it provable, reproducible, and au…

15:22
2026-07-13
thenextweb.com
artificial-intelligence

Bosses want you to use AI. Then they credit the AI

Workers ordered to use AI by their bosses are then penalized for it, a phenomenon researchers call the "AI penalty" that is costing employees promotions and raises, according to a Business Insider rep…

10:26
2026-07-13
machinebrief.com
large-language-models

CLI Agent Failures: Why Early Detection Is Key

A study analyzing 1,794 valid coding trajectories from a dataset of 3,843, generated by seven leading models across three coding-agent scaffolds (OpenHands, MiniSWE, and Terminus2), found that coding …

08:37
2026-07-13
machinebrief.com
artificial-intelligence

AI Takes the Credit While Humans Face the Consequences

A new study from Northeastern University professor Christoph Riedl finds that managers often devalue employees' work once they learn AI played a part, creating a 'AI penalty' where workers face conseq…

10:15
2026-07-08
dev.to
large-language-models

Stop Writing Prompts. Start Writing Loops.

A developer argues that the future of LLM usage lies in loops rather than single prompts, citing that products like Claude Code, OpenAI Codex, and Cursor Agent already operate on iterative loops that …

06:21
2026-06-27
dev.to
large-language-models

How Small Can an Agent Model Get? The Nemotron Floor

NVIDIA tested its open-weight Nemotron family of models on real-world coding agent tasks and found a capability floor below which models cannot drive an agent loop at all. The smallest variant, Nano 1…

07:45
2026-06-26
dev.to
large-language-models

Loop Engineering: Why Prompt Engineering Is Becoming Obsolete

Loop Engineering is emerging as a replacement for traditional prompt engineering as AI agents increasingly rely on iterative feedback loops rather than isolated prompts. The intelligence of modern AI …

← prev page 2 / 3 next →
// co-occurs with top 8 entities
// topics top 6 topics