11:09
2026-09-27
vibeleaderboard.ai
ai-agents
Have a model read full agent traces to rewrite prompts, not just RL
GEPA, a prompt-optimization method in which a model reads an agent's full trace โ reasoning, tool calls and errors โ and rewrites the prompt, reportedly doubled the gains of GRPO after a single round โฆ