GEPA: Reflective Prompt Evolution Can Outperform Reinforcement Learning
Researchers introduce GEPA (Genetic-Pareto), a prompt optimizer that uses natural language reflection to learn from trial and error, outperforming reinforcement learning methods like GRPO by 6% on aveβ¦