Prompt Learning vs GEPA, benchmarked
We ran the same benchmarks used in the GEPA paper, but for Prompt Learning.
With some eval engineering, here are the results we got:
HotpotQA, GPT-4.1 Mini
HoVer, GPT-4.1 Mini
AI text/layout recreation from video frame; verify against source image.