Propel: Breaking the Solver Bottleneck in Task-Generator RL
Training AI systems to generate harder tasks hits a wall when testing each task requires expensive solver runs. PROPEL bypasses this by using a lightweight activation probe on a frozen model to predict task difficulty instantly, replacing costly solver trials and doubling the rate of useful frontier tasks across math, coding, and software engineering benchmarks.