Iterative Generation-Selection Improves LLM Creativity via Small Evaluators
August 10, 2026
Iterative generation-selection can match human creativity benchmarks in recipe generation, but increasing iteration counts yields diminishing returns. The size of the in-loop selection scorer is the most critical factor, as smaller evaluators significantly outperformed larger models across most Torrance Tests of Creative Thinking dimensions.
HOW THIS AFFECTS YOU
●
builderConsider using smaller, specialized models as in-loop scorers to optimize the performance-to-compute ratio in creative agents.
●
researcherYou should focus on evaluator architecture rather than iteration depth to drive creative performance.