Large Language Monkeys

Book

A paper that proposes repeatedly asking the same input problem to an LLM and using a verifier to select the best response, improving performance of smaller models.

Mentioned in 1 video