Reposted by @anaymehrotra.bsky.social
Karan and Du, followed by @haithambouammar.bsky.social et al., showed that inference-time sampling from carefully chosen distributions improves LLM reasoning; no posttraining, reward curation, or verifier needed.
We show smarter test-time budget allocation yields drastic gains!
(1/3)