Skip to content
All tags

#test-time-scaling

1 posts

Test-Time Scaling: BrowseConf and Confidence-Guided Reasoning

The previous articles covered evaluation. This one covers another dimension: how to dynamically allocate compute during reasoning. BrowseConf's core insight is that an agent's self-declared 'confidence' can predict answer accuracy. High confidence uses fewer resources; low confidence searches more rounds.