[BidClub_]

PERSON DIRECTORY

Noam Brown

Noam Brown appears in 3 indexed conversations across Dwarkesh Podcast, Latent Space, No Priors. This directory brings every appearance, source, TL;DR, digest, and transcript into one searchable feed.

3 EPISODES3 SHOWS
1 episode1 active
Language
No PriorsEN · 36 min

Really Big Test-Time Compute in AI Changes Benchmarks, Safety and Research with OpenAI's Noam Brown

Sarah GuoNoam Brown

Noam Brown argues that model quality must be measured as a cost, token, or time curve, because fixed benchmark scores hide gains from test-time compute.Models can keep improving beyond 100 million tokens, while safety policies still lack a clear budget for evaluating cyber, bio, and other dangerous capabilities.Routing and orchestration businesses therefore face a demanding test: outperforming a single stronger model allowed to think longer at the same cost, with gains that transfer beyond benchmarks.