In agentic search, a GPT-OSS-120B agent using the 9B model answers 64.0% of BrowseComp+ questions correctly, 4.9 points ahead of the next ColBERT model, with fewer searches than any baseline.
In agentic search, a GPT-OSS-120B agent using the 9B model answers 64.0% of BrowseComp+ questions correctly, 4.9 points ahead of the next...
AISummary
In agentic search, a GPT-OSS-120B agent using the 9B model answers 64.0% of BrowseComp+ questions correctly, 4.9 points ahead of the next ColBERT model, with fewer searches than any baseline.
Source: Perplexity · x.com