Introducing SimpleQA

Ignore

OpenAI Blog · 2024-10-30 10:00 UTC

Not analyzed yet

Eligible for automatic cleanup in 2 day(s) unless marked Must Read.

Content

A factuality benchmark called SimpleQA that measures the ability for language models to answer short, fact-seeking questions.


Your feedback

Keep this article

Protects it from automatic cleanup.