ResearchOpenAI News
PaperBench: Evaluating AI’s Ability to Replicate AI Research
OpenAI has introduced PaperBench, a benchmark that evaluates the ability of AI agents to replicate state-of-the-art AI research.
Summary written by Kernelia from the original article by OpenAI News. The story and its rights belong to its author.
