We’re introducing GeneBench-Pro, a research-level benchmark for a harder kind of AI progress: how well agents can navig…
By OpenAI · AI Agents
We’re introducing GeneBench-Pro, a research-level benchmark for a harder kind of AI progress: how well agents can navigate messy biological data, choose the right analysis path, and make judgment calls that real computational research depends on.