Frontier ModelsWednesday, July 1, 2026
OpenAI introduces GeneBench Pro to test AI research judgment in computational biology
GeneBench Pro is a research level benchmark of 129 genomics, quantitative biology, and translational medicine problems that require agents to pick a methodology and reach a decision ready conclusion. OpenAI top model GPT-5.6 Sol solved 31.5 percent in Pro mode, underscoring how far agents remain from expert research taste.
Read the original source$ part of the KMM daily AI analysis · published automatically