OpenAI Launches Computational Biology Benchmark Genebench-Pro; GPT-5.6 Full Version Has a Correct Rate of Only 30%
2026-07-30 18:29:10
According to CoinMeta, OpenAI has released a computational biology evaluation benchmark Genebench-Pro to test the multi-step decision-making capabilities of AI agents in complex scientific research scenarios such as genomics and translational medicine. The new benchmark consists of 129 questions (82 of which have been reviewed by external experts), and it generates data with clear causal relationships through computer simulations to prevent models from cheating by taking shortcuts or catering to the preferences of the question setters. Test results show that even top models struggle significantly when dealing with scientific reasoning that involves quantitative uncertainties, with the pass rate for GPT-5.6 being only 31.5%. In addition, former researcher Andrew Ho of OpenAI has left to start his own business, focusing on creating high-quality reinforcement learning data for large models. It is expected that over $100 billion will be invested in precise training data in the future.
Source:Internet
This content is for market information only and does not constitute investment advice.
Follow CoinMeta official accounts to stay updated

Hot Articles
Refresh

Ethereum: Ethereum Enters its Second Decade: Foundation Restructuring and Institutional Adoption in Parallel
29m ago

web3: MoonPay launches AI wallet PayBox, USDC rewards bring in a large number of new users.
39m ago

Zoox Receives US Regulatory Exemption to Advance Fee-Based Robotaxi Services
1h ago

Ethereum: Ripple issues 15 million RLUSD on Ethereum.
1h ago

web3: Avalanche testnet launches Hel token issuance and upgrades.
1h ago



