Remember work history to reduce repeated explanationsCoding tasks become faster and more accurateApproved defenders can advance advanced vulnerability research with GPT-5.6-CyberCyclone forecasts now provide over a day of extra lead timeTalk while reasoning or using tools without conversation breaksAI delivers new results on 10 long-standing math problems with proofs releasedGPT-5.6 Luna and Terra prices drop 80% and 20%, with faster Sol option addedOpus 5 now available on all paid plans and APISecurely link health records to understand symptom changes and test results in contextRun code inside notes for deeper analysisGPT-Red boosts prompt injection resistance significantlyCut lesson prep time with AIYou can move from conversation to documents fasterRun AI inference in the browser and cut wait timeReview how you use Claude and cut wasteLong tasks can move from draft to presentation more easilyTrack the latest safety rules for bigger modelsSee how Anthropic judges risky model misuseEasily automate multi-step daily tasks at lower costMake Claude easier to deploy through AWSRemember work history to reduce repeated explanationsCoding tasks become faster and more accurateApproved defenders can advance advanced vulnerability research with GPT-5.6-CyberCyclone forecasts now provide over a day of extra lead timeTalk while reasoning or using tools without conversation breaksAI delivers new results on 10 long-standing math problems with proofs releasedGPT-5.6 Luna and Terra prices drop 80% and 20%, with faster Sol option addedOpus 5 now available on all paid plans and APISecurely link health records to understand symptom changes and test results in contextRun code inside notes for deeper analysisGPT-Red boosts prompt injection resistance significantlyCut lesson prep time with AIYou can move from conversation to documents fasterRun AI inference in the browser and cut wait timeReview how you use Claude and cut wasteLong tasks can move from draft to presentation more easilyTrack the latest safety rules for bigger modelsSee how Anthropic judges risky model misuseEasily automate multi-step daily tasks at lower costMake Claude easier to deploy through AWS
Official sources only. Rumors, leaks, and get-rich schemes are excluded.
← Back to top
AI BriefingOpenAIFeature Updates00:00

AI summarized from verified sources

Measure how well AI agents handle ambiguous biology research judgments

Delegate biology data analysis judgments to AI, improving research efficiency.

SOURCE CHECK

1 sources

VERIFIED

Sources

Key Points

  • 1129 research-level benchmark questions
  • 2Synthetic data for rigorous evaluation
  • 3GPT-5.6 Sol reaches 31.5%

OpenAI introduced GeneBench-Pro with 129 problems across genomics and clinical genetics. It evaluates AI agents on data exploration, analysis path selection and judgment calls. GPT-5.6 Sol achieves up to 31.5% pass rate, assisting tasks that take human experts 20-40 hours for just a few dollars.

What happened

OpenAI announced GeneBench-Pro on June 30. It benchmarks AI judgment and iterative analysis in computational biology.

Impact

Advances practical AI agent use, making analysis support more accessible for researchers.

What changed

OpenAI introduced GeneBench-Pro with 129 problems across genomics and clinical genetics. It evaluates AI agents on data exploration, analysis path selection and judgment calls. GPT-5.6 Sol achieves up to 31.5% pass rate, assisting tasks that take human experts 20-40 hours for just a few dollars.

Briefs that include this news

Use daily, weekly, and monthly briefs to understand the surrounding context.

h
hayami

Stay on top of OpenAI, Google & Anthropic updates. An AI digest for business professionals.

Source Policy

We use only official sources. Each article links to the original announcement so you can verify it yourself.

© 2026 hayami. All rights reserved.