OpenAI HealthBench: Sam Altman-Run Company Introduces New Evaluation Benchmark To Assess AI in Healthcare; Check Details
OpenAI has introduced HealthBench, a new evaluation benchmark to evaluate AI performance in healthcare. Developed with input from 250 physicians across 60 countries, it features 5,000 health conversations and custom rubrics to grade AI responses.
Sam Altman-run OpenAI has introduced a new evaluation benchmark called HealthBench to better measure how well AI systems work in healthcare. It was created with input from over 250 physicians who have practiced in 60 countries. It is now available on OpenAI’s GitHub repository. OpenAI said, “HealthBench includes 5,000 realistic health conversations, each with a custom physician-created rubric to grade model responses.” ChatGPT New Feature Update: OpenAI Announces Users Can Export Files Generated Using Chatbot's Deep Research Into PDF Format With Tables, Images and More.
OpenAI HealthBench
Evaluations are essential to understanding how models perform in health settings.
HealthBench is a new evaluation benchmark, developed with input from 250+ physicians from around the world, now available in our GitHub repository.https://t.co/s7tUTUu5d3
— OpenAI (@OpenAI) May 12, 2025
(SocialLY brings you all the latest breaking news, fact checks and information from social media world, including Twitter (X), Instagram and Youtube. The above post contains publicly available embedded media, directly from the user's social media account and the views appearing in the social media post do not reflect the opinions of LatestLY.)
(The above story first appeared on LatestLY on May 13, 2025 11:59 AM IST. For more news and updates on politics, world, sports, entertainment and lifestyle, log on to our website latestly.com).