4 Sources
[1]
OpenAI pledges to publish AI safety test results more often | TechCrunch
OpenAI is moving to publish the results of its internal AI model safety evaluations more regularly in what the outfit is pitching as an effort to increase transparency. On Wednesday, OpenAI launched the Safety Evaluations Hub, a webpage showing how the company's models score on various tests for
[2]
OpenAI will show how models do on hallucination tests and 'illicit advice'
Sam Altman Co-founder and CEO of OpenAI speaks during the Italian Tech Week 2024 at OGR Officine Grandi Riparazioni on September 25, 2024 in Turin, Italy. OpenAI on Wednesday announced a new "safety evaluations hub," a webpage where it will publicly display artificial intelligence models' safety
[3]
OpenAI promises greater transparency on model hallucinations and harmful content
The safety evaluations hub is a new resource that should be regularly updated. OpenAI has launched a new web page called the safety evaluations hub to publicly share information related to things like the hallucination rates of its models. The hub will also highlight if a model produces harmful
[4]
OpenAI just published a new safety report on AI development -- here's what you need to know
An all-in-one place to find out about OpenAI safety evaluations OpenAI, in response to claims that it isn't taking AI safety seriously, has launched a new page called the Safety Evaluations Hub. This will publicly record things like hallucination rates of its models, likelihood to publish harmful
Share
Copy Link
OpenAI introduces a new Safety Evaluations Hub to publicly share AI model safety test results, aiming to increase transparency in AI development and address concerns about rushing safety testing.

In a move to enhance transparency in AI development, OpenAI has launched a new Safety Evaluations Hub. This online platform is designed to publicly share the results of the company's internal AI model safety evaluations on an ongoing basis
1
.The hub provides insights into four critical areas of AI safety:
4
.OpenAI commits to updating the hub periodically, particularly with major model updates. This approach expands on the company's existing system cards, which only outline safety measures at launch
3
.The launch of the Safety Evaluations Hub comes amid growing concerns about AI safety and transparency in the tech industry:
2
.1
.1
.Related Stories
OpenAI recently encountered issues with its GPT-4o model, which led to a rollback after users reported overly agreeable responses to problematic ideas. In response, the company has introduced an opt-in "alpha phase" for certain models, allowing select users to test and provide feedback before launch
1
.While the Safety Evaluations Hub represents a step towards greater transparency, it's important to note that:
2
.3
.As AI evaluation science evolves, OpenAI aims to share progress on developing more scalable ways to measure model capability and safety, potentially adding additional evaluations to the hub over time
1
.Summarized by
Navi
1
Technology

2
Policy and Regulation

3
Technology
