Introducing MentalHealthBench
OpenAI has released MentalHealthBench, a benchmark designed with expert input. It assesses whether AI responses are helpful and safe in realistic conversations about mental health.
OpenAI has released MentalHealthBench, a benchmark designed with expert input. It assesses whether AI responses are helpful and safe in realistic conversations about mental health.
OpenAI CEO Sam Altman addressed the UN Security Council. He spoke about AI safety, keeping humans in control and the need for international cooperation.
OpenAI is giving Ukraine’s government access to its Daybreak program. The aim is to support the cyber defense of civilian infrastructure.
Google is adding private, server-side memory to Private AI Compute, its system for personal AI. The feature is designed to keep users’ data protected.
OpenAI sets out what it considers necessary for independent safety assessments of frontier models and their safeguards. It stresses rigor, security and independence.
A Hugging Face post covers work by the UK AI Security Institute and the EvalEval initiative to make AI benchmark results reproducible.