Inside Genebench-Pro
OpenAI just launched Genebench-Pro, a new framework for evaluating AI models on genetic data. This tool helps researchers assess model performance in genomics, making it easier to develop AI applications in healthcare.
More in Research
New benchmark confirms AI models still perform poorly at visual perception
New benchmarks show AI models struggle with visual perception tasks. This highlights ongoing challenges in improving AI's ability to understand and interpret visual information effectively.

Don't classify. Hallucinate!
Simon Willison is pushing for AI to embrace hallucination instead of avoiding it. He argues that accepting these errors can lead to more creative and innovative outputs.
Study contradicts Anthropic and OpenAI claims that autonomous AI research is within reach
A new study challenges Anthropic and OpenAI's assertions that autonomous AI research is close to realization. This suggests that the timeline for achieving true AI autonomy may be longer than previously thought.

What are AI Hallucinations?
Databricks is defining AI hallucinations and their impact on model reliability. Understanding these hallucinations helps developers create more trustworthy AI systems.