New benchmark confirms AI models still perform poorly at visual perception

New benchmarks show AI models struggle with visual perception tasks. This highlights ongoing challenges in improving AI's ability to understand and interpret visual information effectively.
More in Research
Don't classify. Hallucinate!
Simon Willison is pushing for AI to embrace hallucination instead of avoiding it. He argues that accepting these errors can lead to more creative and innovative outputs.
Study contradicts Anthropic and OpenAI claims that autonomous AI research is within reach
A new study challenges Anthropic and OpenAI's assertions that autonomous AI research is close to realization. This suggests that the timeline for achieving true AI autonomy may be longer than previously thought.

What are AI Hallucinations?
Databricks is defining AI hallucinations and their impact on model reliability. Understanding these hallucinations helps developers create more trustworthy AI systems.
Aug 13, 2026Frontier Red TeamPatterns and problems in emerging multiagent systems
Anthropic is analyzing patterns and problems in emerging multiagent systems. Their research aims to improve the design and functionality of these systems for better performance and reliability.