All AI news
Browse, filter, and search every article in the archive. The homepage shows the last 24 hours; everything older lives here.
Security vulnerability reports have exploded since AI models started hunting for bugs
AI models are now actively hunting for security vulnerabilities, leading to a surge in vulnerability reports. This means organizations can identify and address security issues faster than ever before.

Meta's AI agent push is moving slower than Zuckerberg planned
Meta is slowing down its AI agent development, falling behind Zuckerberg's initial timeline. This means users might wait longer for advanced AI features that were expected to roll out sooner.

Vercel's Andrew Qu on why agents are a new kind of software
Vercel's Andrew Qu emphasizes that agents represent a new category of software, focusing on their ability to perform tasks autonomously. This shift could change how developers approach building applications and workflows.
Agent Runs now available in the Vercel MCP and CLI
Vercel just launched Agent Runs in their MCP and CLI. This feature allows developers to run AI agents directly within their infrastructure, streamlining workflows and enhancing automation capabilities.
Mark Zuckerberg tells staff that AI agents haven’t progressed as quickly as he’d hoped
Mark Zuckerberg tells Meta staff that AI agents aren't advancing as quickly as expected. He emphasizes the need for faster progress to stay competitive in the AI landscape.
llm-coding-agent 0.1a0
Simon Willison just released llm-coding-agent 0.1a0. This new coding agent helps automate programming tasks, making it easier for developers to streamline their workflows.
Using DSPy to evaluate and improve Datasette Agent's SQL system prompts
Simon Willison is using DSPy to enhance the SQL system prompts for the Datasette Agent. This improvement aims to make the agent's SQL queries more effective and user-friendly.
The Download: a startup has a solution for AI’s groupthink problem
A startup just unveiled a solution to combat groupthink in AI systems. This innovation aims to enhance decision-making by ensuring diverse perspectives are considered in AI outputs.
Autoresearch: The feedback loop behind self-improving agents
Autoresearch is developing self-improving agents that learn from their own performance feedback. This approach enhances their ability to complete tasks autonomously and adapt over time.
How Cursor deploys AI inside the enterprise
Cursor is integrating AI tools directly into enterprise workflows to enhance productivity. This allows teams to leverage AI for various tasks without needing extensive technical knowledge.
Building a serverless A2A gateway for agent discovery, routing, and access control
AWS just launched a serverless A2A gateway for agent discovery and routing. This simplifies how agents access and communicate with each other, enhancing overall efficiency in multi-agent systems.
Structured memory filtering with metadata in AgentCore Memory
AWS just introduced structured memory filtering with metadata in AgentCore Memory. This enhancement allows AI agents to manage and retrieve information more efficiently, improving their performance in complex tasks.
Anthropic’s long-sidelined Fable 5 is greenlit to return
Anthropic just greenlit the return of Fable 5, which had been sidelined for a while. This means they're moving forward with developing this AI agent, potentially enhancing their offerings in autonomous systems.

Claude Science is Anthropic’s newest flagship product
Anthropic just launched Claude Science, their latest flagship AI product focused on scientific tasks. This tool aims to assist researchers by generating insights and automating data analysis, making scientific workflows more efficient.
ScarfBench: Benchmarking AI Agents for Enterprise Java Framework Migration
Hugging Face just launched ScarfBench, a benchmarking tool for AI agents focused on migrating Enterprise Java frameworks. This tool helps developers evaluate agent performance in real-world migration tasks, making the process smoother and more efficient.
Anthropic launches Claude Sonnet 5 as a cheaper way to run agents
Anthropic just launched Claude Sonnet 5, offering a more cost-effective way to run AI agents. This update makes it easier for developers to implement agent-based solutions without breaking the bank.
Acti puts AI agents directly into your smartphone keyboard
Acti just integrated AI agents into smartphone keyboards. Users can now access AI assistance directly while typing, making it easier to generate text and get suggestions on the go.
NVIDIA BioNeMo Agent Toolkit Brings Accelerated AI to Life Sciences Researchers in Claude Science
NVIDIA just launched the BioNeMo Agent Toolkit for life sciences researchers. This toolkit accelerates AI workflows, enabling researchers to leverage AI for complex biological tasks more efficiently.
Anthropic’s Claude Science bets on workflow, not a new model, to win over scientists
Anthropic is focusing Claude Science on enhancing workflow for scientists instead of launching a new model. This shift aims to better integrate AI into scientific processes, making it more useful for researchers.
Have your agent record video demos of its work with shot-scraper video
Simon Willison just introduced a shot-scraper video tool for AI agents to record video demos of their work. This feature allows agents to showcase their capabilities visually, enhancing user understanding and engagement.