AI Search Agents Rely on Memory, Not Web Research
A study reveals leading AI search agents often confirm existing knowledge rather than research the web, with models like GPT-5.4 and Kimi-K2.6 scoring high on static benchmarks.
A study reveals leading AI search agents often confirm existing knowledge rather than research the web, with models like GPT-5.4 and Kimi-K2.6 scoring high on static benchmarks.
A new review paper by Meta, Stanford, and the University of Illinois Urbana-Champaign highlights code as the foundation for AI agents’ reasoning and actions, citing real-world examples like Claude Code and Codex.
A large-scale study shows that making AI chatbots helpful reduces their ability to mimic human behavior, with the effect worsening across generations.
Mathematician Terence Tao suggests AI could transform math research by enabling collaboration, as seen in industry and natural sciences. He argues AI can fill skill gaps in math teams, though challenges remain in verification.
In February 2026, METR found most developers would not work without AI, even for limited tasks, raising concerns about long-term code quality.