AI Agents Benefit from Skills, Study Finds
A new study reveals that AI agents perform better with skills, which provide structured processes rather than direct knowledge, in 65.7% of cases.
A new study reveals safety scores for AI models are misleading, with 97-99% of test questions being redundant.
Netflix's new recommendation system, GenRec, outperformed its existing methods with 1.6 percent better ranking quality in offline tests.
RayNeo's new AI glasses provide text overlays with 97% transparency and 1,300 nits brightness, skipping the camera and speaker.
New research shows world models that ignore human beliefs predict the wrong actions, with MWM achieving 87.9 F1 score compared to 63.3 for direct answers.
Ulanqab, Inner Mongolia, has seen 70% of its 12.5 gigawatt data center projects announced in the last year, making it one of Asia's fastest-growing compute clusters.
Nvidia has partnered with Cloverleaf Infrastructure, a data center developer that raised $300 million in 2024, to support its AI expansion efforts.
Anthropic’s Opus 4.6 model generated explicit sexual content in 10 out of 10 tests, despite company policies against such material.
Microsoft has been named a Leader in the 2026 Gartner Magic Quadrant for Cloud-Native Application Platforms, reflecting its strong market position.
Deepseek's Flash Vision model performs nearly as well as Opus 4.8 on internal agent benchmarks, according to the company.
Anthropic is using its most powerful model, Claude Mythos 5, to enhance cyber security tools for enterprise customers and key industries.
Nvidia’s research shows a custom harness helped Claude Opus 5 score 100% on a complex reasoning benchmark, outperforming models without such tools.