NVIDIA Nemotron 3 Embed Ranks #1 on RTEB
NVIDIA's Nemotron 3 Embed model tops the RTEB leaderboard with 78.5% accuracy, advancing agentic retrieval capabilities.
432 articles
New AI models as they ship — large language models, multimodal systems, and open-weight releases. Capabilities, context windows, pricing, and availability, summarized so you can compare what is actually new.
NVIDIA's Nemotron 3 Embed model tops the RTEB leaderboard with 78.5% accuracy, advancing agentic retrieval capabilities.
Analysis suggests smaller, open-source models can already handle many tasks efficiently, with costs and energy use remaining manageable.
Kimi's K3 model, with 2.8 trillion parameters and a one-million-token context window, outperforms most systems in its own benchmarks, signaling a shift in Chinese AI pricing.
xAI’s Grok 4.3 is now generally available on Amazon Bedrock, ranking #1 on several benchmarks and offering a 1 million token context window.
Google rebrands NotebookLM as Gemini Notebook, used by 30 million people and 600,000 organizations.
Google announced on Thursday that users can now create AI videos featuring their own digital avatars, using a selfie and voice recording.
Hugging Face releases LightOn-rerank, a 2B model achieving 62.66 NDCG on ViDoRe V3, outperforming existing open-source 2B-class multimodal rerankers.
Google Vids now lets users generate and edit high-quality videos using natural language, with Gemini Omni and personal avatars available for AI Pro and Ultra users.
Sakana AI adds Nvidia's Nemotron models to its Fugu orchestrator, claiming its system can match single frontier models in performance.
Moonshot AI’s Kimi 3 is set to close the gap with Anthropic’s Opus 4.8, with a parameter count between 2 trillion and 3 trillion, according to Financial Times.
Google announced on Thursday that users can now link and interact with select apps like Instacart, Canva, and YouTube within AI Mode, its conversational search experience.
Google released a stealth update for Gemma 4, improving tool calling and response completion, with performance boosts up to 70% on Nvidia Hopper GPUs.