48 articles
Qwen3.5-Omni enters the omni-modal AI arena with strong benchmark results across audio, video, and speech tasks - putting direct pressure on Google's Gemini lineup.
Gemini leads AI platform conversion rates at 6.8%, outpacing Perplexity, ChatGPT, and Claude in new aggregated data.
A new independent benchmark reveals CodeRabbit leading AI code review tools with 51.30% F1 score, while Gemini places third at 49.70%. The evaluation tracks both controlled tests and real-world bug fixes in open-source projects.
Google's Gemini 3.1 Pro High model produced high-quality voxel art, but frequent timeouts and reasoning errors are leaving users frustrated - raising questions about whether raw AI power means anything without a stable platform to run it on.
Google's Gemini 3 Deep Think achieved 84.6% on the ARC-AGI-2 reasoning benchmark, marking a significant leap in AI capabilities and intensifying the race among tech giants.
Leaked details show Google's Gemini 3.5 Pro features specialized reasoning modes, modular sub-models, and can generate 3,000 lines of working code from a single prompt.
Google has launched Gemini CLI v0.24.0, a free command-line interface that lets developers code, debug, and automate workflows directly from the terminal without switching tools.
Audience data shows a growing share of ChatGPT users also visiting Google's Gemini during 2025. The increase became more pronounced after the Nano Banana release in late August.
Google's Gemini 3 update brings independent daily quotas for "Thinking" and "Pro" models, giving users more control over complex problem-solving and coding tasks with limits ranging from 300 to 1,500 prompts depending on the tier.
Google's Stitch platform now supports Gemini 3 Flash, a faster AI model designed to speed up design iteration while maintaining quality. The Pro model remains available for tasks requiring deeper analytical thinking.
Google DeepMind launched the Gemini Interactions API in beta—a unified interface for Gemini models and AI agents—with plans to prioritize agent-focused features through 2026.
Google DeepMind just launched Gemma Scope 2, bringing interpretability tools to every layer of its Gemma 3 models—giving researchers unprecedented access to see how AI actually thinks.
New FrontierMath data from EpochAI reveals Chinese open-weight models are roughly seven months behind frontier AI systems in mathematical reasoning tasks, with the performance gap widening significantly on the most difficult problems.
Google launched Gemini 3 Flash, a new AI model that beats Gemini 2.5 Pro in performance while running three times faster and costing significantly less. The model is now live across Google's AI platforms.
Google is set to launch its first Gemini-powered smart glasses in 2026, featuring two product lines developed with partners like Samsung, Warby Parker, and Gentle Monster.
Grok 4.20 dominated Alpha Arena Season 1.5, taking first place with $16,171 in total equity while claiming four spots in the top ten rankings ahead of GPT-5.1, Gemini 3 Pro, and DeepSeek Chat V3.1.
Fresh evaluation data positions Google's Gemini 3 ahead of competing AI models in accuracy and speed during live browser testing, outperforming GPT-5 and Claude in real-world interaction scenarios.
Poetiq sets a new benchmark record on ARC-AGI-2 with a verified 54% score, surpassing Alphabet's Gemini 3 Deep Think. The system uses an open-source scaffold combining Gemini 3 Pro and GPT-5.1.
Google has released Gemini 3 Deep Think for Ultra subscribers. The model shows stronger reasoning - it scores 41 percent on Humanity's Last Exam and 45.1 percent on ARC-AGI-2, the highest results reported so far.
Fresh measurements from Google show Gemini now uses far less water, electricity, and CO₂ than earlier reports suggested. The updated data reveals a 33-fold jump in energy efficiency compared to widely circulated 2023 figures.
We use cookies to improve your experience on our site and to show you relevant advertising. To find our more, read our privacy policy and cookie policy