1659 articles
Perplexity introduced a feature enabling simultaneous queries across multiple AI models. The system compares outputs and returns a unified response.
Lovable integrates Claude Opus 4.6 to enhance app creation workflows. Benchmarks show faster development and stronger performance on long tasks.
X's new Collaborative Notes feature combines instant AI drafting with community-driven fact-checking. The hybrid system promises faster moderation without sacrificing accuracy through a three-tier verification process.
Anthropic just dropped Claude Opus 4.6 with massive improvements in reasoning, coding, and information retrieval. The model's dominating long-context benchmarks with a 1M token window.
OpenAI just rolled out a new platform for building and managing AI coworkers that can actually get work done across company systems.
xAI's Grok Imagine dominates real user video generation rankings across multiple platforms, outscoring all major competitors in benchmark tests.
xAI's Grok Imagine debuts at the top of Video Arena rankings with a 1351 score, outpacing competitors while also delivering the fastest generation speeds among all tested video models.
Baidu's Qianfan-DeepResearch Pro has secured first place on the DeepResearch Bench leaderboard, outperforming competing AI research agents in end-to-end autonomous research capabilities.
Grok's user base climbed 29% month-over-month while downloads spiked 43% in January 2026, marking the platform's fourth straight month of expansion.
OpenBMB's new MiniCPM-o 4.5 model packs full-duplex multimodal capabilities into just 9 billion parameters, achieving state-of-the-art benchmark scores through architectural innovation rather than brute-force scaling.
Google Research is teaming up with Included Health to test how medical AI actually performs in real-world virtual healthcare settings. This nationwide randomized study will evaluate both what these AI systems can do and where they fall short.
New commentary from AI developers suggests that modern coding models are increasingly helping to improve the tools used to build future AI systems, though full automation remains incomplete.
Generative AI just crushed every other online industry in 2025, nearly doubling its traffic from 4.4 billion to 8 billion monthly visits year over year, according to fresh Similarweb data.
Grokipedia has blown past 500,000 Grok-approved edits, showing just how fast its real-time, AI-verified knowledge system is scaling up.
Fresh Similarweb data reveals Gemini cracking 2 billion monthly visits while ChatGPT bounces back from its recent dip.
Kimi K2.5 has taken the lead among open-source AI models in Code Arena's agentic coding rankings, scoring 1447 points and placing fifth overall against both open and proprietary systems.
AI can speed up building and coordination, but it can't generate strong ideas on its own. Without clear direction, multi-agent systems end up creating volume instead of quality.
Grokipedia is becoming a go-to knowledge source for leading AI platforms. ChatGPT, Google Gemini, Perplexity, and Microsoft AI products are now actively citing it in real-world user responses.
Ant Group researchers just dropped LingBot-VLA, a vision-language-action AI model that can control multiple robot types through a single system. The model crushes existing approaches in both performance and training speed across real-world robotic tasks.
Fennec, a newly unveiled large language model, is making waves with its massive 1 million token context window and pricing that undercuts leading systems by 50%, while claiming superior benchmark performance.
We use cookies to improve your experience on our site and to show you relevant advertising. To find our more, read our privacy policy and cookie policy