1841 articles
A new AI framework called OpenNovelty helps peer reviewers spot unoriginal research by checking claims against real published papers. The system uses multi-stage verification to make academic review more transparent and reliable.
New 2025 web traffic numbers reveal ChatGPT's grip on consumer AI isn't loosening, even as Grok posts explosive growth. The data paints a picture of one dominant player and a handful of scrappy challengers trying to break through.
A senior researcher at DeepMind says we're closing in on artificial general intelligence faster than many expected. He's putting a coin-flip probability—50 percent—on minimal AGI arriving by 2028, sparking fresh debate about how soon transformative AI might reshape the economy.
A fresh early-access chapter on LLM self-refinement just dropped, pushing inference-time scaling techniques way beyond basic self-consistency and voting. The update brings an iterative critique-and-refine loop with fully functional code implementations.
Qwen3-TTS just dropped on ModelScope with a complete open-source text-to-speech model family. The release includes five models, full weights, source code, and technical documentation—everything developers need to start building.
Enterprise SSD prices have skyrocketed 257% in under a year due to AI infrastructure demand, while HDD costs rose just 35%. The widening price gap is making all-flash data centers dramatically more expensive to build and operate.
A compact 10-billion-parameter Chinese AI model delivers impressive benchmark performance while running entirely on consumer laptops, marking a shift toward efficient, locally-deployable AI systems.
Evernorth partners with t54 Labs to put AI-powered financial agents straight onto the XRP Ledger, automating treasury operations and compliance checks without human oversight.
Grok Code grabbed the top spot on Kilo Code leaderboard at launch and hasn't budged since. The model's also crushing it as 2025's most-used AI coding tool, racking up usage numbers that leave competitors in the dust.
Google's launching a new "Personal Intelligence" feature in AI Mode that lets you connect apps like Gmail and Photos for more personalized search results. The feature's rolling out to subscribers first as an experimental Labs option.
Grok just pulled ahead of DeepSeek in global website traffic for the first time, hitting 3.5% market share in mid-January 2026. The data from Similarweb shows how fast things can shift in the AI space.
OpenAI's GPT-5.2 just crushed the FrontierMath benchmark, landing near the top across multiple math-focused tests. The results show serious progress in advanced reasoning capabilities.
New open-source AI model matches human players in real-time 3D games by learning from massive gameplay datasets rather than traditional reinforcement learning methods.
PostgreSQL gains native BM25 relevance ranking through a newly open-sourced extension, eliminating the need for external search infrastructure.
Anthropic released new interpretability research outlining the "Assistant Axis," a method to measure and constrain AI behavior across 275 distinct archetypes. The work aims to reduce persona drift without harming model capability.
Suzhou Lumos Robotics and Shanghai Damon Technology introduced a wheeled humanoid robot for intelligent logistics, demonstrating automated handling of loads up to 50kg in operational warehouse conditions.
McKinsey data reveals a dramatic acceleration in AI skill requirements, with demand for AI fluency jumping sevenfold in just two years.
Unreleased Google Gemini models codenamed Snowbunny topped the Hieroglyph benchmark with 16 out of 20 points, outpacing other frontier AI systems in lateral reasoning tests.
OpenAI's revenue hit $20 billion in 2025, but the company burned through $9 billion while projecting expenses to balloon to $47 billion by 2028—raising questions about long-term sustainability.
Researchers from Microsoft and CUHK Shenzhen have introduced a new training approach to improve reasoning in speech-based AI models. The method significantly narrows the gap between speech and text reasoning performance.
We use cookies to improve your experience on our site and to show you relevant advertising. To find our more, read our privacy policy and cookie policy