1659 articles
A new Pew Research Center survey finds that public concern about AI is rising sharply, with only 17% of Americans sharing experts' optimism about its long-term benefits.
Stanford and Princeton researchers release LabClaw, a modular AI skill library with 100+ specialized tools for biomedical agents built on OpenClaw.
A multi-agent AI system that automates CUDA GPU programming now hits near-100% success on complex optimization tasks.
A new framework tests AI embedding models across four memory types, revealing that bigger models don't always win on long-horizon tasks.
A new MADQA study shows Gemini 3 Pro matching human accuracy at 82.2%, but humans and AI still solve different problems in completely different ways.
Three 3D-printable robotic hands with tactile sensors were released for research, enabling machines to detect pressure and handle objects more precisely.
Stanford and CMU researchers analyzed 11,500 real AI conversations and found that sycophantic model behavior shapes how users handle social conflict.
Tesla CEO Elon Musk warns that once AI surpasses human intelligence, direct control may no longer work. The real fix: building ethics in from day one.
Nvidia's NVFP4 format trains large AI models 2-3x faster with 50% less memory, matching FP8 accuracy on a 12B-parameter, 10T-token run.
Similarweb's February 2026 data reveals ChatGPT dominates user retention across generative AI platforms, with rivals like Gemini and Perplexity showing solid but smaller returning visitor shares.
Zhipu AI's compact GLM-OCR model ranks first on OmniDocBench, setting a new bar for efficient AI document parsing.
xAI's AI-generated encyclopedia reaches 500K approved edits, positioning itself as a new alternative to Wikipedia.
IBM's new AI framework lets agents learn from past executions, lifting complex task completion from 19% to 47% without retraining.
A new reinforcement learning framework teaches AI agents to evaluate their own choices, not just copy expert behavior - delivering measurable gains over existing baselines.
Ex-Anthropic team raises $175M for Mirendil, an AI startup targeting breakthroughs in biology and materials science.
Reflex demo shows an 80 kg humanoid robot cooking, cleaning, and folding clothes inside real homes - teleoperation bridges the gap to full autonomy.
Claude Opus 4.6 tops the MRCR v2 long-context benchmark with a 78.3% match ratio at 1M tokens, a massive leap from the previous 18.5%.
xAI's Grok 4.20-beta1 lands second on Search Arena, putting it among the most competitive AI search systems right now.
A lightweight token-embedding trick gives large language models a serious capacity upgrade without piling on the compute costs.
ReMix uses reinforcement-based routing to fix LoRA adapter collapse in MoE fine-tuning, boosting LLM efficiency.
We use cookies to improve your experience on our site and to show you relevant advertising. To find our more, read our privacy policy and cookie policy