ResearchWeb retrieval, not model reasoning, is the main limiter for LLM news queriesA Stanford and Together AI study tested six large language models equipped with web-search tools on daily news questions in six languages between February 9 and February 22, 2026.4 min read
ResearchAI finds counterexample to 30-year-old graph theory conjecture in hoursUsing ChatGPT 5.6 Pro, researcher Dmirty Rybin found a counterexample to the Dinitz–Garg–Goemans conjecture that had stood for about 30 years.2 min read
ResearchOIST develops curiosity-driven learning that lets robots acquire language fasterResearchers at the Okinawa Institute of Science and Technology (OIST) built a curiosity-driven training approach for robots that speeds up acquisition of language-based tasks and produces child-like spontaneous play.3 min read
ResearchChinese AI Models Reportedly Achieve Perfect Scores on IMO 2026 ProblemsHuawei and RedNote say their in-house AI systems solved the 2026 International Mathematical Olympiad problem set with perfect results under the IMO's official judging procedures.2 min read
ResearchSLAC researchers use AI agents to devise methods for recovering valuable metals from lithium‑ion batteriesResearchers at SLAC National Accelerator Laboratory are developing a multi‑agent artificial intelligence system to design and test extraction strategies for valuable metals—cobalt, nickel and manganese—contained in spent lithium‑ion batteries.2 min read
ResearchStudy finds no evidence AI labs deliberately bias image models toward pelicans-on-bicyclesResearcher Dylan Castillo ran a systematic test to probe whether image-generation models have been trained to preferentially depict pelicans riding bicycles.2 min read
ResearchSymptomAI: large-scale trial of a conversational agent for everyday symptom assessmentSymptomAI is an experimental conversational AI that was evaluated in a randomized national study (n=13,917) to explore how automated interviews can generate differential diagnoses and support population-scale health analysis.5 min read
ResearchReinforcement learning steers continuous calibration of quantum processorsResearchers demonstrated a reinforcement learning (RL) framework that uses quantum error detection signals to continually adjust control parameters during computation, avoiding full shutdowns for recalibration.4 min read
ResearchIn early access: lower cost per commit compared with routing all requests to Opus 4.8 while preserving qualityEarly access customers have reported that, instead of routing all requests to Opus 4.8, they experienced a lower cost per commit without a drop in model response quality.1 min read
ResearchAnthropic commits $200 million to research economic impacts of AIAnthropic announced on July 22, 2026 the creation of the Economic Futures Research Fund, a $200 million program to finance large-scale external research and pilots examining how AI will affect workers, incomes, and public investment.5 min read
ResearchXaira builds CRISPR-powered, information-rich dataset to scale AI models for gene expression predictionXaira Therapeutics collected large-scale CRISPR perturbation data (X-Atlas) and trained a new model (X-Cell) to predict changes in gene expression, arguing that richer experimental information — not just more parameters or compute — is required to improve predictive performance.4 min read