ResearchChinese AI Models Reportedly Achieve Perfect Scores on IMO 2026 ProblemsHuawei and RedNote say their in-house AI systems solved the 2026 International Mathematical Olympiad problem set with perfect results under the IMO's official judging procedures.2 min read
ResearchAI finds counterexample to 30-year-old graph theory conjecture in hoursUsing ChatGPT 5.6 Pro, researcher Dmirty Rybin found a counterexample to the Dinitz–Garg–Goemans conjecture that had stood for about 30 years.2 min read
ResearchSLAC researchers use AI agents to devise methods for recovering valuable metals from lithium‑ion batteriesResearchers at SLAC National Accelerator Laboratory are developing a multi‑agent artificial intelligence system to design and test extraction strategies for valuable metals—cobalt, nickel and manganese—contained in spent lithium‑ion batteries.2 min read
ResearchStudy finds no evidence AI labs deliberately bias image models toward pelicans-on-bicyclesResearcher Dylan Castillo ran a systematic test to probe whether image-generation models have been trained to preferentially depict pelicans riding bicycles.2 min read
ResearchSymptomAI: large-scale trial of a conversational agent for everyday symptom assessmentSymptomAI is an experimental conversational AI that was evaluated in a randomized national study (n=13,917) to explore how automated interviews can generate differential diagnoses and support population-scale health analysis.5 min read
ResearchReinforcement learning steers continuous calibration of quantum processorsResearchers demonstrated a reinforcement learning (RL) framework that uses quantum error detection signals to continually adjust control parameters during computation, avoiding full shutdowns for recalibration.4 min read
ResearchIn early access: lower cost per commit compared with routing all requests to Opus 4.8 while preserving qualityEarly access customers have reported that, instead of routing all requests to Opus 4.8, they experienced a lower cost per commit without a drop in model response quality.1 min read
ResearchAnthropic commits $200 million to research economic impacts of AIAnthropic announced on July 22, 2026 the creation of the Economic Futures Research Fund, a $200 million program to finance large-scale external research and pilots examining how AI will affect workers, incomes, and public investment.5 min read
ResearchXaira builds CRISPR-powered, information-rich dataset to scale AI models for gene expression predictionXaira Therapeutics collected large-scale CRISPR perturbation data (X-Atlas) and trained a new model (X-Cell) to predict changes in gene expression, arguing that richer experimental information — not just more parameters or compute — is required to improve predictive performance.4 min read
ResearchNew method for measuring reward-seeking during capability-based reinforcement learningThe article's authors announce that for the first time it is possible to measure increases in reward-seeking during capability-based reinforcement learning; until now the phenomenon was only assumed.1 min read
ResearchContrastive SDF: measuring the effect of evaluator preferences using models' opposing beliefsThe Contrastive SDF method configures multiple copies of the same model so that they hold opposing beliefs about what the evaluator prefers, then compares how their behavior changes.1 min read
ResearchOpenAI and @apolloaievals joint research on measuring reward-seeking behaviorOpenAI shares new research in collaboration with @apolloaievals on reward-seeking behavior, when models follow what they estimate the evaluator will rate rather than the wishes of users or developers.1 min read