ResearchGenerative causal testing translates LLM brain-prediction models into testable hypothesesGenerative causal testing (GCT), developed by teams at Microsoft Research, UC Berkeley, UCSF and Columbia University, converts black-box language-model brain predictors into short, testable verbal explanations and verifies them by generating new stories for fMRI experiments.4 min read
ResearchOpenAI: GPT-5 Pro revealed the key mechanisms of an earlier experimentOpenAI said in an article and infographic that researchers used GPT-5 Pro to interpret the results of an experiment from several years ago; according to the article, the AI was able to identify the…1 min read
ResearchHow reasoning prompts unlock stored facts in large language modelsA new study shows that asking reasoning-capable large language models (R-LLMs) to generate stepwise ‘‘chain-of-thought’’ traces can surface correct factual answers that are otherwise unrecoverable from the model’s parametric memory.5 min read
ResearchHassabis on Creativity: AI, Simulation and the 'Einstein Test'At Cannes Lions, Google DeepMind CEO Demis Hassabis framed creativity as a form of simulation that AI must master to push science forward.3 min read
ResearchSelf-taught researcher says he has deciphered 3,500-year-old Linear A using AI toolsTom Di Mino, a self-taught AI engineer and amateur linguist from New York's Hudson Valley, claims to have systematically decoded the 3,500-year-old Minoan script Linear A after five months of solo work aided by Python scripts and an AI agent.3 min read
ResearchOpen Far-Field ASR Leaderboard (FFASR) Measures Real-World RobustnessTreble Technologies and Hugging Face have launched the Far-Field ASR (FFASR) Leaderboard, an open, community-driven benchmark that evaluates speech recognition models under realistic far-field acoustic conditions.4 min read
ResearchAI model GPT‑5 Pro helped reveal how deoxyglucose steers T‑cell specializationImmunologist Derya Unutmaz used GPT‑5 Pro to revisit a 2022 experiment and uncovered a mechanistic explanation for why deoxyglucose drives developing T cells toward an inflammatory Th17 fate.4 min read
ResearchDFlash block-diffusion drafter boosts LLM inference throughput up to 15× on NVIDIA BlackwellDFlash, an open-source block-diffusion model for speculative decoding, converts sequential token drafting into parallel block generation and verification, improving throughput while preserving target-model quality.4 min read
ResearchStudy finds advanced AI more persuasive than human experts in text conversationsResearchers from the University of Oxford, the UK AI Security Institute, Stanford University and the London School of Economics ran four experiments with 6,923 participants and 18,978 conversations and found that current AI systems outperform humans in text-based persuasion, including real-money charitable donations.4 min read
ResearchDeepMind outlines possible routes from AGI to superintelligenceResearchers at Google DeepMind published a paper mapping ways the field could move from human-level artificial general intelligence (AGI) to artificial superintelligence (ASI).4 min read
ResearchHumanoid robot progress will shape timelines for self-sustaining AIForecasters argue that whether artificial intelligence becomes self-sustaining depends largely on the development of humanoid robots and physical infrastructure.3 min read
ResearchStartup Recursive demonstrates automated-research system improving language-model training and GPU kernelsRecursive, a newly founded AI research startup, demonstrated that its automated research system achieved new best results on three benchmarks: NanoChat Autoresearch, NanoGPT Speedrun, and SOL-ExecBench.2 min read