Microsoft Research, UC Berkeley, UCSF, and Columbia researchers introduced generative causal testing. The method converts black-box language-brain prediction models into short hypotheses, then tests them by having an LLM write stories meant to activate targeted cortical regions.
Sciences
RSS FeedThe HN thread cared less about a brand extension and more about whether AI-assisted ultrasound can change the cost and access curve for medical imaging.
OpenAI says GPT-5.5 Instant has pushed health responses close to its frontier Thinking models while reaching free ChatGPT users. The bigger signal is production data: flagged factuality issues in health answers fell 71% over two months.
AI is moving into one of medicine’s slowest review loops: revisiting old genome cases as the science changes. OpenAI and Boston Children’s Hospital report 18 new diagnoses after reanalyzing 376 unresolved pediatric rare-disease cases.
OpenAI is presenting a more concrete test for AI-assisted science: a chemistry project that reached a validated experimental result. The tweet says GPT-5.4 worked with Molecule.one’s Maria AI and a specialized lab on a drug-discovery reaction.
AI for life sciences is getting a more realistic yardstick. OpenAI says LifeSciBench was built with 173 biotech and pharma scientists and spans 750 expert-written tasks across seven biological research workflows.
A $400M round would lift CuspAI to a $2.6B valuation and put generative materials discovery back on the AI funding map. The company points to customers including ASML and Meta, plus a PFAS project that screened 300 trillion molecular structures.
Google-backed UC San Diego researchers plan to build a low-carbon cloud platform from 2,000 retired Pixel phones. The design strips devices to motherboards, groups 25-50 phones into Kubernetes-managed clusters, and targets teaching, grading, and research workloads.
Google Research is framing dermatology AI around user understanding, not just condition labels. A JAMA Dermatology study with 2,345 participants tested whether an AI-powered informational tool helped people identify skin concerns and choose better next steps.
The r/artificial thread focused less on banning AI tools and more on author responsibility when unchecked model output reaches the scholarly record.
Anthropic points to infrastructure, not only model intelligence, as the bottleneck for scientific agents. In an NCBI Virus retrieval task, accuracy rose to nearly 100% after adding a deterministic gget virus layer.
NMR analysis is a slow chemistry bottleneck, and Anthropic says Opus 4.7 matched or beat specialist tools on parts of a 20-compound test. Its hydrogen NMR average error was about plus or minus 0.079 ppm.