The conversation around the evidence
Reader commentary
What does the research add, miss or leave unresolved? Share an experience, evidence link or question on any article. No account is required; every note is reviewed before it appears here.
Reader views are their own. Publication does not verify a claim or mean the newsroom endorses it. Disagreement is welcome; personal attacks, spam and confidential information are not.
The conversation is just starting. There are no approved reader notes yet; choose an article below to contribute.
Health & Life Sciences
A deep-learning model separated cure from relapse across 298 MRI examinations, but the effective test cohort was only 10 treated mice from one laboratory. It is preclinical evidence, not a patient-ready predictor.
Read and comment →Health & Life Sciences
Seventy-one Beijing university students rated an avatar-based CBT system easier to use than the same text chatbot after one session. The study measured experience—not symptom improvement, long-term safety or therapeutic effectiveness.
Read and comment →Government & Policy
Ministers meeting in Kyoto endorsed a shared vision for AI-enabled discovery, new research funding models and wider access to compute. The one-page declaration has no budget, deadline, enforcement mechanism or country-by-country delivery plan.
Read and comment →Education
Researchers tested a three-part AI competency scale with 470 nurse educators in Egypt. The structure held across two split samples, but it measures perceived skills, one subscale had borderline reliability and the authors warn against certification or staff appraisal.
Read and comment →Technology
A peer-reviewed benchmark put one 3-billion-parameter model through 90 simulated smart-home commands and synthetic preference changes. AdaHome beat reimplemented baselines, but no people, homes or physical devices were tested—and responses still took roughly 10 to 14 seconds.
Read and comment →AI Risks & Safety
A peer-reviewed video-language model matched or exceeded reported human scores on five-choice intention questions. When answer options disappeared, text-overlap scores fell below 20—leaving open-vocabulary claims, cultural bias and surveillance risk unresolved.
Read and comment →