Chubby♨️@kimmonismus
36AI 编辑部评分,满分 100

Google 用模拟住院实习训练 Gemini,临床问诊能力显著提升

2026-08-12 21:32· 3小时前
AI 导读

Google 通过 ResidencyRL 让 Gemini 3.5 Flash 在 49,870 次模拟远程诊疗中训练,AI 患者会隐瞒症状、抗拒建议,迫使模型主动收集信息。训练后对抗条件下诊断准确率从 81% 升至 88%,漏诊红旗信号下降 31%;97 例盲评中,临床医生在 87.6% 的对比中更偏好训练后模型,且提升迁移至未见肿瘤病例。

Thats the reaserach i love to see: Google put Gemini through a simulated medical residency, and it became markedly better at conducting clinical consultations.

ResidencyRL trained Gemini 3.5 Flash across 49,870 simulated telehealth encounters covering 81 conditions.

The AI patients hid symptoms, resisted advice and requested inappropriate treatments, forcing the model to gather information over conversations of up to 60 turns.

After training, diagnostic accuracy under adversarial conditions rose from 81% to 88%, while missed red flags fell by 31%. In a blinded evaluation of 97 cases, clinicians preferred the trained agent over the base model in 87.6% of comparisons.

The improvement also transferred to unseen oncology cases and external benchmarks.

AI health agents incoming <3

Samuel SchmidgallDoctors don't become doctors from textbooks. They become doctors through residency, requiring years of practice across thousands of patient encounters. Can AI l...

来源:Chubby♨️ · x.com

Google 用模拟住院实习训练 Gemini,临床问诊能力显著提升

Chubby♨️ · @kimmonismus · X·2026-08-12 21:32·3小时前
AI 导读

Google 通过 ResidencyRL 让 Gemini 3.5 Flash 在 49,870 次模拟远程诊疗中训练,AI 患者会隐瞒症状、抗拒建议,迫使模型主动收集信息。训练后对抗条件下诊断准确率从 81% 升至 88%,漏诊红旗信号下降 31%;97 例盲评中,临床医生在 87.6% 的对比中更偏好训练后模型,且提升迁移至未见肿瘤病例。

Thats the reaserach i love to see: Google put Gemini through a simulated medical residency, and it became markedly better at conducting clinical consultations.

ResidencyRL trained Gemini 3.5 Flash across 49,870 simulated telehealth encounters covering 81 conditions.

The AI patients hid symptoms, resisted advice and requested inappropriate treatments, forcing the model to gather information over conversations of up to 60 turns.

After training, diagnostic accuracy under adversarial conditions rose from 81% to 88%, while missed red flags fell by 31%. In a blinded evaluation of 97 cases, clinicians preferred the trained agent over the base model in 87.6% of comparisons.

The improvement also transferred to unseen oncology cases and external benchmarks.

AI health agents incoming <3

Samuel SchmidgallDoctors don't become doctors from textbooks. They become doctors through residency, requiring years of practice across thousands of patient encounters. Can AI l...

来源:Chubby♨️· x.com