AI medical tools match or surpass doctors for advice - FT中文网
登录×
电子邮件/用户名
密码
记住我
请输入邮箱和密码进行绑定操作:
请输入手机号码,通过短信验证(目前仅支持中国大陆地区的手机号):
请您阅读我们的用户注册协议隐私权保护政策,点击下方按钮即视为您接受。
商业快报

AI medical tools match or surpass doctors for advice

Two health models matched or surpassed doctors across a range of diagnostic and treatment decisions, studies show
00:00

{"text":[[{"start":6.5,"text":"Two AI medical tools matched or surpassed doctors across a range of diagnostic and treatment decisions, in the latest sign that specialist health large language models are moving closer to demonstrating clinical value. "}],[{"start":19.6,"text":"Mira, developed by researchers in Germany, outperformed physicians in analyses of diseases including pancreatic cancer and pneumonia, while Google’s Amie produced more precise treatments and investigation plans than humans, according to results published in Nature on Wednesday.  "}],[{"start":36.1,"text":"The studies suggest specialist health AI tools can give better medical advice than general consumer AI models. But their inventors and independent experts warned that the tests were conducted in controlled simulations and did not mean the tools were ready for real-world clinical use. "}],[{"start":52.45,"text":"“We are getting a preview of how AI could transform medicine,” said Jakob Kather, whose academic group at TUD Dresden University of Technology and Heidelberg University co-developed Mira. "}],[{"start":65.35000000000001,"text":"“I see AI agents as being similar to the autopilot system in an airplane. These systems can support and relieve medical professionals by taking over routine tasks, but ultimate responsibility will always remain with the physicians,” he added."}],[{"start":80.9,"text":"Mira draws on patient data from an electronic health record system and can choose from more than 85,000 options, including ordering diagnostic tests, prescribing medication and scheduling procedures. The researchers tested it using information from more than 500 emergency department clinical cases, which were passed to it via chats with AI agents acting as patients."}],[{"start":102.80000000000001,"text":"Mira notched a diagnostic accuracy of 87.1 per cent across eight conditions including appendicitis and lung embolism, according to the Nature paper. That compared with 78.1 per cent achieved by a panel of six physicians across specialities."}],[{"start":119.60000000000001,"text":"Amie used Google’s Gemini AI model to respond to data given to it by actors role-playing patients. The scientists tested Amie against 21 primary care physicians on 100 multi-visit case scenarios, which were grounded in current UK clinical practice guidelines and drug recommendations."}],[{"start":137.85000000000002,"text":"Amie matched real physicians in patient management reasoning capabilities and aligned its plans more closely with the guidelines than they did, the scientists found. It outperformed human professionals’ reasoning on medication in difficult cases.  "}],[{"start":151.15000000000003,"text":"Both AI models had limitations, their inventors acknowledged. Mira still suggested “care that deviated from best practices” for a “small but non-zero” fraction of patients, the researchers said. "}],[{"start":163.35000000000002,"text":"The case information offered by the AI agents might have been “more structured than real speech of patients in emergency departments”, with fewer omissions and inconsistencies, they added."}],[{"start":173.95000000000002,"text":"The Amie study represented a “milestone” but neither the case mix nor the text-based patient scenarios were representative of a real clinical setting, the AI tool’s developers said. "}],[{"start":184.55,"text":"Amie exhibited “promising capabilities” but was “not ready for real-world translation” and required more work to curb problems such as latent reasoning errors, the scientists said."}],[{"start":195.65,"text":"Researchers not involved in the studies praised their rigour but echoed the caveat that both were based on carefully regulated simulations of patients. "}],[{"start":204.4,"text":"“This is some remove from the messy, complex, human world of everyday healthcare,” said Catherine Pope, professor of medical sociology at the University of Oxford."}],[{"start":215.85,"text":"Many of the reported instances of the AI models’ superiority reflected the “precision and completeness of plans” they offered, rather than showing “clear differences in clinical correctness”, said Julie Jacko, chaired professor of health informatics and data science at the University of Edinburgh."}],[{"start":234.6,"text":"“Overall, this is a strong experimental study and a meaningful step forward, but it demonstrates performance against a structured standard rather than fully capturing the complexity of real clinical decision-making,” Jacko said."}],[{"start":249,"text":"There was also a “question about where Amie’s advantage actually comes from”, given that on one benchmark general-purpose AI models had scored similarly, said Wei Xing, assistant professor in the University of Sheffield’s School of Mathematical and Physical Sciences."}],[{"start":264.35,"text":"“This suggests Amie’s edge may reflect the rapid general progress of AI models, more than the specific system built around it,” he said."}],[{"start":279.5,"text":""}]],"url":"https://audio.ftcn.net.cn/album/a_1781766171_6654.mp3"}

版权声明:本文版权归FT中文网所有,未经允许任何单位或个人不得转载,复制或以任何其他方式使用本文全部或部分,侵权必究。

能源危机加剧,燃料补贴拖累公共财政

过去四个月,出台燃料补贴以保护消费者免受价格飙升影响的国家数量增加了一倍多,各国财政压力进一步加重。

全球最火热股市为何反成韩国之累

韩国股价的剧烈波动正在损害国家形象。

必须采用不同方式监管金融领域的AI

在我们急于监管之前,我们应该思考如何不剥夺这项工具的益处,又管理好其造成伤害的风险。

他会成为印度尼西亚下一任总统吗?

德迪•穆利亚迪在社交媒体上的高度活跃,帮助他与选民建立起深厚联系。在许多人眼中,他是一个真正贴近民众的“自己人”。
10小时前

多边主义不是理想主义,而是现实必需

我们需要加强现有合作体系,而不是另起炉灶。

一周展望:日本央行担心通胀超调有没有道理?

投资者正评估日本央行将以多大力度继续加息,以及该行能否跑赢曲线,从而遏制通胀、支撑日元。
设置字号×
最小
较小
默认
较大
最大
分享×