Daily Tech Briefing
AI 科技速览
每天 5 分钟内学习 AI。获取最新的人工智能新闻,理解其重要性,并学习如何将其应用于您的工作。
MIT News AI · 2026/7/29 14:00:00
How a medical database developed at MIT evolved into a global standard of data-sharing
AI 中文解读
核心亮点:上世纪70年代MIT科学家手工复制的心电图磁带,如今竟成为全球医学数据共享的“开山鼻祖”,催生了影响世界的PhysioNet平台。
通俗解读:以前医生做研究,数据都锁在自己抽屉里,想对比不同医院的结果难如登天。1975年,MIT的一群研究人员决定改变这局面:他们自己造电脑、一盘盘复制磁带,花了五年时间把心脏电信号数据整理出来,免费寄给同行。没想到索要的人越来越多,最后这批数据发展成全球公开数据库PhysioNet,如今上面有几百个临床数据集,去年被超过1.5万篇论文引用,180多个国家的用户都在用。
实际影响:你可能觉得这事离自己很远,但PhysioNet上的数据帮助医生更精准地诊断心脏病、改进急救设备算法,甚至训练AI监测重症病人。以后你在医院做心电图,医生用的分析软件很可能就是基于这些公开数据训练出来的。当年那盘通过邮局寄出的磁带,悄悄推动了整个医疗行业的数据共享习惯,让今天的医学研究不再“闭门造车”,最终受益的是每个普通患者。
<p>Before the advancement of scientific data storage and collaboration via the cloud, medical investigators seeking health research breakthroughs had to overcome significant obstacles to collaboration and key clinical data gathering. </p><p>Data were siloed and difficult to distribute, so those looking to undertake research had no option but to gather them themselves. This not only made research more expensive, but it was challenging to compare findings across datasets. </p><p>In 1975, researchers studying arrhythmias at MIT and Boston’s Beth Israel Hospital envisioned another way: the team began collecting and digitizing electrocardiogram recordings with the intention of not only studying them, but of also making them available to the wider research community.</p><p>The team built their own computers for the process, painstakingly duplicated tapes one by one, and created more than 100,000 annotations for the recordings. The process took years, but by summer 1980, the tapes were finally ready. The team initially thought their tool would reach fewer than a dozen academic and industry groups. But interest kept pouring in. Over the next decade, they went on to mail about 100 copies. </p><p>The data eventually became the first database of the global platform <a href="https://physionet.org/" target="_blank">PhysioNet</a> — founded in 1999 at the Harvard-MIT program in Health Sciences and Technology — as a clinical data repository for complex physiological signals. </p><p>At the time, that type of data-sharing, which may seem like the default today, was a near-revolutionary idea. PhysioNet’s “founding was incredibly visionary,” says <a href="https://imes.mit.edu/people/heldt-thomas">Thomas Heldt</a>, Richard J. Cohen (1976) Professor in Medicine and Biomedical Physics, associate director of MIT’s Institute for Medical Engineering and Science, and the senior author of a <a href="https://www.nature.com/articles/s44360-026-00096-z.epdf?sharing_token=JCCZrQYhSx2k8UGOFhiGu9RgN0jAjWel9jnR3ZoTv0OYh1eA8TCaCZrKOgyrlAWDjgy7w_drrFSNFpUt0FQCSWnyEBf-N24xBfnUgrckU7Nvv5kveP7nAH8GUBLcCkBB-Zk5QeCIQo6BrBQbL5w9kiBhn5cI6ISA33jSrZQOHZg%3D" target="_blank">recent paper in <em>Nature Health</em></a> examining the platform’s impact. </p><p>Eventually, those magnetic tapes sent through the mail became burned CD-ROMs, which then evolved into FTP servers hosted on the newly minted internet. Today, as PhysioNet looks back at over 25 years of operation, the platform hosts hundreds of databases, and has become one of the most comprehensive biomedical and clinical data repositories in existence. Last year, more than 15,000 scientific publications cited PhysioNet, and users from more than 180 countries have registered on the platform. It is widely used by researchers, manufacturers, and clinical decision-makers.</p><p>“The research impact is truly significant,” says Heldt, who is also a professor in the MIT Department of Electrical Engineering and Computer Science and a principal investigator at the Research Laboratory of Electronics, “and quite humbling.” </p><p>“It is really beautiful to see that such a vision has proven right and so enabling for so many people.”</p><p><strong>Setting a standard </strong></p><p>Around 2009, a PhD student named Tom Pollard was conducting research on critically ill patients at one of London’s leading hospital systems. Although the hospital generated large volumes of valuable clinical data, the infrastructure and processes needed to curate and support their wider research use were still developing. </p><p>“Hospital data were collected primarily to support immediate patient care, with less attention given to how they might be curated and reused for research,” says Pollard, now a research scientist at MIT’s <a href="https://lcp.mit.edu/">Laboratory for Computational Physiology (LCP)</a>, technical director of PhysioNet, and the lead author on the <em>Nature Health</em> paper. </p><p>The problem was not simply privacy. Hospital inform
分享
阅读原文 ↗