EHRSQL: A Practical Text-to-SQL Benchmark for Electronic Health Records

We present a new text-to-SQL dataset for electronic health records (EHRs). The utterances were collected from 222 hospital staff members, including physicians, nurses, and insurance review and health records teams. To construct the QA dataset on structured EHR data, we conducted a poll at a university hospital and used the responses to create seed questions. We then manually linked these questions to two open-source EHR databases, MIMIC-III and eICU, and included various time expressions and held-out unanswerable questions in the dataset, which were also collected from the poll. Our dataset poses a unique set of challenges: the model needs to 1) generate SQL queries that reflect a wide range of needs in the hospital, including simple retrieval and complex operations such as calculating survival rate, 2) understand various time expressions to answer time-sensitive questions in healthcare, and 3) distinguish whether a given question is answerable or unanswerable. We believe our dataset, EHRSQL, can serve as a practical benchmark for developing and assessing QA models on structured EHR data and take a step further towards bridging the gap between text-to-SQL research and its real-life deployment in healthcare. EHRSQL is available at https://github.com/glee4810/EHRSQL.

翻译：我们提出一个新的面向电子健康记录（EHR）的文本到SQL数据集。话语数据来自222名医院工作人员，包括医生、护士以及保险审核和健康记录团队。为了构建结构化EHR数据上的问答数据集，我们在大学医院进行了一项民意调查，并利用调查回复生成种子问题。随后，我们手动将这些问题关联到两个开源EHR数据库（MIMIC-III和eICU），并在数据集中纳入多种时间表达式及从调查中收集的不可回答问题。我们的数据集带来一系列独特挑战：模型需要1）生成反映医院中广泛需求的SQL查询，包括简单检索和计算存活率等复杂操作；2）理解多种时间表达式以回答医疗领域中的时间敏感问题；3）区分给定问题是否可回答。我们相信，本数据集EHRSQL可作为在结构化EHR数据上开发与评估问答模型的实用基准，并进一步弥合文本到SQL研究与其在医疗领域实际部署之间的差距。EHRSQL可通过https://github.com/glee4810/EHRSQL获取。

相关内容

数据集

关注 88

数据集，又称为资料集、数据集合或资料集合，是一种由数据所组成的集合。
Data set（或dataset）是一个数据的集合，通常以表格形式出现。每一列代表一个特定变量。每一行都对应于某一成员的数据集的问题。它列出的价值观为每一个变量，如身高和体重的一个物体或价值的随机数。每个数值被称为数据资料。对应于行数，该数据集的数据可能包括一个或多个成员。

Linux导论，Introduction to Linux，96页ppt

专知会员服务

82+阅读 · 2020年7月26日

【跨语言BERT模型大集合】Transfer learning is increasingly going multilingual with language-specific BERT models

专知会员服务

54+阅读 · 2020年1月30日

FlowQA: Grasping Flow in History for Conversational Machine Comprehension

专知会员服务

34+阅读 · 2019年10月18日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

50+阅读 · 2019年10月17日