Efficient Few-Shot Clinical Task Adaptation with Large Language Models

Few-shot learning has been studied to adapt models to tasks with very few samples. It holds profound significance, particularly in clinical tasks, due to the high annotation cost of medical images. Several works have explored few-shot learning on medical images, yet they still require a large number of medical images for pre-training models to gain domain-specific priors. Vision foundation models recently have achieved remarkable success in natural images. Hence, adapting rapidly advancing vision foundation models from natural images to few-shot clinical tasks holds great promise. MedFMC has recently organized a challenge to shed more light on this topic at NeurIPS 2023. In this work, we present our challenge solution. We observe that a simple variant of fine-tuning with partial freezing shows remarkable performance. Empirical evidence demonstrates that this approach could outperform various common fine-tuning methods under limited sample sizes. Additionally, we explore enhanced utilization of semantic supervision to boost performance. We propose a novel approach that contextualizes labels via large language models (LLMs). Our findings reveal that the context generated by LLMs significantly enhances the discrimination of semantic embeddings for similar categories, resulting in a notable performance improvement of 3%-5% in 1-shot settings compared to commonly employed one-hot labels and other semantic supervision methods. Our solution secures the 1st place in the MedFMC challenge.

翻译：小样本学习旨在研究如何利用极少样本将模型适配至特定任务。由于医学图像标注成本高昂，该方法在临床任务中具有深远意义。已有研究探索了医学图像的小样本学习，但仍需大量医学图像进行预训练以获取领域先验知识。近年来，视觉基础模型在自然图像领域取得了显著成功，因此将快速发展的视觉基础模型从自然图像适配至小样本临床任务具有巨大潜力。MedFMC近期在NeurIPS 2023上组织了相关挑战赛以深化该领域研究。本文呈现了我们的解决方案：观察到部分冻结参数的简单微调变体展现出卓越性能。实验证据表明，在样本量有限条件下，该方法可超越多种常见微调策略。此外，我们探索了语义监督的强化应用以提升性能，提出一种通过大语言模型（LLMs）对标签进行情境化处理的新方法。研究发现，LLMs生成的语义上下文能显著增强相似类别语义嵌入的区分度，在1-shot设置下相比常用独热编码及其他语义监督方法实现了3%-5%的性能提升。本方案最终在MedFMC挑战赛中荣获第一名。

相关内容

大语言模型

关注 67

大语言模型是基于海量文本数据训练的深度学习模型。它不仅能够生成自然语言文本，还能够深入理解文本含义，处理各种自然语言任务，如文本摘要、问答、翻译等。2023年，大语言模型及其在人工智能领域的应用已成为全球科技研究的热点，其在规模上的增长尤为引人注目，参数量已从最初的十几亿跃升到如今的一万亿。参数量的提升使得模型能够更加精细地捕捉人类语言微妙之处，更加深入地理解人类语言的复杂性。在过去的一年里，大语言模型在吸纳新知识、分解复杂任务以及图文对齐等多方面都有显著提升。随着技术的不断成熟，它将不断拓展其应用范围，为人类提供更加智能化和个性化的服务，进一步改善人们的生活和生产方式。

【NeurIPS2021】用于文本图表示学习的 GNN 嵌套 Transformer 模型：GraphFormers

专知会员服务

46+阅读 · 2021年11月24日

【亚马逊-WWW2020】不解析,生成!用于面向任务的语义分析的序列到序列体系结构，Don't Parse, Generate! A Sequence to Sequence Architecture for Task-Oriented Semantic Parsing

专知会员服务

15+阅读 · 2020年2月1日

FlowQA: Grasping Flow in History for Conversational Machine Comprehension

专知会员服务

35+阅读 · 2019年10月18日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

50+阅读 · 2019年10月17日