使用向导想象扩大小型数据集 (Expanding Small-Scale Datasets with Guided Imagination) - 专知论文

会员服务 ·

0

GIF · INFORMS · MoDELS · 数据集 · 模型评估 ·

2022 年 12 月 8 日

Expanding Small-Scale Datasets with Guided Imagination

翻译：使用向导想象扩大小型数据集

Yifan Zhang,Daquan Zhou,Bryan Hooi,Kai Wang,Jiashi Feng

The power of Deep Neural Networks (DNNs) depends heavily on the training data quantity, quality and diversity. However, in many real scenarios, it is costly and time-consuming to collect and annotate large-scale data. This has severely hindered the application of DNNs. To address this challenge, we explore a new task of dataset expansion, which seeks to automatically create new labeled samples to expand a small dataset. To this end, we present a Guided Imagination Framework (GIF) that leverages the recently developed big generative models (e.g., DALL-E2) and reconstruction models (e.g., MAE) to "imagine" and create informative new data from seed data to expand small datasets. Specifically, GIF conducts imagination by optimizing the latent features of seed data in a semantically meaningful space, which are fed into the generative models to generate photo-realistic images with new contents. For guiding the imagination towards creating samples useful for model training, we exploit the zero-shot recognition ability of CLIP and introduce three criteria to encourage informative sample generation, i.e., prediction consistency, entropy maximization and diversity promotion. With these essential criteria as guidance, GIF works well for expanding datasets in different domains, leading to 29.9% accuracy gain on average over six natural image datasets, and 12.3% accuracy gain on average over three medical image datasets. The source code will be released: \url{https://github.com/Vanint/DatasetExpansion}.

翻译：深神经网络(DNNS)的力量在很大程度上取决于培训数据的数量、质量和多样性。然而,在许多真实的情景中,收集和批注大型数据既费钱又费时。这严重阻碍了DNNS的应用。为了应对这一挑战,我们探索了扩大数据集的新任务,即自动创建标签标签的新样本,以扩大小型数据集。为此,我们提出了一个指导想象框架(GIF),利用最近开发的大型基因化模型(例如DALL-E2)和重建模型(例如MAE)来“想象”和从种子数据中创建信息性的新数据以扩大小型数据集。具体地说,GIF通过在具有语义意义的空间优化种子数据的潜在特征以生成一个带有新内容的图像。为了引导想象力,我们将CLIP的零光识别能力引入三个标准,鼓励信息性样本生成, i. IMED 的准确度, 将GIF 数据流流流化为三个基本版本, 将数据升级为基础版本。

0

相关内容

GIF

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

专知会员服务

60+阅读 · 2022年4月22日

50+篇《神经架构搜索NAS》2020论文合集

专知会员服务

61+阅读 · 2020年3月19日

【新书】数字图像(影像)处理手第二版，2176pdf，Mathematical Methods in Imaging

【新书】数字图像(影像)处理手第二版，2176pdf，Mathematical Methods in Imaging

专知会员服务

93+阅读 · 2020年2月12日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

专知会员服务

160+阅读 · 2019年10月12日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

专知会员服务

65+阅读 · 2019年10月9日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

105+阅读 · 2019年10月9日

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

专知会员服务

41+阅读 · 2019年10月9日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium8

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium8

中国图象图形学学会CSIG

0+阅读 · 2021年11月16日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium6

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium6

中国图象图形学学会CSIG

2+阅读 · 2021年11月12日

【ICIG2021】Latest News & Announcements of the Plenary Talk1

【ICIG2021】Latest News & Announcements of the Plenary Talk1

中国图象图形学学会CSIG

0+阅读 · 2021年11月1日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

强化学习的Unsupervised Meta-Learning

强化学习的Unsupervised Meta-Learning

CreateAMind

18+阅读 · 2019年1月7日

无监督元学习表示学习

无监督元学习表示学习

CreateAMind

27+阅读 · 2019年1月4日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

最新5篇生成对抗网络相关论文推荐—FusedGAN、DeblurGAN、AdvGAN、CipherGAN、MMD GANS

最新5篇生成对抗网络相关论文推荐—FusedGAN、DeblurGAN、AdvGAN、CipherGAN、MMD GANS

专知

23+阅读 · 2018年1月18日

具有抗氧化功能的壳聚糖衍生物细胞支架的研究

国家自然科学基金

0+阅读 · 2015年12月31日

内质网应激IRE1－XBP1S通路在高糖引起肾脏及系膜细胞发生氧化应激及损伤中的机制研究

国家自然科学基金

1+阅读 · 2014年12月31日

基于信号放大技术的表面增强拉曼成像分析法用于肿瘤细胞检测及单细胞分析

国家自然科学基金

0+阅读 · 2013年12月31日

基于酶循环放大表面增强拉曼检测人乳腺癌细胞及细胞表面活性物质

国家自然科学基金

0+阅读 · 2013年12月31日

(规范)超引力黑洞与黑环

国家自然科学基金

0+阅读 · 2012年12月31日

泛素化蛋白酶A20抑制caspase-8活化在脑肿瘤干细胞发生TRAIL抵抗中的作用机制

国家自然科学基金

0+阅读 · 2012年12月31日

PGRMC1介导孕酮抗氧化的作用及其机制的研究

国家自然科学基金

0+阅读 · 2012年12月31日

RhoA/ROCK促巨噬细胞活化在糖尿病动脉粥样硬化中的作用及其机制研究

国家自然科学基金

0+阅读 · 2009年12月31日

长寿基因SIRT1在细胞衰老过程中的转录调控研究

国家自然科学基金

0+阅读 · 2009年12月31日

hTERT转染雪旺细胞与复合FK506壳聚糖支架构建人工神经的研究

国家自然科学基金

0+阅读 · 2009年12月31日

Zero-shot Clarifying Question Generation for Conversational Search

Zero-shot Clarifying Question Generation for Conversational Search

Arxiv

0+阅读 · 2023年2月10日

Designing Robust Transformers using Robust Kernel Density Estimation

Arxiv

0+阅读 · 2023年2月9日

Is This Loss Informative? Speeding Up Textual Inversion with Deterministic Objective Evaluation

Is This Loss Informative? Speeding Up Textual Inversion with Deterministic Objective Evaluation

Arxiv

0+阅读 · 2023年2月9日

Cooperative Open-ended Learning Framework for Zero-shot Coordination

Arxiv

0+阅读 · 2023年2月9日

ChatGPT and Software Testing Education: Promises & Perils

Arxiv

0+阅读 · 2023年2月8日

Multi-sensor large-scale dataset for multi-view 3D reconstruction

Arxiv

0+阅读 · 2023年2月8日

Expanding Accurate Person Recognition to New Altitudes and Ranges: The BRIAR Dataset

Expanding Accurate Person Recognition to New Altitudes and Ranges: The BRIAR Dataset

Arxiv

16+阅读 · 2022年11月3日

Learning to Learn and Predict: A Meta-Learning Approach for Multi-Label Classification

Learning to Learn and Predict: A Meta-Learning Approach for Multi-Label Classification

Arxiv

17+阅读 · 2019年9月9日

Low-Shot Learning from Imaginary Data

Arxiv

15+阅读 · 2018年4月3日

DOTA: A Large-scale Dataset for Object Detection in Aerial Images

Arxiv

19+阅读 · 2018年1月27日

VIP会员

文章信息

相关主题

相关VIP内容

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

专知会员服务

60+阅读 · 2022年4月22日

50+篇《神经架构搜索NAS》2020论文合集

专知会员服务

61+阅读 · 2020年3月19日

【新书】数字图像(影像)处理手第二版，2176pdf，Mathematical Methods in Imaging

【新书】数字图像(影像)处理手第二版，2176pdf，Mathematical Methods in Imaging

专知会员服务

93+阅读 · 2020年2月12日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

专知会员服务

160+阅读 · 2019年10月12日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

专知会员服务

65+阅读 · 2019年10月9日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

105+阅读 · 2019年10月9日

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

专知会员服务

41+阅读 · 2019年10月9日

热门VIP内容

开通专知VIP会员享更多权益服务

操作系统智能体：基于多模态大模型（MLLM）的通用计算设备智能体综述

《美国太空军系统全生命周期建模、仿真与分析效能提升方案》最新84页报告

【博士论文】推进数据高效的深度学习：非参数 Transformer、主动测试与上下文学习

自主人工智能：未来战争是否将是自主化的？

相关资讯

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium8

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium8

中国图象图形学学会CSIG

0+阅读 · 2021年11月16日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium6

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium6

中国图象图形学学会CSIG

2+阅读 · 2021年11月12日

【ICIG2021】Latest News & Announcements of the Plenary Talk1

【ICIG2021】Latest News & Announcements of the Plenary Talk1

中国图象图形学学会CSIG

0+阅读 · 2021年11月1日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

强化学习的Unsupervised Meta-Learning

强化学习的Unsupervised Meta-Learning

CreateAMind

18+阅读 · 2019年1月7日

无监督元学习表示学习

无监督元学习表示学习

CreateAMind

27+阅读 · 2019年1月4日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

最新5篇生成对抗网络相关论文推荐—FusedGAN、DeblurGAN、AdvGAN、CipherGAN、MMD GANS

最新5篇生成对抗网络相关论文推荐—FusedGAN、DeblurGAN、AdvGAN、CipherGAN、MMD GANS

专知

23+阅读 · 2018年1月18日

相关论文

Zero-shot Clarifying Question Generation for Conversational Search

Zero-shot Clarifying Question Generation for Conversational Search

Arxiv

0+阅读 · 2023年2月10日

Designing Robust Transformers using Robust Kernel Density Estimation

Arxiv

0+阅读 · 2023年2月9日

Is This Loss Informative? Speeding Up Textual Inversion with Deterministic Objective Evaluation

Is This Loss Informative? Speeding Up Textual Inversion with Deterministic Objective Evaluation

Arxiv

0+阅读 · 2023年2月9日

Cooperative Open-ended Learning Framework for Zero-shot Coordination

Arxiv

0+阅读 · 2023年2月9日

ChatGPT and Software Testing Education: Promises & Perils

Arxiv

0+阅读 · 2023年2月8日

Multi-sensor large-scale dataset for multi-view 3D reconstruction

Arxiv

0+阅读 · 2023年2月8日

Expanding Accurate Person Recognition to New Altitudes and Ranges: The BRIAR Dataset

Expanding Accurate Person Recognition to New Altitudes and Ranges: The BRIAR Dataset

Arxiv

16+阅读 · 2022年11月3日

Learning to Learn and Predict: A Meta-Learning Approach for Multi-Label Classification

Learning to Learn and Predict: A Meta-Learning Approach for Multi-Label Classification

Arxiv

17+阅读 · 2019年9月9日

Low-Shot Learning from Imaginary Data

Arxiv

15+阅读 · 2018年4月3日

DOTA: A Large-scale Dataset for Object Detection in Aerial Images

Arxiv

19+阅读 · 2018年1月27日

相关基金

具有抗氧化功能的壳聚糖衍生物细胞支架的研究

国家自然科学基金

0+阅读 · 2015年12月31日

内质网应激IRE1－XBP1S通路在高糖引起肾脏及系膜细胞发生氧化应激及损伤中的机制研究

国家自然科学基金

1+阅读 · 2014年12月31日

基于信号放大技术的表面增强拉曼成像分析法用于肿瘤细胞检测及单细胞分析

国家自然科学基金

0+阅读 · 2013年12月31日

基于酶循环放大表面增强拉曼检测人乳腺癌细胞及细胞表面活性物质

国家自然科学基金

0+阅读 · 2013年12月31日

(规范)超引力黑洞与黑环

国家自然科学基金

0+阅读 · 2012年12月31日

泛素化蛋白酶A20抑制caspase-8活化在脑肿瘤干细胞发生TRAIL抵抗中的作用机制

国家自然科学基金

0+阅读 · 2012年12月31日

PGRMC1介导孕酮抗氧化的作用及其机制的研究

国家自然科学基金

0+阅读 · 2012年12月31日

RhoA/ROCK促巨噬细胞活化在糖尿病动脉粥样硬化中的作用及其机制研究

国家自然科学基金

0+阅读 · 2009年12月31日

长寿基因SIRT1在细胞衰老过程中的转录调控研究

国家自然科学基金

0+阅读 · 2009年12月31日

hTERT转染雪旺细胞与复合FK506壳聚糖支架构建人工神经的研究

国家自然科学基金

0+阅读 · 2009年12月31日

微信扫码咨询专知VIP会员