GAMMT: 利用多个变换器生成多义模型 (GAMMT: Generative Ambiguity Modeling Using Multiple Transformers) - 专知论文

会员服务 ·

0

变换 · 数据生成过程 · 序列 · 不确定 · 概率 ·

2023 年 4 月 4 日

GAMMT: Generative Ambiguity Modeling Using Multiple Transformers

翻译：GAMMT: 利用多个变换器生成多义模型

from arxiv, 14 pages, 1 figure, 3 algorithms

We introduce a novel model called GAMMT (Generative Ambiguity Models using Multiple Transformers) for sequential data that is based on sets of probabilities. Unlike conventional models, our approach acknowledges that the data generation process of a sequence is not deterministic, but rather ambiguous and influenced by a set of probabilities. To capture this ambiguity, GAMMT employs multiple parallel transformers that are linked by a selection mechanism, allowing for the approximation of ambiguous probabilities. The generative nature of our approach also enables multiple representations of input tokens and sequences. While our models have not yet undergone experimental validation, we believe that our model has great potential to achieve high quality and diversity in modeling sequences with uncertain data generation processes.

翻译：我们介绍了一种新型的基于概率集的序列数据模型，称为 GAMMT （利用多个变换器生成多义模型），它与传统模型不同，因为我们的方法认识到一个序列的数据生成过程并非是确定的，而是存在歧义并受到一组概率的影响。为了捕捉这种不确定性，GAMMT 采用多个并行的变换器，通过选择机制相互链接，以近似不确定概率。我们方法的生成性质还允许输入标记和序列的多重表示。虽然我们的模型尚未经过实验验证，但我们相信我们的模型具有在对具有不确定数据生成过程的序列进行建模方面实现高质量和多样性的巨大潜力。

0

相关内容

【Hugging Face】指导文本生成与约束波束搜索🤗Transformers，Guiding Text Generation with Constrained Beam Search in 🤗 Transformers

【Hugging Face】指导文本生成与约束波束搜索🤗Transformers，Guiding Text Generation with Constrained Beam Search in 🤗 Transformers

专知会员服务

22+阅读 · 2022年3月18日

最新《Transformers模型》教程，64页ppt

最新《Transformers模型》教程，64页ppt

专知会员服务

325+阅读 · 2020年11月26日

【文本生成现代方法】Modern Methods for Text Generation

【文本生成现代方法】Modern Methods for Text Generation

专知会员服务

44+阅读 · 2020年9月11日

【上海交大】可解释CNN的对象分类，Interpretable CNNs for Object Classification

专知会员服务

54+阅读 · 2020年3月14日

Transformer文本分类代码

Transformer文本分类代码

专知会员服务

118+阅读 · 2020年2月3日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

逆强化学习-学习人先验的动机

逆强化学习-学习人先验的动机

CreateAMind

16+阅读 · 2019年1月18日

无监督元学习表示学习

无监督元学习表示学习

CreateAMind

27+阅读 · 2019年1月4日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

可解释的CNN

可解释的CNN

CreateAMind

17+阅读 · 2017年10月5日

爆破发展方程的控制理论

国家自然科学基金

1+阅读 · 2014年12月31日

随机变量结构的模型论

国家自然科学基金

0+阅读 · 2013年12月31日

功率变换器非线性不稳定行为的washout滤波器控制方法

国家自然科学基金

0+阅读 · 2012年12月31日

对象模型上交互式修复生成技术研究

国家自然科学基金

0+阅读 · 2012年12月31日

mTOR激活对吗啡耐受的调控及其分子机制

国家自然科学基金

0+阅读 · 2011年12月31日

Uni-ControlNet: All-in-One Control to Text-to-Image Diffusion Models

Arxiv

0+阅读 · 2023年5月25日

Trans-Dimensional Generative Modeling via Jump Diffusion Models

Trans-Dimensional Generative Modeling via Jump Diffusion Models

Arxiv

0+阅读 · 2023年5月25日

Latent Diffusion Model Based Foley Sound Generation System For DCASE Challenge 2023 Task 7

Arxiv

0+阅读 · 2023年5月25日

Dynamic Context Pruning for Efficient and Interpretable Autoregressive Transformers

Arxiv

0+阅读 · 2023年5月25日

Bias-to-Text: Debiasing Unknown Visual Biases through Language Interpretation

Arxiv

0+阅读 · 2023年5月24日

VIP会员

文章信息

相关主题

数据生成过程

相关VIP内容

【Hugging Face】指导文本生成与约束波束搜索🤗Transformers，Guiding Text Generation with Constrained Beam Search in 🤗 Transformers

【Hugging Face】指导文本生成与约束波束搜索🤗Transformers，Guiding Text Generation with Constrained Beam Search in 🤗 Transformers

专知会员服务

22+阅读 · 2022年3月18日

最新《Transformers模型》教程，64页ppt

最新《Transformers模型》教程，64页ppt

专知会员服务

325+阅读 · 2020年11月26日

【文本生成现代方法】Modern Methods for Text Generation

【文本生成现代方法】Modern Methods for Text Generation

专知会员服务

44+阅读 · 2020年9月11日

【上海交大】可解释CNN的对象分类，Interpretable CNNs for Object Classification

专知会员服务

54+阅读 · 2020年3月14日

Transformer文本分类代码

Transformer文本分类代码

专知会员服务

118+阅读 · 2020年2月3日

热门VIP内容

开通专知VIP会员享更多权益服务

【MIT博士论文】弱监督学习：理论、方法与应用

Andrej Karpathy：2025 年 LLM 年度回顾（2025 LLM Year in Review）

锚定情报：合成欺骗时代的地面真相

NeurIPS 2025 | NMKE：基于神经元归因与动态稀疏掩码的终身知识编辑

相关资讯

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

逆强化学习-学习人先验的动机

逆强化学习-学习人先验的动机

CreateAMind

16+阅读 · 2019年1月18日

无监督元学习表示学习

无监督元学习表示学习

CreateAMind

27+阅读 · 2019年1月4日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

可解释的CNN

可解释的CNN

CreateAMind

17+阅读 · 2017年10月5日

相关论文

Uni-ControlNet: All-in-One Control to Text-to-Image Diffusion Models

Arxiv

0+阅读 · 2023年5月25日

Trans-Dimensional Generative Modeling via Jump Diffusion Models

Trans-Dimensional Generative Modeling via Jump Diffusion Models

Arxiv

0+阅读 · 2023年5月25日

Latent Diffusion Model Based Foley Sound Generation System For DCASE Challenge 2023 Task 7

Arxiv

0+阅读 · 2023年5月25日

Dynamic Context Pruning for Efficient and Interpretable Autoregressive Transformers

Arxiv

0+阅读 · 2023年5月25日

Bias-to-Text: Debiasing Unknown Visual Biases through Language Interpretation

Arxiv

0+阅读 · 2023年5月24日

相关基金

爆破发展方程的控制理论

国家自然科学基金

1+阅读 · 2014年12月31日

随机变量结构的模型论

国家自然科学基金

0+阅读 · 2013年12月31日

功率变换器非线性不稳定行为的washout滤波器控制方法

国家自然科学基金

0+阅读 · 2012年12月31日

对象模型上交互式修复生成技术研究

国家自然科学基金

0+阅读 · 2012年12月31日

mTOR激活对吗啡耐受的调控及其分子机制

国家自然科学基金

0+阅读 · 2011年12月31日

微信扫码咨询专知VIP会员