基于影响函数的高效推理数据选择方法 (Influence Functions for Efficient Data Selection in Reasoning) - 专知论文

会员服务 ·

0

影响函数 · 数据选择 · CoT · 微调 · 高效推理 ·

Influence Functions for Efficient Data Selection in Reasoning

翻译：基于影响函数的高效推理数据选择方法

Prateek Humane,Paolo Cudrano,Daniel Z. Kaplan,Matteo Matteucci,Supriyo Chakraborty,Irina Rish

from arxiv, 4 pages, 2 figures; added link to codebase

Fine-tuning large language models (LLMs) on chain-of-thought (CoT) data shows that a small amount of high-quality data can outperform massive datasets. Yet, what constitutes "quality" remains ill-defined. Existing reasoning methods rely on indirect heuristics such as problem difficulty or trace length, while instruction-tuning has explored a broader range of automated selection strategies, but rarely in the context of reasoning. We propose to define reasoning data quality using influence functions, which measure the causal effect of individual CoT examples on downstream accuracy, and introduce influence-based pruning, which consistently outperforms perplexity and embedding-based baselines on math reasoning within a model family.

翻译：在思维链（CoT）数据上微调大型语言模型（LLMs）的研究表明，少量高质量数据的效果可超越海量数据集。然而，“质量”的具体定义仍不明确。现有推理方法依赖问题难度或推理轨迹长度等间接启发式指标，而指令微调虽探索了更广泛的自动选择策略，却极少在推理任务背景下应用。我们提出利用影响函数来定义推理数据的质量——该函数可量化单个CoT样本对下游准确率的因果效应，并引入基于影响的剪枝方法。实验表明，在数学推理任务中，该方法在同一模型家族内持续优于基于困惑度与嵌入向量的基线模型。

0

相关内容

影响函数

【WWW2024】在MOOCs中利用对比学习建模平衡显式与隐式关系以推荐知识概念

【WWW2024】在MOOCs中利用对比学习建模平衡显式与隐式关系以推荐知识概念

专知会员服务

14+阅读 · 2024年2月14日

【CVPR 2022】基于实例深度估计的统一深度感知全景分割 PanopticDepth: Per-Instance Depth Estimation for Unified Depth-Aware Panoptic Segmentation

【CVPR 2022】基于实例深度估计的统一深度感知全景分割 PanopticDepth: Per-Instance Depth Estimation for Unified Depth-Aware Panoptic Segmentation

专知会员服务

18+阅读 · 2022年3月19日

【CVPR2022】MSDN: 零样本学习的互语义蒸馏网络

【CVPR2022】MSDN: 零样本学习的互语义蒸馏网络

专知会员服务

21+阅读 · 2022年3月8日

【ACL2021】基于隐含结构推理网络的事件因果关系识别

专知会员服务

52+阅读 · 2021年8月13日

语义相似性算法演化论文，29页pdf，Evolution of Semantic Similarity - A Survey

语义相似性算法演化论文，29页pdf，Evolution of Semantic Similarity - A Survey

专知会员服务

44+阅读 · 2020年4月30日

【AAAI2021】知识图谱增强的预训练模型的生成式常识推理

【AAAI2021】知识图谱增强的预训练模型的生成式常识推理

专知

29+阅读 · 2021年1月25日

【KDD2020-Tutorial】因果推理与稳定学习，Causal Inference and Stable Learning

【KDD2020-Tutorial】因果推理与稳定学习，Causal Inference and Stable Learning

专知

11+阅读 · 2020年8月28日

【阿里巴巴-WWW2020】对抗性多模态表示学习的点击率预测，Adversarial Multimodal RL

【阿里巴巴-WWW2020】对抗性多模态表示学习的点击率预测，Adversarial Multimodal RL

专知

11+阅读 · 2020年3月17日

论文浅尝 | 当知识图谱遇上零样本学习——零样本学习综述

论文浅尝 | 当知识图谱遇上零样本学习——零样本学习综述

开放知识图谱

22+阅读 · 2018年9月26日

Spark机器学习：矩阵及推荐算法

Spark机器学习：矩阵及推荐算法

LibRec智能推荐

16+阅读 · 2017年8月3日

含非正态及缺失数据的结构方程模型分析

国家自然科学基金

0+阅读 · 2015年12月31日

线性时序关系下推理的概率计量化模型

国家自然科学基金

0+阅读 · 2014年12月31日

一般误差分布下若干半参数模型的复合分位数方法

国家自然科学基金

0+阅读 · 2014年12月31日

变换结构方程模型的非参数贝叶斯分析

国家自然科学基金

4+阅读 · 2014年12月31日

复杂数据下含指标项半参数模型结构的统计推断及应用

国家自然科学基金

0+阅读 · 2014年12月31日

Improving Semantic Uncertainty Quantification in LVLMs with Semantic Gaussian Processes

Arxiv

0+阅读 · 12月16日

Weighted Conformal Prediction for Survival Analysis under Covariate Shift

Arxiv

0+阅读 · 12月3日

Joint Evaluation of Answer and Reasoning Consistency for Hallucination Detection in Large Reasoning Models

Arxiv

0+阅读 · 11月15日

A Unified Framework for Inference with General Missingness Patterns and Machine Learning Imputation

Arxiv

0+阅读 · 11月11日

A Unified Framework for Inference with General Missingness Patterns and Machine Learning Imputation

Arxiv

0+阅读 · 11月7日

VIP会员

文章信息

相关主题

相关VIP内容

【WWW2024】在MOOCs中利用对比学习建模平衡显式与隐式关系以推荐知识概念

【WWW2024】在MOOCs中利用对比学习建模平衡显式与隐式关系以推荐知识概念

专知会员服务

14+阅读 · 2024年2月14日

【CVPR 2022】基于实例深度估计的统一深度感知全景分割 PanopticDepth: Per-Instance Depth Estimation for Unified Depth-Aware Panoptic Segmentation

【CVPR 2022】基于实例深度估计的统一深度感知全景分割 PanopticDepth: Per-Instance Depth Estimation for Unified Depth-Aware Panoptic Segmentation

专知会员服务

18+阅读 · 2022年3月19日

【CVPR2022】MSDN: 零样本学习的互语义蒸馏网络

【CVPR2022】MSDN: 零样本学习的互语义蒸馏网络

专知会员服务

21+阅读 · 2022年3月8日

【ACL2021】基于隐含结构推理网络的事件因果关系识别

专知会员服务

52+阅读 · 2021年8月13日

语义相似性算法演化论文，29页pdf，Evolution of Semantic Similarity - A Survey

语义相似性算法演化论文，29页pdf，Evolution of Semantic Similarity - A Survey

专知会员服务

44+阅读 · 2020年4月30日

热门VIP内容

开通专知VIP会员享更多权益服务

大语言模型中的事件抽取：方法、模态与未来展望的全面综述

美海军作战管理系统：变革战场空间的二十年

【MIT博士论文】以语言为中心的医学影像理解

俄罗斯“沙希德”/“天竺葵”攻击无人机

相关资讯

【AAAI2021】知识图谱增强的预训练模型的生成式常识推理

【AAAI2021】知识图谱增强的预训练模型的生成式常识推理

专知

29+阅读 · 2021年1月25日

【KDD2020-Tutorial】因果推理与稳定学习，Causal Inference and Stable Learning

【KDD2020-Tutorial】因果推理与稳定学习，Causal Inference and Stable Learning

专知

11+阅读 · 2020年8月28日

【阿里巴巴-WWW2020】对抗性多模态表示学习的点击率预测，Adversarial Multimodal RL

【阿里巴巴-WWW2020】对抗性多模态表示学习的点击率预测，Adversarial Multimodal RL

专知

11+阅读 · 2020年3月17日

论文浅尝 | 当知识图谱遇上零样本学习——零样本学习综述

论文浅尝 | 当知识图谱遇上零样本学习——零样本学习综述

开放知识图谱

22+阅读 · 2018年9月26日

Spark机器学习：矩阵及推荐算法

Spark机器学习：矩阵及推荐算法

LibRec智能推荐

16+阅读 · 2017年8月3日

相关论文

Improving Semantic Uncertainty Quantification in LVLMs with Semantic Gaussian Processes

Arxiv

0+阅读 · 12月16日

Weighted Conformal Prediction for Survival Analysis under Covariate Shift

Arxiv

0+阅读 · 12月3日

Joint Evaluation of Answer and Reasoning Consistency for Hallucination Detection in Large Reasoning Models

Arxiv

0+阅读 · 11月15日

A Unified Framework for Inference with General Missingness Patterns and Machine Learning Imputation

Arxiv

0+阅读 · 11月11日

A Unified Framework for Inference with General Missingness Patterns and Machine Learning Imputation

Arxiv

0+阅读 · 11月7日

相关基金

含非正态及缺失数据的结构方程模型分析

国家自然科学基金

0+阅读 · 2015年12月31日

线性时序关系下推理的概率计量化模型

国家自然科学基金

0+阅读 · 2014年12月31日

一般误差分布下若干半参数模型的复合分位数方法

国家自然科学基金

0+阅读 · 2014年12月31日

变换结构方程模型的非参数贝叶斯分析

国家自然科学基金

4+阅读 · 2014年12月31日

复杂数据下含指标项半参数模型结构的统计推断及应用

国家自然科学基金

0+阅读 · 2014年12月31日

微信扫码咨询专知VIP会员