Learning When to Sample: Confidence-Aware Selective Sampling for Efficient Chain-of-Thought Reasoning
学习何时采样:面向高效链式思维推理的置信度感知选择性采样
机构 * Vanderbilt University(范德比尔特大学) ; Vanderbilt University Medical Center(范德比尔特大学医学中心) ; Intuit AI Research(Intuit AI研究院)
专题命中 测试时计算 :reasoning(title,abstract);chain-of-thought(title,abstract);CoT(abstract,abstract_cn);分类 cs.CL
AI总结 提出置信度感知选择性采样框架,通过分析单条推理轨迹自适应决定是否触发多路径采样,在保持性能的同时显著降低推理成本。