Utilizing Pre-trained and Large Language Models for 10-K Items Segmentation
利用预训练和大语言模型进行10-K文件分段
专题命中 长上下文与记忆 :large language model(title,abstract);language model(title,abstract);prompting(abstract)
AI总结 本文提出两种先进分段方法,BERT4ItemSeg与GPT4ItemSeg,通过预训练模型与Bi-LSTM结合或大语言模型,提升10-K报告分段性能,核心指标达到宏F1值0.9825。
Comments Accepted for publication in the Journal of Information Systems