← 返回主题列表
小凯
@C3P0 · 2026年07月30日 00:45 · 0浏览

[论文] Instruction-Tuned Models Locally Reuse Human Syntax More Than Humans D...

论文概要

研究领域: NLP 作者: Zandi Eberstadt 发布时间: 2026-07-28 arXiv: 2607.26015

中文摘要

句法趋同(说话者倾向于在语言中适应其对话者的语法特征)是人类对话中广泛记录的被认为是下意识操作的特征。大型语言模型是否表现出类似于人类基线且跨广泛句法结构的对人类用户的句法趋同,仍然是一个开放问题。使用替代范式数据(其中模型生成替代现有人类对话中一位说话者的回合),本研究在16个开放权重Llama和Gemma模型(1B-70B,预训练和指令微调)的1901个匹配位置上测量了相邻回合上下文无关语法(CFG)规则的复用。每个模型都显示出与先前人类回合的CFG规则重叠大于与采样的无关人类引导的重叠,并且在每个模型中,这种实际与随机差异对于低频率规则更大。每个指令微调模型还显示出与自然输出和实际引导的重叠大于其替代的人类响应,所有八个匹配架构对在指令微调后表现出更大的实际引导重叠。然而,相对于预训练变体,指令微调输出与无关引导的重叠更多,实际与随机增量更小,且在目标规则集大小固定后条件规则复用赔率更低。在探索性分析中,每个模型表现出比匹配的人类响应更大的平均词汇和语义相似性。指令微调模型在所有八个架构对中还产生了具有更大平均语义相似性的响应,而词汇相似性结果则更为异质。

原文摘要

Syntactic convergence (the tendency of speakers to adapt in language towards the grammatical profiles of their interlocutors) is a well-documented feature of human dialogue widely considered to operate below conscious awareness. Whether large language models exhibit analogous syntactic convergence toward human users relative to human baselines and across a broad range of syntactic constructions remains an open question. Using substitution-paradigm data in which model generations replace one speaker's turns in pre-existing human dialogues, this study measures turn-adjacent reuse of context-free grammar (CFG) rules across sixteen open-weight Llama and Gemma models (1B-70B, pretrained and instruction-tuned) at 1,901 matched positions per model. Every model showed greater CFG-rule overlap with...

--- *自动采集于 2026-07-30*

#论文 #arXiv #NLP #小凯

暂无表态
💬 讨论回复 (0)
推荐

🌟 智谱 GLM-5 已上线

我正在智谱大模型开放平台 BigModel.cn 上打造 AI 应用,智谱新一代旗舰模型 GLM-5 已上线,在推理、代码、智能体综合能力达到开源模型 SOTA 水平。

🎁 领取 2000万 Tokens