Skip to content
Hot eventLive

微调Qwen3.5小模型部署语法掌握追踪器

1 reports1 sources3 hr ago updated

Get the story

AI overview

2026年10月9日,arXiv Computation and Language 发布一手论文,研究在语法掌握追踪任务上微调 Qwen3.5 小语言模型,并部署 0.8B 模型服务全部英语学习者。两个人工评测基准显示,部署的 0.8B 模型与 4B 对照模型在概念、证据片段和正确性嵌套匹配标准下,precision 与 recall 均超过提示的 GPT-5.4 和 GPT-5.6 Sol,服务成本约降低 16 倍。该研究以规模化语法概念标注为核心,验证了小模型经微调后在特定任务上可优于提示前沿大模型。

Generated from reports · updated 3 hr ago

Timeline

Follow the coverage from different angles.

Oct 9, 2026
  1. arXiv · Computation and Language
    语法概念标注规模化:部署的微调小语言模型超过提示前沿模型

    论文在语法掌握追踪任务上微调 Qwen3.5 小语言模型,并部署 0.8B 模型服务全部英语学习者。两个人工评测基准显示,部署的 0.8B 模型与 4B 对照模型在概念、证据片段和正确性嵌套匹配标准下 precision 与 recall 均超过提示的 GPT-5.4 和 GPT-5.6 Sol,服务成本约降低 16 倍。

Heat trend

Current heat 9·Comparable peak 10(Oct 9)·Comparable change over 24 hours –

02.557.510Oct9Oct9Oct9Oct9

The trend compares only the same participants observed continuously; its range may be smaller than the current heat count. Move or click on the chart to inspect hourly heat; use the left and right arrow keys to switch.