Reddit r/LocalLLaMA (开源AI)· /u/GodComplecs·· 1 天前AI 评分28
开源实验室还应继续产出优秀的非思考(instruct)模型吗?
Should we plead opensource labs to still produce great non thinking (instruct) models?
AI 导读
Reddit 用户对比发现,新模型的非思考(instruct)模式性能退化比旧模型更严重,如 3.6 在多个 coding benchmark 的 instruct 模式上领先 3.8。作者呼吁实验室继续投入 instruct 模式:不少用户并非所有场景都需要长推理链,为 Agentic 能力牺牲基础模型性能不是好方向。
来源:Reddit r/LocalLLaMA (开源AI) · reddit.com