开源7B与32B参数大模型,专注长文本推理与代码生成。
Olmo 3
- 基于完整构建流程的全开源模型家族,包含所有训练数据与依赖。
- 32B版本在推理与代码任务上表现领先,是当前最强开源思考型模型。
- 适合研究者复现、开发者部署,尤其关注可解释性与可控性的场景。
我们推出Olmo 3,一个在7B和32B参数规模下的前沿全开源语言模型系列。该系列旨在提升长上下文推理、函数调用、编程能力、指令遵循、通用对话及知识召回性能。本次发布涵盖模型全生命周期:包括所有训练阶段、检查点、数据点及依赖项。其旗舰模型Olmo 3 Think 32B是迄今发布的最强全开源思考型模型。
原文摘要 · Abstract (English)
We introduce Olmo 3, a family of state-of-the-art, fully-open language models at the 7B and 32B parameter scales. Olmo 3 model construction targets long-context reasoning, function calling, coding, instruction following, general chat, and knowledge recall. This release includes the entire model flow, i.e., the full lifecycle of the family of models, including every stage, checkpoint, data point, and dependency used to build it. Our flagship model, Olmo 3 Think 32B, is the strongest fully-open thinking model released to-date.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。