arXiv:2506.00019cs.CLcs.AI2025-06被引 2

开源葡萄牙语大模型家族,适配多种应用场景

Amadeus-Verbo Technical Report: The powerful Qwen2.5 family models trained in Portuguese

  • 基于Qwen2.5微调,覆盖0.5B至72B参数规模
  • 提供基础、合并、指令微调三类模型,支持多任务需求
  • 模型全公开于HuggingFace,助力巴西葡语AI生态

本报告介绍了Amadeus Verbo的开发经验,这是一个面向巴西葡萄牙语的大语言模型系列。为应对多样化的使用场景,Amadeus Verbo包含0.5B、1.5B、3B、7B、14B、32B和72B参数规模的基座微调、合并及指令微调模型。其核心目标是展示在具备数据与资源条件下,微调基础模型以实现巴西葡萄牙语开源大模型民主化开发的可行性。所有Amadeus-Verbo系列模型均已在HuggingFace平台开放获取:https://huggingface.co/collections/amadeusai/amadeus-verbo-qwen25-67cf2e7aae69ce2b3bcdcfda。

原文摘要 · Abstract (English)

This report introduces the experience of developing Amadeus Verbo, a family of large language models for Brazilian Portuguese. To handle diverse use cases, Amadeus Verbo includes base-tuned, merged, and instruction-tuned models in sizes of 0.5B, 1.5B, 3B, 7B, 14B, 32B, and 72B parameters. Thus, the main objective is to show how easy it is to fine-tune foundation models to democratize the open-source development of Brazilian Portuguese LLMs when data and resources are available. Amadeus-Verbo family models are all available at HuggingFace at https://huggingface.co/collections/amadeusai/amadeus-verbo-qwen25-67cf2e7aae69ce2b3bcdcfda.

葡萄牙语大模型开源微调

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。