一个可扩展的深度学习工具包,助力音频任务快速训练。
autrainer: A Modular and Extensible Deep Learning Toolkit for Computer Audition Tasks
- 基于PyTorch构建,支持低代码训练与多种神经网络
- 兼容多种预处理流程,提升不同音频任务适配性
- 适合音频研究者快速搭建实验框架
本文介绍了autrainer这一新型深度学习训练框架的核心设计原则,专为计算机听觉任务设计。autrainer是一个基于PyTorch的工具包,支持在多种计算机听觉任务上实现快速、可复现且易于扩展的训练。具体而言,它提供低代码训练模式,兼容广泛神经网络结构及预处理方法。本文概述了其内部机制与核心功能,展示了其在音频任务中的通用性与灵活性。
原文摘要 · Abstract (English)
This work introduces the key operating principles for autrainer, our new deep learning training framework for computer audition tasks. autrainer is a PyTorch-based toolkit that allows for rapid, reproducible, and easily extensible training on a variety of different computer audition tasks. Concretely, autrainer offers low-code training and supports a wide range of neural networks as well as preprocessing routines. In this work, we present an overview of its inner workings and key capabilities.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。