KAN用可学习的样条函数替代传统激活函数,提升模型可解释性与效率。
A Survey on Kolmogorov-Arnold Network
- 用可学习的样条函数代替固定激活函数,实现灵活且可解释的高维函数拟合。
- 在时间序列、生物医学和图学习中表现优异,支持高效参数化与扩展。
- 适合需要可解释性与自适应能力的复杂建模任务,如动态系统与科学计算。
本文系统综述了受柯尔莫戈罗夫-阿诺德表示定理启发的柯尔莫戈罗夫-阿诺德网络(KAN)的理论基础、演进历程、应用前景及未来潜力。KAN通过使用可学习的样条参数化函数替代传统神经网络中的固定激活函数,实现了对高维函数的灵活且可解释的表示。该综述详细阐述了其架构优势,包括基于边的自适应激活函数,显著提升了参数效率与可扩展性,在时间序列预测、计算生物医学和图学习等场景中表现出色。关键进展如Temporal-KAN、FastKAN和PDE-KAN展示了其在动态环境中的广泛应用,增强了可解释性、计算效率与复杂函数逼近的适应性。此外,本文还讨论了KAN与卷积、循环及Transformer等架构的融合,体现了其在混合建模任务中的灵活性。尽管在高维与噪声数据下仍面临计算挑战,推动了优化策略、正则化技术与混合模型的持续研究。本综述强调了KAN在现代神经网络架构中的地位,并指明了提升其计算效率、可解释性与可扩展性的未来方向。
原文摘要 · Abstract (English)
This systematic review explores the theoretical foundations, evolution, applications, and future potential of Kolmogorov-Arnold Networks (KAN), a neural network model inspired by the Kolmogorov-Arnold representation theorem. KANs distinguish themselves from traditional neural networks by using learnable, spline-parameterized functions instead of fixed activation functions, allowing for flexible and interpretable representations of high-dimensional functions. This review details KAN's architectural strengths, including adaptive edge-based activation functions that improve parameter efficiency and scalability in applications such as time series forecasting, computational biomedicine, and graph learning. Key advancements, including Temporal-KAN, FastKAN, and Partial Differential Equation (PDE) KAN, illustrate KAN's growing applicability in dynamic environments, enhancing interpretability, computational efficiency, and adaptability for complex function approximation tasks. Additionally, this paper discusses KAN's integration with other architectures, such as convolutional, recurrent, and transformer-based models, showcasing its versatility in complementing established neural networks for tasks requiring hybrid approaches. Despite its strengths, KAN faces computational challenges in high-dimensional and noisy data settings, motivating ongoing research into optimization strategies, regularization techniques, and hybrid models. This paper highlights KAN's role in modern neural architectures and outlines future directions to improve its computational efficiency, interpretability, and scalability in data-intensive applications.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。