用极坐标几何重构线性算子,提升训练稳定性和可解释性。
Foundations of Polar Linear Algebra
- 将线性算子分解为径向与角向分量,基于极坐标构建新框架。
- 在MNIST上训练可靠,参数量和计算复杂度降低,收敛更稳定。
- 适合关注频谱结构与模型并行的场景,如大规模神经网络设计。
本文从谱视角重新审视算子学习,提出极坐标线性代数(Polar Linear Algebra)框架,该框架基于极几何,结合线性径向分量与周期角向分量。由此定义相关算子并分析其谱性质。以标准基准(MNIST)为验证,结果表明极坐标及全谱算子可被可靠训练,且施加类自伴谱约束能提升稳定性与收敛性。除准确率外,该框架还减少参数量与计算复杂度,实现解耦频谱模式的可解释表示。通过从空间域转向谱域,问题分解为正交特征模态,可作为独立计算流水线处理。此结构自然引入额外的模型并行维度,补充现有策略,无需人为划分。整体上,该工作为算子学习提供新概念视角,尤其适用于频谱结构与并行执行为核心的问题。
原文摘要 · Abstract (English)
This work revisits operator learning from a spectral perspective by introducing Polar Linear Algebra, a structured framework based on polar geometry that combines a linear radial component with a periodic angular component. Starting from this formulation, we define the associated operators and analyze their spectral properties. As a proof of feasibility, the framework is evaluated on a canonical benchmark (MNIST). Despite the simplicity of the task, the results demonstrate that polar and fully spectral operators can be trained reliably, and that imposing self-adjoint-inspired spectral constraints improves stability and convergence. Beyond accuracy, the proposed formulation leads to a reduction in parameter count and computational complexity, while providing a more interpretable representation in terms of decoupled spectral modes. By moving from a spatial to a spectral domain, the problem decomposes into orthogonal eigenmodes that can be treated as independent computational pipelines. This structure naturally exposes an additional dimension of model parallelization, complementing existing parallel strategies without relying on ad-hoc partitioning. Overall, the work offers a different conceptual lens for operator learning, particularly suited to problems where spectral structure and parallel execution are central.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。