arXiv:2601.11531cs.HCcs.AI2026-01AAAI

用自然语言生成IT监控看板,提升运维效率。

NOVAID: Natural-language Observability Visualization Assistant for ITOps Dashboard Widget Generation

  • 基于大模型的语义解析与动态数据获取,理解运维查询意图。
  • 在271个真实查询上实现94.1%的指标提取准确率。
  • 适合运维工程师快速构建可视化看板,降低使用门槛。

手动创建IT监控看板耗时易错,阻碍新手与专家用户。我们提出NOVAID,一个交互式聊天机器人,利用大语言模型(LLMs)直接从自然语言查询生成IT监控组件。不同于通用自然语言转可视化工具,NOVAID解决运维特定挑战:如SLO图表等专用组件类型、动态API数据获取及复杂上下文过滤。系统结合领域感知语义解析、模糊实体匹配与模式补全,生成标准化的组件JSON规范,并通过交互澄清环确保查询不完整时的准确性。在包含271个真实查询的定制数据集上,NOVAID在多个LLM中达到最高94.10%的指标提取准确率。对运维工程师的用户研究显示,其系统可用性量表(SUS)得分为74.2,表明良好可用性。通过连接自然语言意图与运维看板,NOVAID展现出在企业运维平台部署的明确潜力。

原文摘要 · Abstract (English)

Manual creation of IT monitoring dashboard widgets is slow, error-prone, and a barrier for both novice and expert users. We present NOVAID, an interactive chatbot that leverages Large Language Models (LLMs) to generate IT monitoring widgets directly from natural language queries. Unlike general natural language-to-visualization tools, NOVAID addresses IT operations-specific challenges: specialized widget types like SLO charts, dynamic API-driven data retrieval, and complex contextual filters. The system combines a domain-aware semantic parser, fuzzy entity matching, and schema completion to produce standardized widget JSON specifications. An interactive clarification loop ensures accuracy in underspecified queries. On a curated dataset of 271 realistic queries, NOVAID achieves promising accuracy (up to 94.10% in metric extraction) across multiple LLMs. A user study with IT engineers yielded a System Usability Scale score of 74.2 for NOVAID, indicating good usability. By bridging natural language intent with operational dashboards, NOVAID demonstrates clear potential and a path for deployment in enterprise ITOps monitoring platforms.

自然语言运维监控大模型应用

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。