Multimodal Collaborative Debate for Zero-Shot Time Series Reasoning

작성자

카테고리:

← 피드로
arXiv cs.AI · Patara Trirat, Jin Myung Kwak, Jay Heo, Heejun Lee, Sung Ju Hwang · 2026-08-31 AI

[Submitted on 27 Jan 2026 (v1), last revised 28 Aug 2026 (this version, v2)]

View PDF HTML (experimental)

Abstract:Large language models (LLMs) are increasingly used as natural-language interfaces to structured data, yet they remain brittle when reasoning over time series. Visual patterns can be misleading, numerical claims can be hallucinated, and textual context can override evidence from the signal. We study zero-shot time-series reasoning as a multimodal evidence arbitration problem for LLM agents. We propose TS-Debate, an inference-time multi-agent protocol that requires no task-specific fine-tuning. TS-Debate first elicits relevant domain knowledge, then assigns modality-specialized agents to textual context, visual patterns, and numerical signals, and coordinates their interaction through a verification-conflict-calibration procedure. Reviewer agents check decision-critical claims with lightweight code execution and numerical lookup, resolve cross-modal disagreement, and calibrate the final answer. Unlike generic multi-agent debate or unconstrained tool use, TS-Debate specifies how evidence is exposed, which claims are checkable, and how verification outcomes shape synthesis. Across 20 tasks from three public benchmarks, TS-Debate improves classification and question answering performance over strong baselines, while revealing that debate is most useful for global-structure and cross-view reasoning rather than local value reconstruction.

Submission history

From: Patara Trirat [view email]
[v1] Tue, 27 Jan 2026 03:29:22 UTC (1,361 KB)
[v2] Fri, 28 Aug 2026 10:17:09 UTC (1,236 KB)

원문에서 계속 ↗

추출 본문 · 출처: arxiv.org · https://arxiv.org/abs/2601.19151