← 피드로
[Submitted on 13 Sep 2026 (v1), last revised 16 Sep 2026 (this version, v2)]
Abstract:Agentic Network Operations (NetOps) are an emerging paradigm promising to enable workload-aware, self-adjustable, and reliable autonomous networks. While agents have proven their value in incident summarization and telemetry signal extraction, their effectiveness as autonomous control-loop engines heavily relies on their long-horizon reliability. One such setting is the datacenter fabric, where an agent must respond to alarms and operator intents while abstaining from high-risk actions that may cause or extend downtime. Abstention, however, presupposes that an action’s impact is known pre-execution, which necessitates a per-action ground truth that NetOps agent benchmarks do not provide. We construct such a ground truth for the network repair task of NetArena. A symbolic replay of the emulated network, validated against the environment at every turn, yields the exact value of every action. From the action-level value, we derive two pre-execution targets, namely whether an action reduces the repair distance (progress) and whether it increases it (harm). We show across 10 agent models, that agent verifiers leveraging internal signals predict both harm and progress more reliably than a baseline using observable signals only. Perspectively, we aim to use these signals as safety feedback to an agent harness to abstain from risky actions and protect the target system.
Submission history
From: Tobias Labarta [view email]
[v1]
Sun, 13 Sep 2026 10:43:37 UTC (1,291 KB)
[v2]
Wed, 16 Sep 2026 08:21:07 UTC (1,291 KB)
추출 본문 · 출처: arxiv.org · https://arxiv.org/abs/2609.14422