On the Infinite Width and Depth Limits of Predictive Coding Networks

작성자

카테고리:

← 피드로
arXiv cs.AI · Francesco Innocenti, El Mehdi Achour, Rafal Bogacz · 2026-08-11 AI

[Submitted on 7 Feb 2026 (v1), last revised 8 Aug 2026 (this version, v3)]

View PDF HTML (experimental)

Abstract:Predictive coding (PC) is a biologically plausible alternative to standard backpropagation (BP) that minimises an energy function with respect to network activities before updating weights. Recent work has improved the training stability of deep PC networks (PCNs) by leveraging some BP-inspired reparameterisations, but the scalability and theoretical basis of these methods remain unclear. To address this gap, we study the infinite width and depth limits of PCNs. For linear networks, we derive stable and “non-lazy” parameterisations when scaling both the model width and depth, revealing that the output of standard PCNs explodes with width during training. Moreover, under stable parameterisations, we show that the gradients computed by PC at activity equilibrium converge to the BP gradients for networks that are much wider than deep ($depth/widthto0$). Experiments show high gradient alignment between PC and BP at large width for different nonlinear models, including convolutional networks and transformers. Overall, this work constrains the parameterisations that are scalable with PC, while suggesting how BP could be implemented using only local updates in much wider than deep networks like the brain.

Submission history

From: Francesco Innocenti [view email]
[v1] Sat, 7 Feb 2026 20:47:32 UTC (22,122 KB)
[v2] Fri, 22 May 2026 13:35:30 UTC (22,305 KB)
[v3] Sat, 8 Aug 2026 15:22:29 UTC (22,303 KB)

원문에서 계속 ↗

추출 본문 · 출처: arxiv.org · https://arxiv.org/abs/2602.07697

코멘트

답글 남기기