Вход на сайт

Просмотр новости

Найдите то, что Вас интересует

Almost Sure Convergence of Linear Temporal Difference Learning with Arbitrary Features

Дата публикации: 17-08-2026 20:26:00


Temporal difference (TD) learning with linear function approximation (linear TD) is a classic and powerful prediction algorithm in reinforcement learning. While it is well-understood that linear TD converges almost surely to a unique point, this convergence traditionally requires the assumption that the features used by the approximator are linearly independent. However, this linear independence assumption does not hold in many practical scenarios. This work is the first to establish the almost sure convergence of linear TD without requiring linearly independent features. We prove that the weight iterates of linear TD converge to a bounded set, and that the value estimates derived from the weights in that set are the same almost everywhere. We also establish a notion of local stability of the weight iterates. Importantly, we do not impose assumptions tailored to feature dependence and do not modify the linear TD algorithm. Key to our analysis is a novel characterization of bounded invariant sets of the mean ODE of linear TD.

Схожие новости

#Наименование новостиТональностьИнформативностьДата публикации
1 A Two-Timescale Primal-Dual Framework for Reinforcement Learning via Online Dual Variable Guidance 013.1217-08-2026
2 Approximations and Learning for Continuous State and Action MDPs under Average Cost Criteria 04.117-08-2026
3 Finite-Time Decoupled Convergence in Nonlinear Two-Time-Scale Stochastic Approximation 09.6417-08-2026
4 Statistical Learning Theory for Neural Operators 010.2117-08-2026
5 Convergence of Decentralized Stochastic Subgradient-based Methods for Nonsmooth Nonconvex Optimization 08.7817-08-2026
6 Near-optimal Delta-convex Estimation of Lipschitz Functions 09.7117-08-2026
7 A Functional-Space Mean-Field Theory of Partially-Trained Three-Layer Neural Networks 010.9717-08-2026
8 Statistical guarantees for denoising reflected diffusion models 05.317-08-2026
9 High-Dimensional Analysis of Gradient Flow for Extensive-Width Quadratic Neural Networks 08.717-08-2026
10 Convergence of Noise-Free Sampling Algorithms with Regularized Wasserstein Proximals 08.0217-08-2026

Классификация: Наука. Схожих патентов: 0. Схожих новостей: 10. Тональность: 0. Информативность: 4.2. Источник: jmlr.org.