Updated 41 min ago · first seen 11 Sept 2026
paper_01M294FNRTYB10NXD329EGG6PJ
- Published
- 11 Sept 2026
- T1 · 41 min ago
- arXiv
- 2609.10776
- T1 · 41 min ago
- Category
- cs.LG
- T1 · 41 min ago
Abstract
In continual reinforcement learning, carefully managing the stability-plasticity tradeoff remains a core challenge. Recent work by Abel et al. (2025) formalized this dilemma by defining plasticity as the generalized directed information from an agent's observations to its actions, and empowerment as the generalized directed information from its actions to its observations. This formulation successfully reframes the traditional stability-plasticity tradeoff as an empowerment-plasticity tradeoff. However, while extensive literature exists on optimizing for empowerment, there is currently no research addressing the optimization of plasticity under this new definition. This paper presents preliminary work toward optimizing plasticity within Markov decision processes. We show that there exists a Bellman optimality equation for optimizing plasticity similar to previous work for empowerment.
Authors 2
Jeremy Lucas, Doina Precup
Specification
- Official page
Source:arXiv (Atom API + RSS)T1observed 41 min agohigh
- Arxiv announce type
- new
Source:arXiv (Atom API + RSS)T1observed 41 min agohigh
- arXiv id
- 2609.10776
Source:arXiv (Atom API + RSS)T1observed 41 min agohigh
- Categories
- cs.LG
Source:arXiv (Atom API + RSS)T1observed 41 min agohigh
Source:arXiv (Atom API + RSS)T1observed 41 min agohigh
- Primary category
- cs.LG
Source:arXiv (Atom API + RSS)T1observed 41 min agohigh
- Published
- 11 Sept 2026
Source:arXiv (Atom API + RSS)T1observed 41 min agohigh
Each value shows its source, tier and observation time. Conflicting claims are kept side by side and flagged — never averaged. How AI Atlas records facts →
Provenance
Attributed facts
9
Source tiers
T19
Freshest observation
41 min ago
Conflicts
None
No models linked to this paper yet.
- Authors
- Jeremy Lucas, Doina Precup
As of
Rewind the record: see this entity's attributes exactly as AI Atlas knew them on a given day.
Claim history · Official page
Official pageofficial_url1
| Value | Valid from → to | Status | Source | Confidence | Extractor |
|---|---|---|---|---|---|
| https://arxiv.org/abs/2609.10776 | → current | current | arXiv (Atom API + RSS)T1 | high | deterministic |
Claims are temporal and append-only: a new observation closes the previous claim (valid_to) instead of overwriting it. Conflicting claims from different sources are kept side by side and flagged — never averaged. Methodology →
New paper: A Bellman Optimality Equation for Plasticity
arxiv
| Source | Document | Type | Tier | Last observed | Snapshots |
|---|---|---|---|---|---|
| arXiv (Atom API + RSS) | rss.arxiv.org/rss/cs.LG | feed | T1· Official | 41 min ago | 1 |
Tier 1 = official/primary, 2 = quality secondary, 3 = community, 4 = unverified. Every snapshot is archived; see all sources and the methodology.