S3: Stable Subgoal Selection by Constraining Uncertainty of Coarse Dynamics in Hierarchical Reinforcement Learning
Hierarchical Reinforcement Learning (HRL) intends to separate strategic planning from primitive execution. It has been widely successful in solving long-horizon and complex tasks, where flat-RL algorithms have difficulty in learning. However, while the low-level agent in HRL benefits from dense feedback and abundant trial opportunities, the high-level agent receives sparse, delayed feedback from the environment and its performance depends on the low-level execution capability. In this paper, we study whether subgoal selection by the high-level agent can be performed more strategically, by prov
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- FuzzySimilar title/name (fuzzy) · 59%stabilityai/stable-diffusion-xl-base-1.0 →
“Fuzzy title match (0.73): “S3: Stable Subgoal Selection by Constraining Uncertainty of ” ≈ “stabilityai/stable-diffusion-xl-base-1.0””
- FuzzySimilar title/name (fuzzy) · 59%CompVis/stable-diffusion-v1-4 →
“Fuzzy title match (0.73): “S3: Stable Subgoal Selection by Constraining Uncertainty of ” ≈ “CompVis/stable-diffusion-v1-4””
- FuzzySimilar title/name (fuzzy) · 59%stabilityai/stable-diffusion-3.5-large →
“Fuzzy title match (0.73): “S3: Stable Subgoal Selection by Constraining Uncertainty of ” ≈ “stabilityai/stable-diffusion-3.5-large””
- PossiblePossibly related (embedding) · 51%Gradient-based Planning for World Models at Longer Horizons →
- FuzzySimilar title/name (fuzzy) · 59%DLR-RM/stable-baselines3 →
“Fuzzy title match (0.73): “S3: Stable Subgoal Selection by Constraining Uncertainty of ” ≈ “DLR-RM/stable-baselines3””
- FuzzySimilar title/name (fuzzy) · 59%aymericdamien/TopDeepLearning →
“Fuzzy title match (0.73): “S3: Stable Subgoal Selection by Constraining Uncertainty of ” ≈ “aymericdamien/TopDeepLearning””
- LinkedLinked via arxiv author · 85%Kshitij Kumar Srivastava →
“S3: Stable Subgoal Selection by Constraining Uncertainty of Coarse Dynamics in Hierarchical Reinforcement Learning”
- LinkedLinked via arxiv author · 85%Kshitij Jerath →
“S3: Stable Subgoal Selection by Constraining Uncertainty of Coarse Dynamics in Hierarchical Reinforcement Learning”
