ScholarGate
Assistente

Confronta i metodi

Esamina i metodi selezionati fianco a fianco; le righe che differiscono sono evidenziate.

Apprendimento per Rinforzo Semi-Supervisionato×Apprendimento per Rinforzo Debolmente Supervisionato×
CampoApprendimento profondoApprendimento profondo
FamigliaMachine learningMachine learning
Anno di origine2020s2010s–present
IdeatoreMultiple contributors (Laskin, Srinivas, Abbeel et al.)Multiple contributors; reward-learning framing: Christiano et al. (2017)
TipoSemi-supervised training paradigm for RL agentsReinforcement learning with imperfect or partial reward supervision
Fonte seminaleZhan, X., Zhu, X., & Shi, H. (2022). Deepthermal: Combustion optimization for thermal power generating units using offline reinforcement learning. Proceedings of the AAAI Conference on Artificial Intelligence, 36(4), 4680–4688. link ↗Sutton, R. S. & Barto, A. G. (2018). Reinforcement Learning: An Introduction (2nd ed.). MIT Press. ISBN: 978-0-262-03924-6
AliasSSRL, semi-supervised RL, RL with unlabeled data, label-efficient reinforcement learningWSRL, weak-reward RL, imperfect-reward reinforcement learning, reward-impoverished RL
Correlati63
SintesiSemi-supervised reinforcement learning (SSRL) combines standard reinforcement learning — where an agent learns from sparse reward signals — with semi-supervised techniques that extract structure from unlabeled environment interactions. The goal is to improve sample efficiency and generalization when reward feedback is costly, delayed, or available only for a fraction of the agent's experience.Weakly supervised reinforcement learning (WSRL) trains agents in environments where the reward signal is imperfect, sparse, delayed, or only partially informative — unlike dense fully-supervised RL. The agent must learn effective policies despite incomplete feedback, using auxiliary signals, reward modeling, or preference learning to compensate for the weak supervision.
ScholarGateInsieme di dati
  1. v1
  2. 2 Fonti
  3. PUBLISHED
  1. v1
  2. 2 Fonti
  3. PUBLISHED

Vai alla ricerca Scarica le diapositive

ScholarGateConfronta i metodi: Semi-supervised Reinforcement Learning · Weakly supervised reinforcement learning. Consultato il 2026-06-17 da https://scholargate.app/it/compare