Offline reinforcement learning
Policy learning and evaluation from fixed logs of prior decisions, without further interaction.
Learning to act from data someone else collected, for reasons you cannot fully reconstruct, in a world you cannot go back and query.
- RL
- Offline RL