An active‐inference agent, with a language model as its world model, can learn from its own experience in sevenmutually‐falsifying senses — it discovers actions it was never given, overrides its own priors from lived consequence,pays a real cost for wrong beliefs, stays calibrated, and crosses from cannot solve to solve through its own discovery.
Rahul Chouhan (Mon,) studied this question.