Human#
- class statrl.settings.markovdecisionprocess.discrete_nostructure.agents.Human.Human(env)[source]#
Bases:
MDPAgentInteractive agent that asks the user for each action.
Prompts on stdin at every step and blocks until a valid action name is entered.
- Parameters:
env (object) – A wrapped environment;
env.envmust exposenameActions, so a rawDiscreteMDPwill not do.
Methods
__init__(env)play(state)Print the state and block until the user names an action.
reset(inistate)Start a new run.
update(state, action, reward, observation)Ignore the transition, the user is the policy.