unilab.envs.mdp.rewards.action_rate_l2

unilab.envs.mdp.rewards.action_rate_l2(env)[source]

Penalize the first difference of raw policy actions.

Parameters:

env (ManagerBasedRlEnv)

Return type:

ndarray