Methodforward inputs: action: (B, Ta, Da) action chunk time: (B,) or float within [0,1), flow time cond: dict with key
model/flow/mlp_flow.py:474
Methodforward cond: dict with key state/rgb; more recent obs at the end state: (B, To, Do) rgb: (B, To, C, H, W) no_augment
model/common/critic.py:181
Methodloss PPO loss obs: dict with key state/rgb; more recent obs at the end state: (B, To, Do) rgb: (B, To, C, H, W)
model/diffusion/diffusion_ppo.py:79
Methodloss PPO loss obs: dict with key state/rgb; more recent obs at the end "state": (B, To, Do) "rgb": (B, To, C, H, W
model/flow/ft_ppo/ppoflow.py:398
Methodloss PPO loss obs: dict with key state/rgb; more recent obs at the end state: (B, To, Do) rgb: (B, To, C, H, W)
model/gaussian/gaussian_ppo.py:61
Methodloss PPO loss obs: dict with key state/rgb; more recent obs at the end state: (B, To, Do) rgb: (B, To, C, H, W)
model/gaussian/gmm_ppo.py:61
Methodloss PPO loss obs: dict with key state/rgb; more recent obs at the end state: (B, To, Do) rgb: (B, To, C, H, W)
model/rl/gaussian_ppo.py:61