lerobot

Files

AdilZouitine 306c735172 Refactor SAC policy and training loop to enhance discrete action support

- Updated SACPolicy to conditionally compute losses for grasp critic based on num_discrete_actions.
- Simplified forward method to return loss outputs as a dictionary for better clarity.
- Adjusted learner_server to handle both main and grasp critic losses during training.
- Ensured optimizers are created conditionally for grasp critic based on configuration settings.

2025-04-01 11:42:28 +00:00

configuration_sac.py

Refactor SAC policy and training loop to enhance discrete action support

2025-04-01 11:42:28 +00:00

modeling_sac.py

Refactor SAC policy and training loop to enhance discrete action support

2025-04-01 11:42:28 +00:00