lerobot-clone

mirror of https://github.com/huggingface/lerobot.git synced 2026-06-03 12:21:27 +00:00

Author	SHA1	Message	Date
Michel Aractingi	ff47c0b0d3	- Fixed big issue in the loading of the policy parameters sent by the learner to the actor -- pass only the actor to the `update_policy_parameters` and remove `strict=False` - Fixed big issue in the normalization of the actions in the `forward` function of the critic -- remove the `torch.no_grad` decorator in `normalize.py` in the normalization function - Fixed performance issue to boost the optimization frequency by setting the storage device to be the same as the device of learning. Co-authored-by: Adil Zouitine <adilzouitinegm@gmail.com>	2025-02-19 16:22:51 +00:00
AdilZouitine	2f3370e42f	Add maniskill support. Co-authored-by: Michel Aractingi <michel.aractingi@gmail.com>	2025-02-14 19:53:29 +00:00
Michel Aractingi	7ae368e983	Fixed bug in the action scale of the intervention actions and offline dataset actions. (scale by inverse delta) Co-authored-by: Adil Zouitine <adizouitinegm@gmail.com>	2025-02-14 15:17:16 +01:00
Michel Aractingi	d9a70376d8	Changed the init_final value to center the starting mean and std of the policy Co-authored-by: Adil Zouitine <adilzouitinegm@gmail.com>	2025-02-13 16:42:43 +01:00
Michel Aractingi	c462a478c7	Hardcoded some normalization parameters. TODO refactor Added masking actions on the level of the intervention actions and offline dataset Co-authored-by: Adil Zouitine <adilzouitinegm@gmail.com>	2025-02-13 14:27:14 +01:00
Michel Aractingi	dc086dc21f	Added logging for interventions to monitor the rate of interventions through time Added an s keyboard command to force success in the case the reward classifier fails Co-authored-by: Adil Zouitine <adilzouitinegm@gmail.com>	2025-02-13 11:04:49 +01:00
Michel Aractingi	b9217b06db	Added possiblity to record and replay delta actions during teleoperation rather than absolute actions Co-authored-by: Adil Zouitine <adilzouitinegm@gmail.com>	2025-02-12 19:25:41 +01:00
Michel Aractingi	a7db3959f5	- Added JointMaskingActionSpace wrapper in `gym_manipulator` in order to select which joints will be controlled. For example, we can disable the gripper actions for some tasks. - Added Nan detection mechanisms in the actor, learner and gym_manipulator for the case where we encounter nans in the loop. - changed the non-blocking in the `.to(device)` functions to only work for the case of cuda because they were causing nans when running the policy on mps - Added some joint clipping and limits in the env, robot and policy configs. TODO clean this part and make the limits in one config file only. Co-authored-by: Adil Zouitine <adilzouitinegm@gmail.com>	2025-02-11 11:34:46 +01:00
Michel Aractingi	d51374ce12	Several fixes to move the actor_server and learner_server code from the maniskill environment to the real robot environment. Co-authored-by: Adil Zouitine <adilzouitinegm@gmail.com>	2025-02-10 16:03:39 +01:00
Michel Aractingi	12525242ce	- Added `lerobot/scripts/server/gym_manipulator.py` that contains all the necessary wrappers to run a gym-style env around the real robot. - Added `lerobot/scripts/server/find_joint_limits.py` to test the min and max angles of the motion you wish the robot to explore during RL training. - Added logic in `manipulator.py` to limit the maximum possible joint angles to allow motion within a predefined joint position range. The limits are specified in the yaml config for each robot. Checkout the so100.yaml. Co-authored-by: Adil Zouitine <adilzouitinegm@gmail.com>	2025-02-06 16:29:37 +01:00

10 Commits