refactor(processor): clarify action types, distinguish PolicyAction, RobotAction, and EnvAction (#1908)

* refactor(processor): split action from policy, robots and environment - Updated function names to robot_action_to_transition and robot_transition_to_action across multiple files to better reflect their purpose in processing robot actions. - Adjusted references in the RobotProcessorPipeline and related components to ensure compatibility with the new naming convention. - Enhanced type annotations for action parameters to improve code readability and maintainability. * refactor(converters): rename robot_transition_to_action to transition_to_robot_action - Updated function names across multiple files to improve clarity and consistency in processing robot actions. - Adjusted references in RobotProcessorPipeline and related components to align with the new naming convention. - Simplified action handling in the AddBatchDimensionProcessorStep by removing unnecessary checks for action presence. * refactor(converters): update references to transition_to_robot_action - Renamed all instances of robot_transition_to_action to transition_to_robot_action across multiple files for consistency and clarity in the processing of robot actions. - Adjusted the RobotProcessorPipeline configurations to reflect the new naming convention, enhancing code readability. * refactor(processor): update Torch2NumpyActionProcessorStep to extend ActionProcessorStep - Changed the base class of Torch2NumpyActionProcessorStep from PolicyActionProcessorStep to ActionProcessorStep, aligning it with the current architecture of action processing. - This modification enhances the clarity of the class's role in the processing pipeline. * fix(processor): main action processor can take also EnvAction --------- Co-authored-by: Steven Palma <steven.palma@huggingface.co>
2026-06-01 03:11:29 +00:00 · 2025-09-10 22:40:37 +02:00
parent 6745958362
commit 9183083e75
22 changed files with 303 additions and 139 deletions
--- a/tests/processor/test_normalize_processor.py
+++ b/tests/processor/test_normalize_processor.py
@@ -329,14 +329,14 @@ def test_min_max_unnormalization(action_stats_min_max):
    assert torch.allclose(unnormalized_action, expected)


-def test_numpy_action_input(action_stats_mean_std):
+def test_tensor_action_input(action_stats_mean_std):
    features = _create_action_features()
    norm_map = _create_action_norm_map_mean_std()
    unnormalizer = UnnormalizerProcessorStep(
        features=features, norm_map=norm_map, stats={"action": action_stats_mean_std}
    )

-    normalized_action = np.array([1.0, -0.5, 2.0], dtype=np.float32)
+    normalized_action = torch.tensor([1.0, -0.5, 2.0], dtype=torch.float32)
    transition = create_transition(action=normalized_action)

    unnormalized_transition = unnormalizer(transition)