lerobot

mirror of https://github.com/huggingface/lerobot.git synced 2026-05-16 00:59:46 +00:00

Author	SHA1	Message	Date
Pepijn	63fad12e8d	refactor: consolidate VLM classes into single QwenVL implementation Remove Qwen2VL, Qwen3VL, Qwen35VL in favor of one QwenVL class that uses AutoModelForImageTextToText and works with the whole Qwen VL family. Moves shared _parse_skills_response to BaseVLM and extracts _build_messages/_prepare_inputs/_decode helpers to reduce duplication. Made-with: Cursor	2026-03-30 20:37:09 +02:00
Pepijn	2545f1a8ed	fix: route video_metadata through videos_kwargs for Qwen3/3.5 processors The Qwen3VLProcessor distributes kwargs to sub-processors via _merge_kwargs. Flat kwargs like video_metadata and do_sample_frames were not reaching the video processor, causing fps to default to 24 and producing shape mismatches. Pass these kwargs explicitly under videos_kwargs so they reach Qwen3VLVideoProcessor directly. Revert Qwen2VL to its simpler original approach since its processor doesn't use videos_kwargs. Made-with: Cursor	2026-03-30 19:09:46 +02:00
Pepijn	5f85b572d7	fix: unpack video_metadata from tuples and pass separately to processor The Qwen3.5 processor requires video_metadata as a separate parameter, not embedded in the video tensors. Use return_video_metadata=True from process_vision_info, then unpack the (tensor, metadata) tuples into separate videos and video_metadata lists for the processor call. Made-with: Cursor	2026-03-30 17:37:59 +02:00
Pepijn	72692525da	fix: pass fps=1.0 scalar to processor instead of video_metadata tuples The return_video_metadata=True approach causes 'list index out of range' due to (tensor, metadata) tuple format issues. Since all extracted videos are at 1fps (ffmpeg -r 1), directly pass fps=1.0 as a scalar alongside do_sample_frames=False — this gives the processor the exact fps for position embedding computation without format compatibility issues across Qwen processor versions. Made-with: Cursor	2026-03-30 17:32:30 +02:00
Pepijn	9a298524ca	fix: pass video_metadata via process_vision_info for correct position embeddings The Qwen3.5 processor needs video_metadata (fps, frame indices) to compute temporal position embeddings. Use return_video_metadata=True which embeds metadata inside the video tensors as (tensor, metadata) tuples, and return_video_kwargs=True which returns {'do_sample_frames': False} without the problematic fps list. Made-with: Cursor	2026-03-30 17:23:44 +02:00
Pepijn	002a9dd0b9	fix: use do_sample_frames=False instead of video_kwargs fps list The Qwen3.5 processor expects fps as a scalar, not a list, so passing video_kwargs with fps=[...] fails validation. Since process_vision_info already handles frame sampling, we only need do_sample_frames=False to tell the processor to use the pre-sampled frames as-is. Made-with: Cursor	2026-03-30 16:55:46 +02:00
Pepijn	e40985b013	fix: pass video_kwargs from process_vision_info to Qwen processor The Qwen processor needs fps metadata (via return_video_kwargs=True) to compute correct temporal position embeddings. Without it, the processor defaults to fps=24 regardless of the actual video fps, causing shape mismatches between expected and actual video tokens. Made-with: Cursor	2026-03-30 16:50:34 +02:00
Pepijn	d03200bdb3	fix: force torchvision video backend instead of cv2 bypass Replace manual cv2 frame reading with FORCE_QWENVL_VIDEO_READER=torchvision env var. The torchvision backend (PyAV) properly reads video metadata and respects the fps parameter, avoiding the torchcodec fps=24 default issue. Made-with: Cursor	2026-03-30 16:42:52 +02:00
Pepijn	ac41cd6672	fix: bypass torchcodec video decoding by pre-reading frames via cv2 When torchcodec is installed, qwen-vl-utils ignores the fps parameter and defaults to 24fps if video metadata is missing, causing shape mismatches. Fix by reading video frames directly as PIL images and passing them to the processor, bypassing torchcodec entirely. Made-with: Cursor	2026-03-30 16:03:26 +02:00
Pepijn	9b211a45d6	fix: disable thinking mode in Qwen35VL single-episode fallback path The single-episode `segment_skills` method was missing `enable_thinking=False` in `apply_chat_template`, causing the model to output reasoning traces instead of JSON when the batch path fails and falls back to per-episode processing. Made-with: Cursor	2026-03-30 15:31:18 +02:00
root	a6387da464	add license	2026-03-11 23:14:22 +00:00
Jade Choghari	0328b3f4aa	Update src/lerobot/data_processing/data_annotations/vlm_annotations.py Co-authored-by: Copilot <175728472+Copilot@users.noreply.github.com> Signed-off-by: Jade Choghari <chogharijade@gmail.com>	2026-03-11 16:10:37 -07:00
root	cc8e4c0d86	remove folder	2026-03-11 22:49:38 +00:00
root	819c1b9710	add tests/fixes	2026-03-11 22:49:06 +00:00
root	f0848c6887	add subtasl	2026-03-11 19:51:48 +00:00
Silvio Traversaro	19c6adef85	chore(dependencies): Increase opencv-python-headless upper bound (#3120 ) Signed-off-by: Silvio Traversaro <silvio@traversaro.it>	2026-03-09 23:27:18 +01:00
Johnson Sun	96b7f3dae0	Parse HF_USER with NO_COLOR to avoid incorrectly capturing bash ANSI codes (#3119 )	2026-03-09 18:47:58 +01:00
Martino Russi	885ef91892	fix(unitree_g1): correct SDK detection and update installation docs (#3115 ) * update docs * update toml / docs * update docs * fix joystick * Update pyproject.toml Co-authored-by: Copilot <175728472+Copilot@users.noreply.github.com> Signed-off-by: Martino Russi <77496684+nepyope@users.noreply.github.com> * update toml and docs * update docs * clarify robot * update docs * update docs * update pinocchio deps * final touches * Update docs/source/unitree_g1.mdx Co-authored-by: Copilot <175728472+Copilot@users.noreply.github.com> Signed-off-by: Martino Russi <77496684+nepyope@users.noreply.github.com> * move envhub dependencies to docs * point to unitree_sdk docs * upper bound on onnx * chore(docs): small details unitree docs * chore(deps): add version pin and unitree_sdk hint --------- Signed-off-by: Martino Russi <77496684+nepyope@users.noreply.github.com> Co-authored-by: Steven Palma <steven.palma@huggingface.co>	2026-03-09 18:47:12 +01:00
Steven Palma	b0efa73520	chore(dependencies): Bump lerobot to 0.5.1 (#3118 )	2026-03-09 12:43:32 +01:00
Steven Palma	00b662de02	chore(dependencies): Bump lerobot to 0.5.0 (#3117 ) v0.5.0	2026-03-09 11:34:52 +01:00
Steven Palma	5c51a74484	chore(deps): update requirements file (#3114 )	2026-03-09 11:18:05 +01:00
Steven Palma	db8547e35d	test(cameras): skip flaky async_read test (#3106 )	2026-03-08 14:02:33 +01:00
Steven Palma	c17d949531	chore(readme): update citation with ICLR26 paper (#3107 ) * peer reviewed citation 🎉 Signed-off-by: Francesco Capuano <74058581+fracapuano@users.noreply.github.com> * add iclr year Signed-off-by: Francesco Capuano <74058581+fracapuano@users.noreply.github.com> * fix quentin's spelling name Signed-off-by: Francesco Capuano <74058581+fracapuano@users.noreply.github.com> * docs(readme): update citation --------- Signed-off-by: Francesco Capuano <74058581+fracapuano@users.noreply.github.com> Co-authored-by: Francesco Capuano <74058581+fracapuano@users.noreply.github.com>	2026-03-08 14:01:43 +01:00
Steven Palma	1e131f93f8	chore(docs): add uv installation instructions (#3105 ) * chore(docs): add uv installation instructions * fix(docs): format tabs * chore(docs): small details * chore(docs): last details uv installation instructions * chore(docs): last detail --- Co-authored-by: sahilmaniyar888 <156301258+sahilmaniyar888@users.noreply.github.com>	2026-03-08 13:00:06 +01:00
Ignat Georgiev	2fb5c7add0	feat(train): add cudnn_deterministic option for reproducible training (#3102 ) Add a `cudnn_deterministic` flag to `TrainPipelineConfig` (default: False) that sets `torch.backends.cudnn.deterministic = True` and disables benchmark mode, eliminating CUDA floating-point non-determinism at the cost of ~10-20% training speed. When False (default) the existing benchmark=True behaviour is preserved.	2026-03-08 12:29:33 +01:00
Martino Russi	4f2ef024d8	feat(robots): Unitree G1 WBC implementation (#2876 ) * move locomotion from examples to robot, move controller to teleoperator class * modify teleoperate to send back actions to robot * whole body controller * add holosoma to locomotros * various updates * update joint zeroing etc * ensure safefail with locomotion * add unitree locomotion * launch camera from g1 server * publish at varying framerates * fix async read in camera * attempting to fix camera lag * test camera speedup * training * inference works * remove logging from pi0 * remove logging * push local changes * testing * final changes * revert control_utils * revert utils * revert * revert g1 * revert again: * revert utils * push recents * remove examples * remove junk * remove mjlog * revergt edit_dataset * Update lerobot_edit_dataset.py Signed-off-by: Martino Russi <77496684+nepyope@users.noreply.github.com> * undo teleop changes * revert logging * remove loggings * remove loogs * revert dataset tools * Update dataset_tools.py Signed-off-by: Martino Russi <77496684+nepyope@users.noreply.github.com> * move gravity to utils * revert changes * remove matplotlib viewer (rerun works fine) * factory revert * send policy action directly * recent changes * implement flexible action space * send empty command if arms are missing * rename locomotion to controller * add init * implement feedback * add feedback for teleoperator * fix ruff * fix ruff * use read_latest * fix zmq camera * revert exo_serial * simplify PR * revert exo_changes * revert camera_zmq * Update camera_zmq.py Signed-off-by: Martino Russi <77496684+nepyope@users.noreply.github.com> * remove frame duplication from zmq server * revert channerfactoryinitialize * keep channelfactoryinitialize * remove zeroing out logic * fix typo * refactor teleop class * simplify teleop further * import armindex at the top * fix visualizer again * revert ik helper * push stuff * simplify image_server * update image_server * asd * add threading logic * simplify ik helper stuff * simplify holosoma * fix names * fix docs * revert leg override * clean connect * fix controller * fix ruff * clean teleoperator * set_from_wireless * avoid double initializations * refactor robot class * fix pre-commit * update docs * update docs format * add teleop instructions * unitree_g1 specific exception in record/teleoperate * add thumbnail to docs * add thumbnail to doc * refactor(unitree): multiple improvements (#3103) * refactor(unitree): multiple improvements * test(unitree): added tests + improved installation instructions * refactor(robots): minor changes unitree robot kinematic * chore(robots): rename g1 kinematics file --------- Signed-off-by: Martino Russi <77496684+nepyope@users.noreply.github.com> Signed-off-by: Steven Palma <imstevenpmwork@ieee.org> Co-authored-by: Steven Palma <imstevenpmwork@ieee.org> Co-authored-by: Steven Palma <steven.palma@huggingface.co>	2026-03-08 11:33:24 +01:00
Shun.Sasaki	6139b133ca	fix(async_inference): restore robot module imports in robot_client.py (#3081 )	2026-03-06 17:14:14 +01:00
Steven Palma	85de893fa7	fix(ci): skip HF log in (and tests) in forks and community PRs (#3097 ) * fix(ci): skip HF log in (and tests) in forks and community PRs * chore(test): remove comment about test meant to be only run locally * fix(tests): no hf log in decorator for xvla * fix(test): no decorator in yield	2026-03-06 16:33:43 +01:00
Steven Palma	a4c66e530b	chore(docs): remove pi installation note (#3095 )	2026-03-06 15:52:54 +01:00
Steven Palma	a225127527	chore(dependencies): sync intelrealsense + added notes (#3094 )	2026-03-06 10:50:46 +01:00
Steven Palma	e489ba24fc	feat(dependencies): require Python 3.12+ as minimum version (#3023 ) * feat(dependecies): upgrade to python3.12 * fix(test): processor regex message * fix(test): processor regex message * fix(dependecies): resolve all tags in python 3.12 * fix(dependecies): add more hints to faster resolve * chore(dependecies): remove cli tag huggingface-hub dep * refactor(policy): update eagle for python3.12 * chore(docs): update policy creation for python 3.12 * chore(test): skip failing tests in macos	2026-03-06 10:15:13 +01:00
Steven Palma	d324ffe810	fix(ci): test only multi-gpu tests in multi-gpu runner (#3092 )	2026-03-05 19:53:40 +01:00
Pepijn	1a24f770d3	Feat/slurm compute rabc script (#3041 ) * Add SLURM SARM progress annotation script. Provide a standalone two-stage compute/aggregate pipeline for RA-BC progress generation so large datasets can be processed in parallel and optionally uploaded to the Hub. Made-with: Cursor * fix pr comments * remove comments	2026-03-05 18:27:58 +01:00
Caroline Pascal	92fba37225	fix(num_frames): fixing redundant frames count in conversion script (#3091 )	2026-03-05 15:49:50 +01:00
Steven Palma	3e45120272	fix(ci): log in HF for gated repo in nightly workflows (#3089 ) * fix(ci): log in HF for gated repo in nightly workflows * fix(ci): add env var * fix(ci): remove 10 min limit for multi-gpu nightly	2026-03-05 13:22:37 +01:00
Steven Palma	f0d2b37beb	chore(dependencies): bump transformers v5 (#2964 ) * chore(dependencies): upgrade transformers + hggingface-hub + peft + scipy * chore(dependencies): bump pi0 family to transformers v5 * chore(dependencies): bump wall x to transformers v5 * chore(dependencies): bump gr00t to transformers v5 * chore(style): fix pre-commit * fix(policy): xvla forced_bos_token missing * test(rl): skip ci tests for resnet10 * Fix: full pi models support for transformer v5 (#2967) * fix(pi): remove loss truncation * fix(pi): remove state padding before tokenization * fix(pi): fix image padding value * fix from_pretrain * add transformer v5 changes * remove reference * more fixes * make it work * add support for rest of pi family * add pifast work * more changes * more changes * more cleanup * fix torch params * dtype fix * torch compile * embed mismatch fix * revert groot * more nit fixes * remove unused classes * more fixes * revert * nit * torch dtype warning fix * but back dynamic renaming * add tie embedding --------- Co-authored-by: Yufei Sun <skieyfly@gmail.com> * chore: fix XVLA in transformers v5 (#3006) * test(policies): enable wall x CI testing * style(test): pre-commit check * style(test): pre-commit * fix wall x for transformer v5 (#3008) * tv5 fix * various wall x fixes * Delete tests/policies/pi0_pi05/print_pi05_output_logits.py Signed-off-by: Jade Choghari <chogharijade@gmail.com> * sync modeling_florence2.py with chore/bump_transformers_v5 * more * more fixes * more * remove comment * more --------- Signed-off-by: Jade Choghari <chogharijade@gmail.com> * chore(dependencies): adjust dependencies versioning after transformers v5 (#3034) * chore(dependecies): adjust dependecies versioning after transformers v5 * fix(policies): remove deprecated input_embeds * fix(policies): dict _tied_weights_keys * chore(depedencies): common qwen-vl-utils * chore(dependencies): bump transformers to 5.2 * Fix policy testing for tv5 (#3032) * fix ci logger * other fix * fix mypy * change logits to torch2.10 * skip wallx\| * remove logging --------- Co-authored-by: Steven Palma <imstevenpmwork@ieee.org> * feat(ci): log into HF to unblock some CI tests (#3007) * feat(ci): log into HF to unblock some CI tests * chore(ci): change hf call + secret name * fix(ci): temp fix for pi0 rtc test * test(policies): require_cuda for unblocked tests * test(policies): require_cuda wall_x * fic(tests): require_cuda outter most for pi0 * fix(test): return instead of yield --------- Signed-off-by: Steven Palma <imstevenpmwork@ieee.org> * style(test): fix pre-commit * chore(deps): upgrade transformers (#3050) * chore(test): use lerobot model * fix(policies): change default action tokenizer for wall x * sample on cpu * Revert "Merge branch 'chore/bump_transformers_v5' of https://github.com/huggingface/lerobot into chore/bump_transformers_v5" This reverts commit `d9b76755f7`, reversing changes made to `89359cb0b6`. * Reapply "Merge branch 'chore/bump_transformers_v5' of https://github.com/huggingface/lerobot into chore/bump_transformers_v5" This reverts commit `c9914db78b`. --------- Signed-off-by: Jade Choghari <chogharijade@gmail.com> Signed-off-by: Steven Palma <imstevenpmwork@ieee.org> Co-authored-by: Jade Choghari <chogharijade@gmail.com> Co-authored-by: Yufei Sun <skieyfly@gmail.com> Co-authored-by: Pepijn <pepijn@huggingface.co>	2026-03-05 09:25:26 +01:00
Caroline Pascal	cbc8bfb2e6	chore(docstrings): updating v2.1-v3.0 conversion script docstrings to match the new task label (#3077 ) * chore(docstrings): updating v2.1-v3.0 conversion script docstrings to match the new task label * chore(task): renamming the default index label in the tasks DataFrame to task * Revert "chore(docstrings): updating v2.1-v3.0 conversion script docstrings to match the new task label" This reverts commit f55de3255278f23f18b5d955565f6768d094951d. * chore(docstrings): updating docstrings to match dataset v3.0 architecture * chore(format): formatting code	2026-03-04 17:59:03 +01:00
Paul Crook	0d1be72dc8	Fixing metadata indexing when writing new Parquet file (#2941 ) * Fixing metadata indexing when writing new Parquet file Summary: - addressing this issue: https://github.com/huggingface/lerobot/issues/2401 - vibe-coded bugfix by Claude Sonnet 4.5 * Backing out changes to convert_videos_of_camera * Addressing Ruff pre-commit complaint Summary: - addressing "SIM113 Use `enumerate()` for index variable `ep_idx` in `for` loop" --------- Co-authored-by: Paul <238953601+pac-robotics@users.noreply.github.com>	2026-03-04 16:53:34 +01:00
Maxime Ellerbach	96b7c212c4	chore(docs): updating deprecated huggingface-cli to hf (#3071 ) * chore(docs): updating deprecated huggingface-cli to hf * small typo in my-org	2026-03-04 15:08:49 +01:00
Caroline Pascal	4303b3c930	chore(root): fixing `root` semantics in convert_dataset script (#3073 ) * fix(root): fixing root semantincs in convert_dataset script * fix(\): fixing command syntax in dataset conversion script Signed-off-by: Caroline Pascal <caroline8.pascal@gmail.com> --------- Signed-off-by: Caroline Pascal <caroline8.pascal@gmail.com>	2026-03-04 11:11:21 +01:00
Caroline Pascal	63dca86df8	fix(dataset edit tools): clarifying `root` argument usage + adding related features (#3049 ) * fix(root): adding proper support for the root and new_root arguments * feat(roots): adding a roots agrument for the merge operation * chore(clean): cleaning up code * chore(doctrings): updating doctrings with new features * fix(repo_id): setting repo_id to None when not needed * fix(roots/repo_ids): making mypy happy by using repo_ids and roots for merge operation * fix(path): fixing path related issues * fix(repo_id): fixing issues related to repo_id * chore(doctrings): updating docstrings + fix typo * chore(clean): cleaning code * fix(split new_repo_id): reverting new_repo_id addition for split operation * docs(dosctrings): completing docstrings * fix(repo_ids/roots): improving checks for repo_ids/roots lengths * fix(repo_ids): making repo_ids optional in MergeConfig but raise if not given * fix(docstrings): fixing docstrings for split operation * fix(hints): updating get_output_path hints to accept paths as strings too * fix(y/N prompts): removing y/N prompts in lerobot_edit_dataset * fix(merge repo_id): fixing merge operation to use new_repo_id instead of repo_id * fix(typo): fixing typo in doctrings	2026-03-03 15:40:46 +01:00
Caroline Pascal	8a0cc3d664	fix(frame_index): making rerun's "frame_index" timeline compatible with behaviour1k datasets (#3068 ) * fix(frame_index): making rerun's "frame_index" timeline compatible with behaviour1k datasets * fix(segfault risk): removing segfault risk by calling batch["index"] in the dataloader loop	2026-03-03 11:55:09 +01:00
Bernie Telles	8bb8ed4803	Improve policy_device documentation for async.mdx (#3060 )	2026-03-02 15:35:15 +01:00
Steven Palma	095856b06a	chore: add AI policy (#3055 )	2026-02-28 14:41:28 +01:00
Steven Palma	563f42bdb1	chore(dependencies): Bump lerobot to 0.4.5 (#3051 )	2026-02-27 19:29:35 +01:00
Caroline Pascal	8fff0fde7c	chore(docstrings): fixing deprecated `root` argument description in LeRobotDataset class (#3035 ) * chore(docstrings): fixing deprecated `root` argument docstrings in LeRobotDataset class * chore(draccus): updating draccus CLI help * chore(revert): reverting changes in lerobot_dataset_viz.py --------- Co-authored-by: Steven Palma <imstevenpmwork@ieee.org> v0.4.4	2026-02-27 18:22:44 +01:00
Pepijn	04de496547	fix(logging): avoid double-counting samples across processes (#3045 )	2026-02-27 17:45:19 +01:00
Khalil Meftah	baf9b50365	Fix(diffusion): enforce no-crop behavior when crop_ratio=1.0 (#3046 ) * refactor(diffusion): replace crop_shape with resize_shape and crop_ratio * fix(diffusion): address review feedback on resize/crop backward compat * test: regenerate diffusion artifacts for updated default config * fix: disable crop when resize path uses crop_ratio=1.0 --------- Co-authored-by: starlitxiling <1754165401@qq.com>	2026-02-27 17:44:53 +01:00
Jade Choghari	a0fdbf037a	feat(policies): add Smolvla torch compile support (#3043 ) * Change LIBERO init_state_id when reset. Signed-off-by: Aoqun Jin <aojiaojiao@foxmail.com> * Change LIBERO init_state_id when reset. Signed-off-by: Aoqun Jin <aojiaojiao@foxmail.com> * pre-commit run * Add torch.compile for smolvla Signed-off-by: Aoqun Jin <aojiaojiao@foxmail.com> * Add torch.compile for smolvla Add model compilation option for improved performance. Signed-off-by: Aoqun Jin <aojiaojiao@foxmail.com> * first --------- Signed-off-by: Aoqun Jin <aojiaojiao@foxmail.com> Co-authored-by: Aoqun Jin <aojiaojiao@foxmail.com> Co-authored-by: Steven Palma <imstevenpmwork@ieee.org>	2026-02-27 18:58:36 +03:00
Khalil Meftah	c085531b17	fix: add missing openarm_mini import to CLI scripts (#3028 )	2026-02-27 15:46:31 +01:00

1 2 3 4 5 ...

1339 Commits