VIRENA
A minimal Vision-Language-Action model you can read: frozen CLIP + a tiny head on ManiSkill PickCube. LeRobot integration. Runs on a Mac, no GPU.
File Explorer
- results.csv
- results.json
- RESULTS.md
- banner.png
- basecam_01_start.png
- basecam_02_approach.png
- basecam_03_grasp.png
- demo.gif
- so101_pred_vs_true.png
- so101_rollout.gif
- wristcam_01_start.png
- wristcam_02_approach.png
- wristcam_03_grasp.png
- __init__.py
- base_collector.py
- maniskill_collector.py
- __init__.py
- camera_config.py
- maniskill_wrapper.py
- __init__.py
- base_policy.py
- random_policy.py
- scripted_expert.py
- __init__.py
- episode_recorder.py
- __init__.py
- build_dataset.py
- collect.py
- export_lerobot.py
- inspect_data.py
- lerobot_schema.py
- __init__.py
- lerobot_reader.py
- virena_dataset.py
- __init__.py
- probe.py
- rollout.py
- tasks.py
- verdict.py
- ARCHITECTURE.md
- __init__.py
- head.py
- __init__.py
- encoder.py
- __init__.py
- encoder.py
- pool.py
- preprocess.py
- __init__.py
- vla.py
- vla_policy.py
- pickcube.pt
- ablation.py
- diagnose.py
- diagnose_sim.py
- eval_lerobot.py
- eval_sim.py
- predict.py
- record_demo.py
- test_dataloader.py
- test_lerobot_export.py
- test_lerobot_reader.py
- test_vision_encoder.py
- test_vla.py
- train.py
- .gitignore
- config.py
- LICENSE
- pyproject.toml
- README.md
- requirements-model.txt
# Use via CDN
jsDelivrjsDelivr serves any public GitHub repository as a CDN with zero setup. Pick a version and a file to get a ready-to-paste link and snippet.
Link
Example
// repository documentation
Was this content helpful?
(0 ratings)
