Robotics
LeRobot
Safetensors
droid
franka
flux
world-action-model
world-model
robot-learning
black-forest-labs
flux-3
Instructions to use black-forest-labs/flux-3-action-droid with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- LeRobot
How to use black-forest-labs/flux-3-action-droid with LeRobot:
# No code snippets available yet for this library. # To use this model, check the repository files and the library's documentation. # Want to help? PRs adding snippets are welcome at: # https://github.com/huggingface/huggingface.js
- Notebooks
- Google Colab
- Kaggle
maanavdalal commited on
Commit ·
3d0887b
1
Parent(s): 4c8386b
Update FLUX 3 Action model card
Browse files
README.md
CHANGED
|
@@ -16,17 +16,17 @@ FLUX 3 Action is an open weights 7B world action model. It takes camera frames,
|
|
| 16 |
|
| 17 |
For more information, read the [documentation](https://docs.bfl.ai/flux_3/flux3_action_overview).
|
| 18 |
|
| 19 |
-
Fine-tuned on [DROID](https://droid-dataset.github.io/), FLUX 3 Action places first on the RoboLab-120 benchmark at 42.
|
| 20 |
|
| 21 |
## Evaluation
|
| 22 |
|
| 23 |
-
RoboLab-120 is 120 tabletop tasks in Isaac Sim, 10 trials each, on a DROID-style Franka setup; a trial succeeds only if the task is completed as instructed. The full board is on the [RoboLab leaderboard](https://research.nvidia.com/labs/srl/projects/robolab/leaderboard.html).
|
| 24 |
|
| 25 |
-
| Model | Type | Success | Parameters |
|
| 26 |
-
| --- | --- | --- | --- |
|
| 27 |
-
| FLUX 3 Action | WAM | 42.
|
| 28 |
-
| Cosmos3-Nano-Policy | WAM | 36.8% | 16B |
|
| 29 |
-
| π0.5 | VLA | 28.0% | 3.3B |
|
| 30 |
|
| 31 |
WAM: predicts future frames and actions together. VLA: a vision-language model that outputs actions directly.
|
| 32 |
|
|
|
|
| 16 |
|
| 17 |
For more information, read the [documentation](https://docs.bfl.ai/flux_3/flux3_action_overview).
|
| 18 |
|
| 19 |
+
Fine-tuned on [DROID](https://droid-dataset.github.io/), FLUX 3 Action places first on the RoboLab-120 benchmark at 42.92% task success.
|
| 20 |
|
| 21 |
## Evaluation
|
| 22 |
|
| 23 |
+
RoboLab-120 is 120 tabletop tasks in Isaac Sim, 10 trials each, on a DROID-style Franka setup; a trial succeeds only if the task is completed as instructed. The full board is on the [RoboLab leaderboard](https://research.nvidia.com/labs/srl/projects/robolab/leaderboard.html).
|
| 24 |
|
| 25 |
+
| Model | Type | Success | Parameters |
|
| 26 |
+
| --- | --- | --- | --- |
|
| 27 |
+
| FLUX 3 Action | WAM | 42.92% | 7B |
|
| 28 |
+
| Cosmos3-Nano-Policy | WAM | 36.8% | 16B |
|
| 29 |
+
| π0.5 | VLA | 28.0% | 3.3B |
|
| 30 |
|
| 31 |
WAM: predicts future frames and actions together. VLA: a vision-language model that outputs actions directly.
|
| 32 |
|