FLUX 3 Action is an open weights 7B world action model. It takes camera frames, the robot's state and a text instruction, and returns the next chunk of actions, denoised together with the next video frames. This repository is part of the FLUX 3 Action collection.
For more information, read the documentation.
SO-101 policy with checkpoint-owned observation history, camera layout and normalization. See License for usage terms.
from lerobot.policies.flux3 import Flux3Policy
from lerobot.policies.factory import make_pre_post_processors
repo_id = "black-forest-labs/flux-3-action-so101"
policy = Flux3Policy.from_pretrained(repo_id)
preprocessor, postprocessor = make_pre_post_processors(policy.config, pretrained_path=repo_id)
Use the FLUX3 integration with shared-encoder Hub subfolder/revision support. Both frozen encoders load automatically from the pinned base repository. This repository contains no duplicated encoder weights.
Two cameras: observation.images.scene and observation.images.wrist, in that order.
Six state/action dimensions; joint-delta actions with absolute gripper.
Eight observations, two visual snapshots and past-command conditioning.
Predict 42 actions, execute 32 at 30 Hz, then replan. Saved processor pipelines
and normalization state files are required and load from this repository.
Sampling: four Euler steps, shift 6.93, guidance 3, seed 42, BF16.
The checkpoint config and dataset_statistics.json contain the full contract.
Use the included lora.json recipe with your dataset:
hf download black-forest-labs/flux-3-action-so101 lora.json --local-dir models/so101
lerobot-train --config_path=models/so101/lora.json \
--policy.path=black-forest-labs/flux-3-action-so101 \
--dataset.repo_id=YOUR_DATASET --output_dir=outputs/so101-lora --job_name=so101-lora
The recipe inherits the robot/model settings from this checkpoint. It trains a rank-32 LoRA with bf16 mixed precision, batch size 2, gradient accumulation 4 and 10,000 steps.
The policy was trained on the SO-101 episodes of lerobot/community_dataset_v3.
FLUX 3 Action outputs joint targets. Nothing in the model bounds joint velocity, force or workspace; the application must enforce those limits and keep a hardware stop within reach. Validate on a simulator or with the arm's safety limits engaged before running near people.
The model and its derivatives may not be used:
Nothing contained in this model card should be interpreted as or deemed a restriction or modification to the license the model is released under.
Black Forest Labs is committed to responsible model development and deployment. FLUX 3 Action outputs motor commands and, on request, predicted camera frames of the scene it is acting in. For information about our mitigations, evaluation processes and policies, see Capable, Open, and Safe: Combating AI Misuse. To report safety concerns, contact safety@blackforestlabs.ai.
This model falls under the FLUX Kommunity License v.1.0. The text encoder in flux-3-action-base is an unmodified copy of Qwen3-VL-4B-Instruct under Apache-2.0. The code in flux-action has its own license.
This project may contain trademarks or logos for projects, products, or services. Use of Black Forest Labs and FLUX trademarks or logos in modified versions of this project must not cause confusion or imply sponsorship or endorsement. Any use of third-party trademarks, intellectual property or logos are subject to those third-party's policies.
FLUX 3 Action is an open weights 7B world action model. It takes camera frames, the robot's state and a text instruction, and returns the next chunk of actions, denoised together with the next video frames. This repository is part of the FLUX 3 Action collection.
For more information, read the documentation.
SO-101 policy with checkpoint-owned observation history, camera layout and normalization. See License for usage terms.
from lerobot.policies.flux3 import Flux3Policy
from lerobot.policies.factory import make_pre_post_processors
repo_id = "black-forest-labs/flux-3-action-so101"
policy = Flux3Policy.from_pretrained(repo_id)
preprocessor, postprocessor = make_pre_post_processors(policy.config, pretrained_path=repo_id)
Use the FLUX3 integration with shared-encoder Hub subfolder/revision support. Both frozen encoders load automatically from the pinned base repository. This repository contains no duplicated encoder weights.
Two cameras: observation.images.scene and observation.images.wrist, in that order.
Six state/action dimensions; joint-delta actions with absolute gripper.
Eight observations, two visual snapshots and past-command conditioning.
Predict 42 actions, execute 32 at 30 Hz, then replan. Saved processor pipelines
and normalization state files are required and load from this repository.
Sampling: four Euler steps, shift 6.93, guidance 3, seed 42, BF16.
The checkpoint config and dataset_statistics.json contain the full contract.
Use the included lora.json recipe with your dataset:
hf download black-forest-labs/flux-3-action-so101 lora.json --local-dir models/so101
lerobot-train --config_path=models/so101/lora.json \
--policy.path=black-forest-labs/flux-3-action-so101 \
--dataset.repo_id=YOUR_DATASET --output_dir=outputs/so101-lora --job_name=so101-lora
The recipe inherits the robot/model settings from this checkpoint. It trains a rank-32 LoRA with bf16 mixed precision, batch size 2, gradient accumulation 4 and 10,000 steps.
The policy was trained on the SO-101 episodes of lerobot/community_dataset_v3.
FLUX 3 Action outputs joint targets. Nothing in the model bounds joint velocity, force or workspace; the application must enforce those limits and keep a hardware stop within reach. Validate on a simulator or with the arm's safety limits engaged before running near people.
The model and its derivatives may not be used:
Nothing contained in this model card should be interpreted as or deemed a restriction or modification to the license the model is released under.
Black Forest Labs is committed to responsible model development and deployment. FLUX 3 Action outputs motor commands and, on request, predicted camera frames of the scene it is acting in. For information about our mitigations, evaluation processes and policies, see Capable, Open, and Safe: Combating AI Misuse. To report safety concerns, contact safety@blackforestlabs.ai.
This model falls under the FLUX Kommunity License v.1.0. The text encoder in flux-3-action-base is an unmodified copy of Qwen3-VL-4B-Instruct under Apache-2.0. The code in flux-action has its own license.
This project may contain trademarks or logos for projects, products, or services. Use of Black Forest Labs and FLUX trademarks or logos in modified versions of this project must not cause confusion or imply sponsorship or endorsement. Any use of third-party trademarks, intellectual property or logos are subject to those third-party's policies.