FLUX 3 Action removes the usual trade-off between the higher success rates of WAMs and the speed of VLAs.
Its single-step 7B checkpoint outperforms every other open policy on RoboLab, while processing each second of robot motion 1.45× to 1.66× faster than Pi0.5 (the strongest open VLA model).
It retains the core WAM architecture, predicting video and actions jointly, but covers a longer 2.13-second action horizon compared with 1 second for Pi0.5. When latency matters less, our guidance-distilled checkpoint raises the state-of-the-art success rate on RoboLab while running 2.85× to 3.15× faster than the previous leading open WAM.
