Flex-π: Multi-Stream World-Action Model

Loading video
Loading videoJointly predicts RGB, 3D pointmaps, DINO semantics with actions; one checkpoint deploys as VLA, full WAM, or anything in between.
Category: research
Author: @GeYan_21
Date: 2026-08-12T00:00:00
Duration: 50.466s
Reference: https://github.com/geyan21/flex-pi





