Robuta

https://yliu-cs.github.io/MMaDA-VLA/ MMaDA-VLA: Large Diffusion Vision-Language-Action Model vlalargediffusionvisionlanguage