Blog Robótica & RL Multimodal

Guided Action Flow: Q-Guided Inference for Flow-Matching Vision-Language-Action Policies

arXiv:2607.02092v1 Announce Type: new Abstract: Flow-matching vision-language-action policies generate robot action chunks through an iterative transport process, creating an opportunity for test-time guidance without retraining the base policy. We study this opportunity in Guided Action Flow, an inference-time framework that keeps a pretrained SmolVLA policy frozen and uses a learned action-chunk critic to guide its reverse-time flow sampler. The critic is trained from real success and failure ...

arXiv cs.RO ·Liuhaichen Yang, Zhuang Jiang, Chenchao Sheng, Zezhi Tang · 03 de janeiro de 2026

Ver no Hugging Face

// relacionados

Guided Action Flow: Q-Guided Inference for Flow-Matching Vision-Language-Action Policies

Leia também

UWORLD U1: a UBTECH lança o primeiro humanoide "ultra-biônico" em série — e a dança que expôs os limites

Takeda fecha acordo de US$ 600 milhões com a Insilico para descoberta de medicamentos com IA

Conheça o WebBrain: um agente de navegador com IA de código aberto e local-first que lê páginas e automatiza tarefas no Chrome e no Firefox

CoRe: Recompensas Combinadas com Feedback de Modelo de Visão-Linguagem para Aprendizado por Reforço Alinhado a Preferências