OpenCoF: Learning to Reason Through Video Generation
OpenCoF framework introduces a reasoning video dataset and model that improve temporal reasoning through diverse supervision and explicit reasoning tokens for visual and textual cu…
Hugging Face · Daily Papers
·Xinyan Chen, Ziyu Guo
·
·▲ 17 upvotes
Este artigo está em destaque na seleção diária de papers do Hugging Face, curada pela comunidade de pesquisa em IA.
Autores: Xinyan Chen, Ziyu Guo, Renrui Zhang, Dongzhi Jiang, Hongsheng Li
- 17 upvotes da comunidade
- Temas: Chain-of-Frame, video generation models, OpenCoF-17K dataset, Wan-CoF model, temporal supervision, visual reasoning tokens
Resumo
Resumo original (em inglês), extraído do paper:
OpenCoF framework introduces a reasoning video dataset and model that improve temporal reasoning through diverse supervision and explicit reasoning tokens for visual and textual cues.Onde ler
// relacionados
Leia também
Blog
Flight attendants freaked out that Google is buying tons of Spirit employee data
Blog
Anthropic says any lab can now let a language model agent run the whole protein design stack
Blog
Valid Per-Field Selective Risk Control for Document Extraction: Three Failure Modes, a Validity Ladder, and When Conditioning Pays
Blog