// radar de ia

Multimodal

Papers, modelos e datasets em alta no Hugging Face, além do blog oficial — com leitura editorial em português.

Hollywood wants Seedance banned and reportedly also wants to keep using it
Blog LLMs & Texto

Hollywood wants Seedance banned and reportedly also wants to keep using it

Bytedance's AI video tool Seedance is dividing Hollywood. A viral clip featuring AI-generated Brad Pitt and Tom Cruise prompted the Motion Picture Association's first-ever cease-and-desist against an AI company. But behind the scenes, studios are quietly using the tool on a "don't ask, don't tell" basis, says Simpsons animation producer Joel Kuwahara. The article Hollywood wants Seedance banned and reportedly also wants to keep using it appeared first on The Decoder .

05.07.2026
Blog Robótica & RL

Orientação de Segurança Neuro-Simbólica para Modelos de Visão-Linguagem-Ação via Correspondência de Fluxo Restrita

arXiv:2607.01378v1 Tipo de Anúncio: novo Resumo: Modelos de Visão-Linguagem-Ação (VLA) têm demonstrado capacidades promissoras de generalização em tarefas de manipulação robótica, mas sua implantação no mundo real permanece limitada pela falta de medidas de segurança eficazes. Especificamente, as medidas de segurança existentes apenas evitam colisões causadas pela próxima ação do robô. Neste artigo, propomos um mecanismo de orientação de segurança neuro-simbólico para VLAs baseados em correspondência de fluxo que possibilita a prevenção preditiva de colisões...

03.07.2026
Blog LLMs & Texto

PairCoder++: Pair Programming as a Universal Paradigm for Verified Code-Driven Multimodal and Structured-Artifact Generation

arXiv:2607.01883v1 Announce Type: new Abstract: Code is the medium through which large language models generate structured artifacts: charts, scientific figures, vector graphics, CAD models, 3D scenes, and hardware designs are all produced by writing programs. In this regime single pass inference is brittle, because the compiler, renderer, or simulator that decides whether the artifact exists is invisible to the model. We present PairCoder, which grounds review in the toolchain and realizes it a...

03.07.2026
Blog Robótica & RL

Guided Action Flow: Q-Guided Inference for Flow-Matching Vision-Language-Action Policies

arXiv:2607.02092v1 Announce Type: new Abstract: Flow-matching vision-language-action policies generate robot action chunks through an iterative transport process, creating an opportunity for test-time guidance without retraining the base policy. We study this opportunity in Guided Action Flow, an inference-time framework that keeps a pretrained SmolVLA policy frozen and uses a learned action-chunk critic to guide its reverse-time flow sampler. The critic is trained from real success and failure ...

03.07.2026
Blog Robótica & RL

Bridge-WA: Predicting Where and How the World Changes for Robotic Action

arXiv:2607.02195v1 Announce Type: new Abstract: General-purpose vision-language-action models benefit from large vision-language priors, but effective manipulation also requires anticipating action-relevant scene changes. Existing world-action models often rely on large generative world models or dense future rollouts, which are expensive and spend capacity on visual details weakly coupled to control. We present Bridge-WA, a lightweight world-action framework that distills a frozen future-change...

03.07.2026
467 itens no radar