zai-org/SCAIL-2
Modelo de geração de vídeo em alta no Hugging Face — 0 downloads e 252 curtidas da comunidade.
Papers, modelos e datasets em alta no Hugging Face, além do blog oficial — com leitura editorial em português.
Modelo de geração de vídeo em alta no Hugging Face — 0 downloads e 252 curtidas da comunidade.
MeshFlow generates triangle meshes directly using equivariant optimal-transport flow matching models with improved inference speed over autoregressive methods.
VESFlow is a training-free safety method for flow matching-based text-to-image generation that edits velocity fields to ensure safe output while maintaining prompt integrity.
Lift4D presents a test-time optimization framework that combines temporal consistency from single-view 3D reconstruction with deformable 3D Gaussian Splatting and view-conditioned…
FedOT is a novel framework that enables ownership verification and leakage tracing in federated latent diffusion models by introducing chunked watermarking and latent vector transf…
Text-to-image models are enhanced with controlled diversity through semantic browsing capabilities that enable structured navigation of image variations based on meaningful semanti…
ABACUS is a unified vision-language model that performs object counting and related tasks through innovative spatial grounding, boundary-aware counting policies, and self-critical…
Vera is a layered diffusion framework that preserves video content during editing by generating edit layers and alpha mattes through a Mixture-of-Transformers architecture.
Um novo paper gera ilusões visuais tridimensionais de forma rápida e zero-shot. O resultado é menos um truque de festa do que um diagnóstico do estado da arte em síntese 3D.
Dataset em destaque no Hugging Face — 3.4 mil downloads. ISCSLP 2026 CoT-TTS Dataset Dataset Overview This dataset is prepared for the ISCSLP 2026 CoT-TTS Challenge and is designed to support research on con…
BioInsight is a multi-agent system that transforms static biomedical reports into interactive, evidence-centered interfaces by organizing disease-specific evidence through structur…
UnityShots is a memory-driven audio-video generation system that maintains consistent subject appearance and audio across video cuts using fixed-size long-term and short-term memor…