Patil/Krea-2-depth-controlnet
Modelo de edição de imagem — 0 downloads e 91 curtidas no Hugging Face.
Papers, modelos e datasets em alta no Hugging Face, além do blog oficial — com leitura editorial em português.
Modelo de edição de imagem — 0 downloads e 91 curtidas no Hugging Face.
Epoch AI reports a sharp rise in security vulnerability reports. In June 2026, 21 organizations reported about 1,500 high-severity and critical CVEs, more than 3.5 times the previous monthly record. The surge lines up with the launch of AI-powered bug-hunting programs. The article Security vulnerability reports have exploded since AI models started hunting for bugs appeared first on The Decoder .
Mark Zuckerberg admitiu falhas na reestruturação da empresa durante uma reunião interna com funcionários. Os agentes de IA em torno dos quais a Meta se reorganizou estão avançando mais devagar do que o planejado, disse Zuckerberg. Seu chefe de IA, por sua vez, pintou um quadro mais otimista. O artigo O avanço dos agentes de IA da Meta está mais lento do que Zuckerberg planejava apareceu primeiro em The Decoder.
A Krea publicou os pesos de um modelo de 12,9 bilhões de parâmetros capaz de gerar imagens em 2K quase instantaneamente — e colocou uma licença cheia de letras miúdas junto do presente.
arXiv:2607.01709v1 Announce Type: new Abstract: Agents are increasingly used to construct workflows and assist humans in completing recurring tasks more efficiently. As these workflows become repeated and domain-specific, agent memory and reusable skills become increasingly important: agents should be able to recall workflow patterns, execution constraints, and user preferences from previous runs. We study this problem in workflow-based image generation and introduce COMFYCLAW, an agentic skill ...
arXiv:2607.01436v1 Announce Type: new Abstract: Diffusion language models, which generate text by denoising a token canvas bidirectionally instead of emitting tokens left to right, have become competitive with autoregressive (AR) generation. Medical foundation models, however, remain almost entirely autoregressive. We adapt a mixture-of-experts diffusion language model, DiffusionGemma-26B, and benchmark it against its same-size AR sibling Gemma-4-26B under an identical LoRA recipe on medical vis...
A 3D latent diffusion model for chest CT generation that achieves high-fidelity results while enabling control over clinical attributes through metadata conditioning and reinforcem…
Vidu S1 is a real-time interactive video generation model that supports voice-controlled digital character animation with infinite-length output and high frame rate on consumer har…
Modelo de geração de imagem — 7.8 mil downloads e 53 curtidas no Hugging Face.
Frameworks like Lean Six Sigma and business process management (BPM) first gained traction because they promised clarity in the chaos—a structured way to bring order to messy, sprawling operations. Lean Six Sigma emphasized statistical rigor and quality control; BPM created end-to-end maps of how work should flow across departments. Both offered a repeatable way to…
O Google lançou em 30 de junho uma versão enxuta do Nano Banana 2 que gera imagens em quatro segundos por menos de meio centavo de dólar cada — e o Gemini Omni Flash, que transforma essas imagens em vídeo de dez segundos editável por linguagem natural.
arXiv:2607.00183v1 Announce Type: new Abstract: Adapting pre-trained text-to-image diffusion models, whether to learn new visual concepts or erase unwanted ones, is routinely evaluated on its intended effects alone. We argue this framing is incomplete. Through sparse autoencoder analysis and zero-shot classification, we demonstrate that adaptation systematically damages semantically unrelated concepts in ways that aggregate metrics structurally cannot surface: when damage is severe enough for FI...