// radar de ia

Robótica & RL

Papers, modelos e datasets em alta no Hugging Face, além do blog oficial — com leitura editorial em português.

Google Deepmind unveils Gemini Robotics 2 to power robots of all shapes from tabletop arms to humanoids
Blog Robótica & RL

Google Deepmind unveils Gemini Robotics 2 to power robots of all shapes from tabletop arms to humanoids

Google Deepmind's Gemini Robotics 2 is its most advanced vision-language-action model yet, built to control everything from tabletop robots to full-body humanoids. Gemini Robotics ER 2 adds a higher-level reasoning layer for robotics tasks. The article Google Deepmind unveils Gemini Robotics 2 to power robots of all shapes from tabletop arms to humanoids appeared first on The Decoder .

31.07.2026
Anthropic follows OpenAI in admitting its Claude models reached out of test environments and attacked real-world systems
Blog Robótica & RL

Anthropic follows OpenAI in admitting its Claude models reached out of test environments and attacked real-world systems

Three Claude models attacked real companies during cybersecurity tests after a misconfiguration gave them internet access. One published malware on PyPI that infected 15 systems. Another kept attacking after recognizing its target was real. Anthropic calls it an operational error. The article Anthropic follows OpenAI in admitting its Claude models reached out of test environments and attacked real-world systems appeared first on The Decoder .

31.07.2026
Aschenbrenner's AI thesis could be correct, his timing and leverage were not
Blog Dados & Embeddings

Aschenbrenner's AI thesis could be correct, his timing and leverage were not

Leopold Aschenbrenner's AI hedge fund Situational Awareness had to unload nearly its entire publicly traded portfolio to Ken Griffin's Citadel after racking up heavy losses on leveraged AI stock positions. Just days earlier, Aschenbrenner had reported a six-month return of 439 percent and pulled in fresh capital. Then margin calls forced the fire sale. The article Aschenbrenner's AI thesis could be correct, his timing and leverage were not appeared first on The Decoder .

31.07.2026
Blog Robótica & RL

JetBrains Open-Sources KotlinLLM: Smart Macros That Generate Kotlin Source Code at Runtime and Hot-Reload It Through JDI

JetBrains Research has open-sourced KotlinLLM under the Apache License 2.0. The IntelliJ IDEA plugin prototype adds Smart macros, asLlm and mockLlm, whose bodies are generated Kotlin source rather than live model calls. The plugin captures runtime values through JDI, asks an LLM agent for a narrow code update, compiles it, and redefines the loaded class. Covered scenarios then run as plain Kotlin with no further inference call. On an adapted Spring Petclinic project, 24 of 24 scenarios completed...

31.07.2026
Building a Policy-Governed Multi-Agent Financial Research Workflow with Omnigent
Blog Robótica & RL

Building a Policy-Governed Multi-Agent Financial Research Workflow with Omnigent

In this tutorial, we demonstrate how to build and execute a multi-agent workflow with Omnigent in a secure, isolated Python environment. Learn to integrate live exchange-rate data, implement hierarchical agent delegation for financial text auditing, and apply hard governance policies—such as cost budgets and tool call limits—to your research pipeline directly from Google Colab. The post Building a Policy-Governed Multi-Agent Financial Research Workflow with Omnigent appeared first on MarkTec...

31.07.2026
Blog LLMs & Texto

BridgeAlign: Bridging Preference Alignment for Humanities and Social Sciences

arXiv:2607.27366v1 Announce Type: new Abstract: While data synthesis for large language models (LLMs) is prevalent, it primarily targets domains with verifiable answers, overlooking open-ended humanities and social sciences (HSS), where nuanced quality judgments matter more than objective correctness. This makes preference alignment a natural paradigm for broad HSS tasks. Yet existing methods are either costly or not tailored to broad HSS disciplines. We thus propose BridgeAlign, among the first...

31.07.2026
Blog LLMs & Texto

Prompt Chaining in Practice: A Case Study in Automated Scholarly Report Generation

arXiv:2607.27210v1 Announce Type: new Abstract: The exponential growth of scholarly publications requires automated tools for effective information synthesis. However, simple, single-shot prompting methods often lack the reliability and quality required for complex synthesis tasks. This paper introduces and empirically evaluates a multi-stage prompt chaining methodology as a more reliable architectural pattern for such tasks. This approach is implemented in our system, AI SciBrief, which automat...

31.07.2026
1635 itens no radar