// radar de ia

Visão Computacional

Papers, modelos e datasets em alta no Hugging Face, além do blog oficial — com leitura editorial em português.

Blog Robótica & RL

Streamlining stereo differentiable rendering for marker-free real-time tracking of surgical robots

arXiv:2607.12604v1 Announce Type: new Abstract: Purpose: Marker-based tracking of surgical robots is occlusion-prone in cluttered operating rooms. We evaluate stereo differentiable rendering for marker-free, real-time robot pose tracking, potentially improving safety, reducing setup time, and enabling multi-robot interaction. Methods: We extend the markerless pose estimation framework roboreg to online dynamic tracking via (i) sequential optimisation that propagates pose estimates across frames ...

15.07.2026
Blog LLMs & Texto

Decodificação de Resposta em Bordas Semânticas do SAM3 para Segmentação de Trincas Zero-Shot

arXiv:2607.12292v1 Tipo de anúncio: novo Resumo: A segmentação de trincas é essencial para a inspeção de infraestrutura e a avaliação da saúde estrutural, mas os métodos de alto desempenho existentes normalmente exigem anotações e treinamento em nível de pixel específicos para a tarefa. Os modelos de fundação de visão que aceitam comandos de texto possibilitam a implantação em modo zero-shot, mas suas propostas finais de máscara são pouco adequadas a trincas finas, fragmentadas e de baixo contraste, cujas evidências podem ser suprimidas, truncadas ou excessivamente expandidas durante a geração da máscara...

15.07.2026
Blog LLMs & Texto

Format Sensitivity Index: Token-Controlled Prompt Wrapper Robustness and Schema Compliance in LLM Benchmarking

arXiv:2607.09665v1 Announce Type: new Abstract: Prompt wrappers often differ only in formatting, yet they can change model scores enough to flip leaderboard conclusions. We study this variance under a token-controlled protocol and introduce two complementary metrics: the Format Sensitivity Index (FSI), the accuracy range induced by wrapper choice, and the Parseability Sensitivity Index (PSI), the corresponding range in answer parseability. Across 140,000 OpenRouter generations spanning 7 QA task...

14.07.2026
Blog LLMs & Texto

LLM-Centric Agentic AI for UAV Swarms: Architecture, Enabling Technologies, and Open Problems

arXiv:2607.09756v1 Announce Type: new Abstract: Uncrewed Aerial Vehicle (UAV) swarms have significant potential for applications such as Search and Rescue (SAR) and environmental monitoring, but their real-world deployment is limited by a lack of situational awareness, intermittent connectivity, and significant cybersecurity risks. Agentic Artificial Intelligence (AI) represents a shift from standalone Large Language Model (LLM) toward closed-loop cognitive architectures that integrate perceptio...

14.07.2026
Blog Visão Computacional

Does YOLO26 Truly Offer Advantages Over Its Predecessors for Edge Deployment? A Benchmark Study in Aquaculture

arXiv:2607.09835v1 Announce Type: new Abstract: The recently introduced YOLO26 architecture incorporates NMS-free end-to-end inference and is optimized for deployment on resource-constrained CPU-based devices, making it well-suited for edge-based aquaculture applications. However, its performance, operational efficiency, and deployment suitability have not been systematically validated in aquaculture-specific scenarios. This study presents a comprehensive benchmark of YOLO26 against three Ultral...

14.07.2026
Blog Geração de Imagem

Adversarially Guided Diffusion for LiDAR Range Image Synthesis

arXiv:2607.09787v1 Announce Type: new Abstract: LiDAR semantic segmentation is a key perception task in autonomous driving, where false predictions can affect downstream planning and safety-critical decision-making. Although adversarial attacks, and specifically adversarial examples, have been widely studied for image classification and 3D point cloud segmentation, unrestricted adversarial examples remain largely unexplored in the space of 2D range images, which are projections of 3D point cloud...

14.07.2026
Blog LLMs & Texto

Consensus vs. Dissent: Dynamic LLM Modeling of Subjective Preferences in Group Recommenders

arXiv:2607.10235v1 Announce Type: new Abstract: Previous work in group recommender systems has demonstrated a sensitivity to the distribution of preferences within a group. Specifically, the selection of the preference aggregation strategy benefits from considering such group configurations. In this paper, we study whether LLMs are able to mimic this sensitivity and to select the ideal aggregation strategy (and corresponding recommendation) according to nuanced human perceptions of fairness, sat...

14.07.2026
Blog LLMs & Texto

RSLoRA: Training-free Rank Allocation for LoRA via Representational Sensitivity Probing

arXiv:2607.09757v1 Announce Type: new Abstract: Low-Rank Adaptation (LoRA) has become a cornerstone of parameter-efficient fine-tuning (PEFT); however, the conventional practice of uniform rank assignment ignores the functional heterogeneity of neural layers. Existing rank allocation methods typically struggle with a trade-off between computational intensity and heuristic simplicity: training-based methods suffer from prohibitive overhead, while pre-allocation methods fail to capture the dynamic...

14.07.2026
Blog Visão Computacional

Cross-Subject Modeling for Widefield Calcium Imaging via Atlas-Aligned Spatiotemporal Tokenization

arXiv:2607.09754v1 Announce Type: new Abstract: Large-scale, multi-subject widefield calcium imaging provides unprecedented access to brain-wide cortical dynamics. However, the high dimensionality, complex spatiotemporal structure, and substantial task-irrelevant activity in widefield recordings have largely restricted modeling efforts to single-session analyses, limiting scalability and generalization. While multi-subject pretrained models have been explored for some neural modalities, multi-su...

14.07.2026
350 itens no radar