Blog
LLMs & Texto
Chain-of-Models: Cross-Model Auditing for Bias-Robust LLM Judges
arXiv:2607.28636v1 Announce Type: new Abstract: LLMs increasingly serve as automated judges, but their judgments remain vulnerable to cognitive biases. Existing mitigations mostly rely on prompt-driven debiasing, which is brittle across bias types, or human evaluation, which does not scale. We study \emph{Chain-of-Models} (CoM), an automated audit pipeline in which a second model inspects the first model's reasoning trace before producing the final judgment. The key design question is whether th...
arXiv cs.CL
·Qian Wang, Zhanzhi Lou, Zhenheng Tang, Nuo Chen, Bingsheng He
·