Harmonizing AI Safety Thresholds

arXiv:2607.16112v1 Announce Type: new Abstract: Frontier AI companies have published capability thresholds that differ substantially, making it difficult for third parties to verify whether a threshold has been crossed or to compare requirements across companies. Moreover, without common minimum thresholds, risk mitigation may be inconsistent, creating a potential race to the bottom in safety standards. We develop a methodology for deriving harmonized thresholds across three risk domains. For mi...

arXiv cs.AI ·Wilber Sean Anterola, Matthew Ball, Luis F. Lafuerza, Markov Grey ·
compartilhar: