VeriSpec audits AI behavior rules for hidden conflicts
VeriSpec checks written AI rules and leaves final judgments to reviewers

An October 1 preprint introduces VeriSpec, a method for checking whether written AI behavior rules clash. The authors report five manually confirmed conflicts among 13 flagged pairs in OpenAIโs August 18 Model Spec.
Checking instructions in context
VeriSpec groups related rules at equal authority levels and asks a model for a concrete case where both cannot be satisfied. Examples and exceptions remain part of that check.

One example combines preserving supplied code with returning executable code. Broken input can make those obligations conflict.
Human review still matters
Two reviewers assessed the findings. Reported precision was 38.5 percent and cost per confirmed case was $11.12.
The evidence covers one document version and researcher designed baselines. It does not establish deployed chatbot behavior. ByteForward has not replicated the study.
The authors provide code and examples. The repositoryโs first commit is dated September 30.
Figure 3 from the VeriSpec paper by Zichen Xie, Mrigank Pawagi, Lize Shao, Yang Hu and Wenxi Wang. Used under CC BY 4.0. Padded and converted to WebP.



