10:42
官方账号arXiv cs.AI@Cheng Xu, Nan Yan, Liming Chen, M-Tahar Kechadi
精选
推荐理由:This paper provides valuable insights into the challenges of auditing self-improvement in language models, particularly the importance of controlling for measurement artifacts. It's a must-read for anyone interested in the field of AI auditing and model evaluation.