AntiLeak-Bench: Preventing Data Contamination by Automatically Constructing Benchmarks with Updated Real-World Knowledge
原文英文,约100词,阅读约需1分钟。
📝
内容提要
本研究提出了AntiLeak-Bench框架,旨在通过自动构建新知识样本防止数据污染,确保大型语言模型(LLM)评估的无污染性。该框架实现了完全自动化的工作流程,显著降低了基准维护成本,有效应对数据污染问题。
🏷️