Hugging Face Blog·· 2024-04-16AI 评分40
Hugging Face 推出 LiveCodeBench 排行榜:面向代码 LLM 的无污染整体评估
Introducing the LiveCodeBench Leaderboard - Holistic and Contamination-Free Evaluation of Code LLMs
AI 导读
Hugging Face 上线 LiveCodeBench 排行榜,该基准由 UC Berkeley、MIT 和 Cornell 研究者开发,从 LeetCode、AtCoder、CodeForces 持续收集带发布日期的编程题,通过“随时间滚动”评估窗口检测并防止数据污染。
来源:Hugging Face Blog · huggingface.co