问HN:还有哪些“无法破解的基准”?
我根据自己的博士论文设定了一个关于人工智能的基准,但由于大学的限制,这个基准并不公开。我在想是否还有其他人也有一些“隐藏”的基准?我自己有几个基准,想知道谈论它们是否会贬值,以及是否可以制作一个不会被泄露的良好测试集?
查看原文
I have my own benchmark for AI just based off my Doctoral Thesis, which isn't public(university constraints). I am wondering if other people have other "hidden" benchmarks? I have a few I have made, and one wonder if talking about them devalues them, and two can a good test set be made that can't be leaked?