Comment on UCSD: Large Language Models Pass the Turing TestparentComments−rfoo1yI still believe that larger models are better at covering the long tail. Our benchmarks are saturated, but actual model capability is not.
Comments
I still believe that larger models are better at covering the long tail. Our benchmarks are saturated, but actual model capability is not.