Ever feel like it's too hard to keep track of what LLMs cannot do as well as humans? We're making your life easier over at: what-llms-can-not-do.github.io
We're compiling a list of papers testing the abilities of LLMs against humans. Check it out! And you can help contribute too!
what-llms-can-not-do.github.io
What LLMs Can(not) Do
A living survey of benchmarks that compare large language models with humans.