MLST — AI benchmarks are broken! [Prof Melanie Mitchell]
I really love this part of the MLST interview in which Prof Mitchell says the key LLM question is: what kind of “understanding,” if any, is really going on?
They don’t and can’t truly “understand” — it’s just word statistics.
They do form rich, concept-like mental models.
Or their huge correlations amount to a new, non-human kind of understanding.