The problem
Conventional decoders stop at performance.
Deep networks can match short MEG segments to the speech a person is hearing, but their learned representations do not correspond to cortical locations, neural rhythms, or time courses. A correct retrieval therefore says little about the neural evidence that produced it.