If we accept that the LLM can "understand" at all, why would we reject the possibility of this understanding living in "latent space", before tokens are output?
If we accept that the LLM can "understand" at all, why would we reject the possibility of this understanding living in "latent space", before tokens are output?