What lies outside the "regular" embeddings space of an LLM?
By definition an llm is just a manifold in a space with (whatever dimension of a single token)* times (context length) dimensions. human text is naturally going to cluster over certain regions and since neural networks are defined over the entire space…