Neuralese is a term used in Artificial Intelligence (AI) research to describe a method where Large Language Models (LLMs) perform intermediate reasoning steps in their high-dimensional latent space (vector embeddings) rather than outputting human-readable text tokens. While standard Chain of thought (AI) reasoning forces a model to generate a sequence of words, neuralese allows the model to pass raw, continuous vectors between computational layers, creating a high-bandwidth, non-linguistic reasoning channel.
See also Chain of thought (AI) Large language model Artificial intelligence safety Interpretability (machine learning) Wiktionary:neuralese
References
