‘Future Lens: Anticipating Subsequent Tokens From a Single Hidden State’

“We conjecture that hidden state vectors corresponding to individual input tokens encode information sufficient to accurately predict several tokens ahead. More concretely, in this paper we ask: Given a hidden (internal) representation of a single token at position t in an input, can we reliably anticipate the tokens that will appear at positions ≥t+2? … We find that, at some layers, we can approximate a model’s output with more than 48% accuracy with respect to its prediction of subsequent tokens through a single hidden state.”

Find the paper and full list of authors at ArXiv.

View on Site: ‘Future Lens: Anticipating Subsequent Tokens From a Single Hidden State’