1
Your Transformer Can Hold Two Thoughts at Once: Evidence of Linear Superposition in LLMs
The paper demonstrates that transformer LLMs exhibit a linear superposition property where combined inputs produce a blended next‑token distribution, and that lightweight fine‑tuning can restore this linearity. It also introduces a guided decoding algorithm that extracts two distinct, coherent continuations from one forward pass.
Hugging Face Daily Papersarxiv.org1 minpaper
