interpretabilityfactneutralIn sparse mixture-of-experts language models, expert subspaces overlap substantially despite the expectation that co-selected experts should contribute distinct representation directionsMachine Learning02 Aug 2026http://arxiv.org/abs/2607.28308v1