Hanyu (Olivia) Yu, a high school sophomore, is seeking an arXiv endorser in cs.LG or cs.CV to upload her independent preprint titled "AttnRoute-MoE: Attention-Prior Routing for Mixture-of-Experts Vision Transformers."

The paper introduces three scalar signals from the self-attention matrix—entropy, cross-head variance, and CLS-similarity—as a routing prior that anneals to zero during training. The author reports that routing diversity at initialization prevents early collapse, while semantic attention shapes post-training expert specialization. The work includes 3-seed experiments, four ablation conditions, and honest reporting of a non-significant p > 0.05 result.

arXiv requires endorsers to have submitted 3+ papers to cs.* categories within the last 5 years, with at least one older than 3 months.