三篇论文被ICML 2026接收
三篇关于高效推理、稀疏注意力和扩散大语言模型的论文被ICML'2026接收。
被接收的论文如下:
- Attention Sink Forges Native MoE in Attention Layers: Sink-Aware Training to Address Head Collapse(由Zizhuo Fu主导)
- TEAM: Temporal–Spatial Consistency Guided Expert Activation for MoE Diffusion Language Model Acceleration(由Linye Wei主导)
- HyPER: Bridging Exploration and Exploitation for Scalable LLM Reasoning with Hypothesis Path Expansion and Reduction(由Shengxuan Qiu、Haochen Huang主导)