2026 ICML 2026 Poster SimpleGPT: Improving GPT via A Simple Normalization Strategy Marco Chen*, Xianbiao Qi*, Yelin He, and 2 more authors In International Conference on Machine Learning, 2026 arXiv HTML Code ICML 2026 Spotlight Delving into Muon and Beyond: Deep Analysis and Extensions Xianbiao Qi*, Marco Chen*, Jiaquan Ye, and 2 more authors In International Conference on Machine Learning, 2026 arXiv HTML Code arXiv 2026 Preprint SageBwd: A Trainable Low-Bit Attention Jintao Zhang*, Marco Chen*, Haoxu Wang*, and 5 more authors arXiv preprint arXiv:2603.02170, 2026 arXiv HTML ICLR 2026 Poster DNT: A Deeply Normalized Transformer That Can Be Trained by Momentum SGD Xianbiao Qi, Marco Chen, Wenjie Xiao, and 4 more authors In International Conference on Learning Representations, 2026 arXiv HTML arXiv 2026 Preprint Vidu S1: A Real-Time Interactive Video Generation Model Jintao Zhang, Kai Jiang, Jintao Chen, and 24 more authors arXiv preprint arXiv:2607.03118, 2026 arXiv HTML 2025 NeurIPS 2025 Poster SeƱorita-2M: A High-Quality Instruction-Based Dataset for General Video Editing by Video Specialists Bojia Zi*, Penghui Ruan*, Marco Chen, and 7 more authors In Advances in Neural Information Processing Systems, 2025 arXiv HTML