Yaqi Xia (夏亚奇)
HongYi Postdoc Research Fellow, School of Computer Science, Wuhan University
I am Yaqi Xia, a HongYi Postdoc Research Fellow at Wuhan University, working with Prof. Dazhao Cheng. My research focuses on distributed and high-performance systems for AI/ML, with an emphasis on efficient long-context LLM inference, scalable graph learning systems, and ML systems optimization. My work has appeared in top systems venues and journals including SC, PPoPP, ATC, IEEE TPDS, and HPDC, with Best Paper Runner-up and Best Paper Award Nomination honors. I received my Ph.D. from Wuhan University in 2025. Prior to that, I obtained both my Bachelor's and Master's degrees from Xidian University under the supervision of Prof. Rui Song.
Wuhan University
Ph.D. in Artificial Intelligence Sep. 2021 - Dec. 2025
Xidian University
M.S. in Electronics and Communication Engineering Sep. 2018 - Jul. 2021
Xidian University
B.S. in Communication Engineering Sep. 2014 - Jul. 2018
Research Center for Graph Computing, Zhejiang Lab
Research Intern Aug. 2023 - Dec. 2023
Chengyu Sun, Yaqi Xia†, Ruirui Pan, Donglin Yang, Xiaobo Zhou, Dazhao Cheng†(† corresponding author)
2026 International Conference for High Performance Computing, Networking, Storage, and Analysis (SC) 2026 ConferenceCCF-ABest Paper Award Nomination
We present DySpin, a plug-and-play library that advances dynamic sparse long-context inference.
Weihu Wang*, Yaqi Xia*, Donglin Yang, Xiaobo Zhou, Dazhao Cheng†(* equal contribution; † corresponding author)
2025 International Conference for High Performance Computing, Networking, Storage, and Analysis (SC) 2025 ConferenceCCF-A
We present MXBLAS, a high-performance MX-GEMM library that unifies support across the full spectrum of MX-format variations.
Yaqi Xia*, Weihu Wang*, Donglin Yang, Xiaobo Zhou†, Dazhao Cheng†(* equal contribution; † corresponding author)
2025 USENIX Annual Technical Conference (ATC) 2025 ConferenceCCF-A
We introduce Voltrix-SpMM, a revolutionary GPU kernel design for sparse matrix-matrix multiplication.
Yaqi Xia, Zheng Zhang, Donglin Yang, Chuang Hu, Xiaobo Zhou, Hongyang Chen, Qianlong Sang, Dazhao Cheng†(† corresponding author)
IEEE Transactions on Parallel and Distributed (TPDS) 2024 JournalCCF-A
This work introduces Sven, a co-designed algorithm-system library aimed at accelerating TGNN training on a multi-GPU platform.
Yaqi Xia, Donglin Yang, Xiaobo Zhou, Dazhao Cheng†(† corresponding author)
The International Conference for High Performance Computing, Networking, Storage, and Analysis (SC) 2024 ConferenceCCF-A
In this paper, we introduced HyDRA, a pioneering framework for sampling-based GNN training on large-scale graphs.
Zheng Zhang, Yaqi Xia, Hulin Wang, Donglin Yang, Chuang Hu, Xiaobo Zhou, Dazhao Cheng†(† corresponding author)
IEEE Transactions on Parallel and Distributed (TPDS) 2024 JournalCCF-ABest Paper Runner-up
In this paper, we present the design and implementation of MPMoE, a high-performance library that accelerates MoE training with adaptive and memory-efficient pipeline parallelism.
Yaqi Xia, Zheng Zhang, Hulin Wang, Donglin Yang, Xiaobo Zhou, Dazhao Cheng†(† corresponding author)
The 32nd International Symposium on High-Performance Parallel and Distributed Computing (ACM HPDC) 2023 ConferenceCCF-ABest Paper Nomination
This paper presents Sven, an algorithm and system co-designed TGNN training library for the end-to-end performance optimization on multi-node multi-GPU systems.
Yaqi Xia*, Yan Xia*, Wei Li, Rui Song, Kailang Cao, Uwe Stilla†(* equal contribution; † corresponding author)
Proceedings of the 29th ACM international conference on multimedia (ACM MM) 2021 ConferenceCCF-A
We tackle the problem of object completion from point clouds and propose a novel point cloud completion network employing an Asymmetrical Siamese Feature Matching strategy, termed as ASFM-Net.