I am Qiang Wang (王强), an Associate Professor at Department of Computer Science and Technology, Harbin Institute of Technology, Shenzhen. Before that, he was a Research Assistant Professor at Department of Computer Science, Hong Kong Baptist University and a Senior Engineer at Tencent (Shenzhen). He received his B.E. degree from South China University of Technology in 2014, and his Ph.D. degree at Department of Computer Science, Hong Kong Baptist University, under supervision of Prof. Xiaowen Chu, in 2020. His research interests include GPU Computing, Energy Efficiency, Distributed Computing, and High Performance Machine Learning.
News
- [Aug 2026] The paper “Laplacian Frequency Hierarchies for Efficient 3D Gaussian Splatting Training” has been accepted by Pacific Graphics (PG) 2026. Congratulations to Dr. Yang Yixiong and Sisheng Zhang.
- [July 2026] The paper “SpatialGrammar: A Domain-Specific Language for LLM-Based 3D Indoor Scene Generation” has been accepted by ACMMM 2026.
- [July 2026] The paper “Same Cache, Different Latency: Understanding and Exploiting L2 Domain Locality in GPUs” has been accepted by MICRO 2026. Congratulations to Weile Luo and Yibin Ma.
- [July 2026] The paper “Exploiting Low-Level Sparsity for Efficient Large Language Model Inference on GPU with SpInfer” has been accepted by TOCS 2026.
- [July 2026] The paper “WAQ-LLM: Optimizing Multi-Instance LLM Deployment via Workload-Aware Queueing Model” has been accepted by ICPP 2026. Congratulations to Jiaxin Lai and Yizhou Luo.
- [April 2026] Our paper “DOSA: Optimizing DVFS Policy for Stochastic Workloads with Analytical Solutions” has been accepted by IWQoS 2026. Congratulations to Xuanhao Feng.
- [Dec 2025] The paper “FastTT: Accelerating Shift-XOR Erasure Coding for Data Storage” has been accepted by IPDPS 2026. Congratulations to Duo Sun.
- [Nov 2025] The paper “ZipServ: Fast and Memory-Efficient LLM Inference with Hardware-Aware Lossless Compression” has been accepted by ASPLOS 2026.
- [Nov 2025] The paper “ROME: Maximizing GPU Efficiency for All-Pairs Shortest Path via Taming Fine-Grained Irregularities” has been accepted by PPoPP 2026.
- [Nov 2025] The papers “SALR: Sparsity-Aware Low-Rank Representation for Efficient Fine-Tuning of Large Language Models” and “PipeDiT: Accelerating Diffusion Transformers in Video Generation with Task Pipelining and Model Decoupling” have been accepted by AAAI 2026.
- [Oct 2025] The paper “Castor: Optimizing Deep Learning Job Scheduling in Multi-Tenant GPU Clusters via Intelligent Colocation” has been accepted by TCC. Congratulations to Yizhou Luo.
- [Oct 2025] The paper “AUE: A Normalized Energy Efficiency Metric for AI Servers under LLM Workloads” has been accepted by ICPADS 2025. This paper proposes a new metric to evaluate the energy efficiency of AI servers, and has received the Best Presentation Award.
- [July 2025] The paper “City-VLM: Towards Multidomain Perception Scene Understanding via Multimodal Incomplete Learning” has been accepted by ACMMM 2025.
- [June 2025] The papers “SAFormer: Spatially Adaptive Transformer for Efficient and Multi-Resolution Occupancy Prediction” and “RA-NeRF: Robust Neural Radiance Field Reconstruction with Accurate Camera Pose Estimation under Complex Trajectories” have been accepted by IROS 2025.
- [May 2025] The paper “UnrealLLM: Towards Highly Controllable and Interactable 3D Scene Generation by LLM-powered Procedural Content Generation” has been accepted by ACL 2025 (Findings).
- [May 2025] The paper “BurstGPT: A Real-World Workload Dataset to Optimize LLM Serving Systems” has been accepted by KDD 2025 (Datasets and Benchmarks Track).
- [May 2025] The paper “SCFusion: Enhance Infrared and Visible Modality Fusion by Preserving Salient Object Consistency” has been accepted by IoTJ.
- [April 2025] The paper “SpInfer: Leveraging Low-Level Sparsity for Efficient Large Language Model Inference on GPUs” has received the Best Paper Award of EuroSys 2025.
- [March 2025] The paper “CMRFusion: Efficient Feature Decomposition for RGB-T Fusion via Cross Modality Mask Reconstruction” has been accepted by ICME 2025.
- [Jan 2025] The paper “SpInfer: Leveraging Low-Level Sparsity for Efficient Large Language Model Inference on GPUs” has been accepted by EuroSys 2025.
- [Jan 2025] The paper “STBLLM: Breaking the 1-Bit Barrier with Structured Binary LLMs” has been accepted by ICLR 2025.
Biography
- 2015-2020, Ph.D., Hong Kong Baptist University, supervised by Prof. Xiaowen Chu
- 2010-2014, B.E., Computer Science and Technology, South China University of Technology, China
Work Experience
- 2025.01-present, Associate Professor, Harbin Institute of Technology (Shenzhen)
- 2022.05-2024.12, Assistant Professor, Harbin Institute of Technology (Shenzhen)
- 2022.01-2022.05, Senior Engineer, Tencent (Shenzhen)
- 2020.09-2022.01, Research Assistant Professor, Hong Kong Baptist University
- 2020.06-2020.08, Research Intern, Alibaba Ant Technology Group (Beijing)
- 2019.08-2020.05, Research Intern, Nvidia Corporation
- 2019-2020/2014-2015, Research Assistant, Hong Kong Baptist University, supervised by Prof. Xiaowen Chu
Awards
- 2025, HW Talent Funding Awardee
- 2025, EuroSys 2025, Best Paper Award
- 2020, IEEE GreenCom 2020, Best Paper Award
- 2020, RPg Performance Award Scheme, Hong Kong Baptist University [link]
- 2018, IEEE DataCom 2018, Best Paper Award
- 2016-2019, Excellent Teaching Assistant Performance Award [link]
- 2015, Hong Kong PhD Fellowship Awardee [link] [link]
- 2013, American Mathematical Contest in Modeling, Honorable Prize
- 2011, National Scholarship
Professional Activities
- Conference Organizer
- IEEE MetaCom 2025, Track Chair of “Metaverse Computing, Architectures, and Applications”
- IWQoS 2024, Web Chair
- Invited Program Committee Member/Reviewer for Conferences (the number counts the contributed reviews.)
- 2026: CVPR, ICLR(1), AAAI(4), ICDCS, IROS
- 2025: ICDCS(8), NeurIPS D&B(3), ICCV(2), ACMMM(3), CVPR(3), WACV(2), AAAI(4), ICME(3), ACL(2)
- 2024: ECCV(3), CVPR(2), AAAI(2), IWQoS(5)
- 2023: AAAI(3), WACV(1), HiPC, BigCom(5)
- 2022: GreenCom, HPCC, HiPC(3)
- 2021: HPCC(6), HiPC(4), ICPADS(2), ICRA(1)
- Invited Reviewer for Journals
- IEEE Transactions on Computers (TC)
- IEEE Transactions on Parallel and Distributed Systems (TPDS)
- IEEE Transactions on Network Science and Engineering (TNSE)
- IEEE Transactions on Cloud Computing (TCC)
- IEEE Transactions on Sustainable Computing (TSUSC)
- IEEE Transactions on Big Data (TBD)
- IEEE/ACM Transactions on Networking (ToN)
- ACM Transactions on Design Automation of Electronic Systems (TODAES)
- IEEE Transactions on Circuits and Systems for Video Technology (TCSVT)
- IEEE Transactions on Dependable and Secure Computing (TDSC)
- IEEE Robotics and Automation Letters (RA-L)
- IEEE Network
- IEEE Pervasive Computing (PC)
- IEEE Access
- ACM Computing Surveys (CSUR)
- Computer Vision and Image Understanding (CVIU)
- Future Generation Computer Systems (FGCS)
- Expert Systems With Applications (ESWA)
- Journal of System Architecture (JSA)