Hao Chen (陈浩)
I obtained my Ph.D. degree in Computer Science from University of Science and Technology of China (USTC) in 2021, guided by Professors Yinlong Xu and Cheng Li. During my Ph.D. studies, I interned at Qatar Computing Research Institute (QCRI) as a research assistant from 2019 to 2021, mentored by Xiaosong Ma.
Currently, I’m part of the PolarDB MySQL Department at Alibaba Cloud, focusing on the development and research of next-generation cloud-native databases. My research interests encompass cloud-native databases, RDMA (Remote Direct Memory Access), and CXL, with a particular focus on their evolving advancements and practical applications. I am also expanding my research interests toward AI infrastructure, with a focus on scalable model serving, KV cache management, and efficient resource utilization. I am keenly interested in exploring innovative partnerships and research opportunities in this field.
Research Intern Openings: We are looking for research interns interested in AI-native data systems and AI infrastructure. Potential topics include vector databases, agent memory systems, KV cache management, RDMA/CXL-based disaggregated memory systems, and efficient LLM inference serving. If chasing down bottlenecks in large-scale cloud and AI infrastructure sounds like fun, we would be happy to hear from you.
Feel free to reach out via email: cighao@gmail.com.
Awards
- SIGMOD 2025 Industry Track Best Paper Award (Corresponding Author)
- SIGMOD 2024 Industry Track Best Paper Award (Corresponding Author)
- ICDE 2024 Best Industry and Application Paper Award (Corresponding Author)
Experience
Research and development of next-generation cloud-native databases and AI infrastructure.
- KVCache for LLM Serving [PVLDB'26]
- Cloud-Native Database Architecture: PolarDB Multi-Primary [SIGMOD'24 Best Paper], PolarDB Serverless [ICDE'24 Best Paper]
- Distributed Database Systems: PolarDB Strongly Consistent Reads [PVLDB'23], PolarDB Limitless [PVLDB'25], Geo-Replication [PVLDB'24]
- Memory-Centric Database Systems: Persistent Memory [ASPLOS'23], CXL Memory Pool [SIGMOD'25 Best Paper], CXL-based Locking [TC'26], AP Memory Pool [SIGMOD'26]
Ph.D. Research Intern.
- SpanDB: Hybrid-Storage KV Store [FAST'21]
- QarSUMO: Parallel Traffic Simulator [SIGSPATIAL'20]
Publications
PolarKV: Tier Locally, Serve Globally—A KV Cache over Cloud Memory and Storage
LakeMem: An Elastic Disaggregated-Memory Caching Layer for Analytical Processing Systems
CXLock: Efficient and Scalable Lock Management for CXL-enabled Distributed Systems
From Scale-Up to Scale-Out: PolarDB's Journey to Achieving 2 Billion tpmC
Unlocking the Potential of CXL for Disaggregated Memory in Cloud-Native Databases
PolyBase: Adapting to Data Affinity Changes in Geo-Replicated Database via Row-Level Paxos-Group Affiliation Re-Assignment
PolarDB-MP: A Multi-Primary Cloud-Native Database via Disaggregated Shared Memory
Towards a Shared-Storage-Based Serverless Database Achieving Seamless Scale-Up and Read Scale-Out
PolarDB-SCC: A Cloud-Native Database Ensuring Low Latency for Strongly Consistent Reads
Persistent Memory Disaggregation for Cloud-Native Relational Databases
HCFTL: A Locality-Aware Flash Translation Layer for Efficient Address Translation
Leveraging NVMe SSDs for Building a Fast, Cost-effective, LSM-tree-based KV Store
SpanDB: A Fast, Cost-Effective LSM-Tree Based KV Store on Hybrid Storage
GFTL: Group-Level Mapping in Flash Translation Layer to Provide Efficient Address Translation for NAND Flash-Based SSDs
QarSUMO: A Parallel, Congestion-optimized Traffic Simulator
ECR: Eviction-cost-aware cache management policy for page-level flash-based SSDs
HCFTL: A Locality-Aware Page-Level Flash Translation Layer
LCR: Load-Aware Cache Replacement Algorithm for Flash-Based SSDs
Boosting performance of SSD with chip-level RAID by deferring garbage collection
Conferences & Talks
Places, people, and ideas along my research journey.
- [Sponsor Talk] Where Data Gravity Meets Intelligence: Building AI-Native Databases for Agentic EraVLDB 2026Boston
- PolarKV: Tier Locally, Serve Globally—A KV Cache over Cloud Memory and StorageVLDB 2026Boston
- Towards a Shared-Storage-Based Serverless Database Achieving Seamless Scale-Up and Read Scale-OutICDE 2024Utrecht
- HCFTL: A Locality-Aware Page-Level Flash Translation LayerDATE 2019Florence
No talks yet.