LLM Systems
Resource-aware serving for multi-model and multi-agent workloads under strict latency and memory budgets.
- Agentic workflows
- KV-cache management
- Cross-cluster serving
PhD Candidate in Software Engineering · Beihang University
I am a systems researcher working on cloud computing, GPU scheduling, and resource management for AI workloads.
My current research focuses on performance prediction, cluster scheduling, and resource-aware serving for deep learning and multi-agent LLM workloads.
01 Research
My research studies resource management and performance in cloud and AI infrastructure.
Resource-aware serving for multi-model and multi-agent workloads under strict latency and memory budgets.
Assurance-driven scheduling and resource coordination for deep learning jobs in shared clusters.
Performance prediction and resource estimation for complex cloud-native and microservice applications.
02 Publications
Conference papers on scheduling, serving, anomaly detection, and resource management.
Assurance-driven scheduling for deep learning jobs sharing multi-tenant cluster resources.
Publication materials will be linked when publicly available.
A graph-based prediction framework for estimating heterogeneous resource demand across microservice topologies.
03 Research Software
Open-source infrastructure for AI training, inference, and resource orchestration.
A cloud-native platform for AI training and inference, built for practical experimentation with shared infrastructure.
04 Experience & Education
Beihang University · Advised by Prof. Chunming Hu
ByteDance · TikTok Lumen AI Infra
Contributed to the development and operation of an AI training platform, focusing on job submission, scheduling, preemption, and production operations.
ByteDance
Beihang University
Chongqing University
05 Academic Service
Journal and conference reviewing in distributed systems, cloud computing, and intelligent systems.
TPDS
TCC
JCC 2026
COINS 2026
06 Contact
For research collaboration or academic inquiries, please contact me by email.