Dr. Xiaoyang Lu is an Assistant Research Professor in the Department of Computer Science at Illinois Institute of Technology and a member of the Gnosis Research Center. His research investigates computer architecture and high-performance systems, specifically optimizing memory hierarchies, processing-in-memory architectures, and cache management for concurrent workloads. He leads the memory and architecture research within the center.
Joined: Aug 2016Graduated: May 2024
Now: Assistant Research Professor at Illinois Institute of Technology
18Publications
3Projects
Research Interests
- Computer Architecture
- Memory Systems
- Processing-in-Memory
- Cache Optimization
Projects
News Mentions
Awards & Honors
News & Announcements
Nov 2025ProMiner Appears in IEEE TCAD
Publications
2026
Improving Data Reuse across Blocks for Efficient Block-Sparse Transformers on GPUs
2026
I/O-Aware PIM Acceleration for Long-Sequence LLM Inference with Hybrid Sparse Attention
2026
HyPIM: Accelerating Hyperbolic Machine Learning via Processing-In-Memory
2026
Zion: A Comprehensive, Adaptive, and Lightweight Hardware Prefetcher
2026
I/O Analysis is All You Need: An I/O Analysis for Long-Sequence Attention
2025
COSMOS: RL-Enhanced Locality-Aware Counter Cache Optimization for Secure Memory
2025
Concurrency-Aware Cache Miss Cost Prediction with Perceptron Learning
2025
ProMiner: Enhancing Locality, Parallelism, and Offloading for Graph Mining on Processing-in-Memory Systems
2025
Pyramid: Accelerating LLM Inference with Cross-Level Processing-in-Memory
2024
AceMiner: Accelerating Graph Pattern Matching using PIM with Optimized Cache System
2024
ACES: Accelerating Sparse Matrix Multiplication with Adaptive Execution Flow and Concurrency-Aware Cache Optimizations
2024
CHROME: Concurrency-Aware Holistic Cache Management Framework with Online Reinforcement Learning
2023
CARE: A Concurrency-Aware Enhanced Lightweight Cache Management Framework
2023
The Memory-Bounded Speedup Model and Its Impacts in Computing
2022
A Generalized Model For Modern Hierarchical Memory System
2021
Premier: A Concurrency-Aware Pseudo-Partitioning Framework for Shared Last-Level Cache
2021
CoPIM: A Concurrency-aware PIM Workload Offloading Architecture for Graph Applications
2020
