HKUST Computer Architecture Group
HKUST Computer Architecture Group
News
People
Events
Publications
Contact
P. Luo
Latest
A 14.08-to-135.69Token/s ReRAM-on-Logic Stacked Outlier-Free Large-Language-Model Accelerator with Block-Clustered Weight-Compression and Adaptive Parallel-Speculative-Decoding
A 28nm 0.22uJ/Token Memory-Compute-Intensity-Aware CNN-Transformer Accelerator with Hybrid-Attention-Based Layer-Fusion and Cascaded Pruning for Semantic-Segmentation
CoXplorer: Multi-staged Co-exploration Framework for AI Model Compression and Accelerator Design
Cite
×