Alert button
Picture for Mengdi Wu

Mengdi Wu

Alert button

FlexLLM: A System for Co-Serving Large Language Model Inference and Parameter-Efficient Finetuning

Add code
Bookmark button
Alert button
Feb 29, 2024
Xupeng Miao, Gabriele Oliaro, Xinhao Cheng, Mengdi Wu, Colin Unger, Zhihao Jia

Viaarxiv icon

Finding the Task-Optimal Low-Bit Sub-Distribution in Deep Neural Networks

Add code
Bookmark button
Alert button
Jan 13, 2022
Runpei Dong, Zhanhong Tan, Mengdi Wu, Linfeng Zhang, Kaisheng Ma

Figure 1 for Finding the Task-Optimal Low-Bit Sub-Distribution in Deep Neural Networks
Figure 2 for Finding the Task-Optimal Low-Bit Sub-Distribution in Deep Neural Networks
Figure 3 for Finding the Task-Optimal Low-Bit Sub-Distribution in Deep Neural Networks
Figure 4 for Finding the Task-Optimal Low-Bit Sub-Distribution in Deep Neural Networks
Viaarxiv icon