Picture for Yunan Zhang

Yunan Zhang

FSE: Continual Learning for Named Entity Recognition by Fast-Slow Experts

Add code
Jul 24, 2026
Viaarxiv icon

GeRe: Towards Efficient Anti-Forgetting in Continual Learning of LLM via General Samples Replay

Add code
Aug 06, 2025
Viaarxiv icon

Phi-4-Mini Technical Report: Compact yet Powerful Multimodal Language Models via Mixture-of-LoRAs

Add code
Mar 03, 2025
Figure 1 for Phi-4-Mini Technical Report: Compact yet Powerful Multimodal Language Models via Mixture-of-LoRAs
Figure 2 for Phi-4-Mini Technical Report: Compact yet Powerful Multimodal Language Models via Mixture-of-LoRAs
Figure 3 for Phi-4-Mini Technical Report: Compact yet Powerful Multimodal Language Models via Mixture-of-LoRAs
Figure 4 for Phi-4-Mini Technical Report: Compact yet Powerful Multimodal Language Models via Mixture-of-LoRAs
Viaarxiv icon

A Little Goes a Long Way: Efficient Long Context Training and Inference with Partial Contexts

Add code
Oct 02, 2024
Figure 1 for A Little Goes a Long Way: Efficient Long Context Training and Inference with Partial Contexts
Figure 2 for A Little Goes a Long Way: Efficient Long Context Training and Inference with Partial Contexts
Figure 3 for A Little Goes a Long Way: Efficient Long Context Training and Inference with Partial Contexts
Figure 4 for A Little Goes a Long Way: Efficient Long Context Training and Inference with Partial Contexts
Viaarxiv icon

Efficient LLM Training and Serving with Heterogeneous Context Sharding among Attention Heads

Add code
Jul 25, 2024
Figure 1 for Efficient LLM Training and Serving with Heterogeneous Context Sharding among Attention Heads
Figure 2 for Efficient LLM Training and Serving with Heterogeneous Context Sharding among Attention Heads
Figure 3 for Efficient LLM Training and Serving with Heterogeneous Context Sharding among Attention Heads
Figure 4 for Efficient LLM Training and Serving with Heterogeneous Context Sharding among Attention Heads
Viaarxiv icon

Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone

Add code
Apr 23, 2024
Figure 1 for Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone
Figure 2 for Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone
Figure 3 for Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone
Figure 4 for Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone
Viaarxiv icon

GenSERP: Large Language Models for Whole Page Presentation

Add code
Feb 22, 2024
Figure 1 for GenSERP: Large Language Models for Whole Page Presentation
Figure 2 for GenSERP: Large Language Models for Whole Page Presentation
Figure 3 for GenSERP: Large Language Models for Whole Page Presentation
Figure 4 for GenSERP: Large Language Models for Whole Page Presentation
Viaarxiv icon

Model Tells You What to Discard: Adaptive KV Cache Compression for LLMs

Add code
Oct 07, 2023
Figure 1 for Model Tells You What to Discard: Adaptive KV Cache Compression for LLMs
Figure 2 for Model Tells You What to Discard: Adaptive KV Cache Compression for LLMs
Figure 3 for Model Tells You What to Discard: Adaptive KV Cache Compression for LLMs
Figure 4 for Model Tells You What to Discard: Adaptive KV Cache Compression for LLMs
Viaarxiv icon

A Neural Span-Based Continual Named Entity Recognition Model

Add code
Feb 23, 2023
Figure 1 for A Neural Span-Based Continual Named Entity Recognition Model
Figure 2 for A Neural Span-Based Continual Named Entity Recognition Model
Figure 3 for A Neural Span-Based Continual Named Entity Recognition Model
Figure 4 for A Neural Span-Based Continual Named Entity Recognition Model
Viaarxiv icon

Towards Disentangling Relevance and Bias in Unbiased Learning to Rank

Add code
Dec 28, 2022
Figure 1 for Towards Disentangling Relevance and Bias in Unbiased Learning to Rank
Figure 2 for Towards Disentangling Relevance and Bias in Unbiased Learning to Rank
Figure 3 for Towards Disentangling Relevance and Bias in Unbiased Learning to Rank
Figure 4 for Towards Disentangling Relevance and Bias in Unbiased Learning to Rank
Viaarxiv icon