Research Interests
Large Language Models
Instruction tuning, alignment, long-context reasoning, and domain-specific LLMs.
Efficient AI
KV cache compression, speculative decoding, prompt compression, and quantization.
Multimodal Models
Vision-language, document intelligence, audio-language, and music AI.
NLP & Structure Parsing
Information extraction, syntactic/semantic parsing, and structured reasoning.
Selected Publications Full list →
(# Equal Contribution; * Corresponding Author)
2025 – 2026
⚡
🔬
📖
Segment First or Comprehend First? Explore the Limit of Unsupervised Word Segmentation with Large Language Models
ACL 2025
Oral
🚀
Scaling LLM Speculative Decoding: Non-Autoregressive Forecasting in Large-Batch Scenarios
AAAI 2026
2024
✂️
News
- Apr 2026 10 papers accepted to ACL 2026: 5 main conference papers and 5 findings papers.10 篇论文入选 ACL 2026:主会长文 5 篇、Findings 长文 5 篇。
- Jan 2026 3 papers accepted to AAAI 2026: Scaling LLM Speculative Decoding, End-to-end Contrastive Language-Speech Pretraining, and Ghost in the Transformer.3 篇论文入选 AAAI 2026:大模型投机解码扩展、端到端对比语言—语音预训练、Ghost in the Transformer。
- Dec 2025 Paper accepted to NeurIPS 2025: SmallKV — small model assisted KV cache compression.1 篇论文入选 NeurIPS 2025:SmallKV —— 小模型辅助的 KV 缓存压缩。
- Nov 2025 6 papers accepted to EMNLP 2025: ToM, XQuant, Faster In-Context Learning, CoViPAL, and more.6 篇论文入选 EMNLP 2025:ToM、XQuant、Faster In-Context Learning、CoViPAL 等。
- Jul 2025 6 papers accepted to ACL 2025 (1 Oral) and 2 papers to NAACL 2025.6 篇论文入选 ACL 2025(1 篇 Oral)、2 篇入选 NAACL 2025。
- May 2025 Paper accepted to ICML 2025: Uni-Bi-Directional Mixture-of-Expert method.1 篇论文入选 ICML 2025:双向统一的混合专家方法。
- Feb 2025 2 papers accepted to AAAI 2025: Imitate Before Detect and SongSong.2 篇论文入选 AAAI 2025:Imitate Before Detect 与 SongSong。
- Jul 2024 3 papers accepted to ACL 2024: SirLLM, Hypergraph Document Understanding, and Selective Prefix Tuning.3 篇论文入选 ACL 2024:SirLLM、超图文档理解、Selective Prefix Tuning。
- May 2024 Paper accepted to ICML 2024: SIFT — Sparse is Enough in Fine-tuning Pre-trained LLMs.1 篇论文入选 ICML 2024:SIFT —— 稀疏足以微调预训练大模型。
- Jan 2024 2 papers accepted to AAAI 2024: PromptKD and Prompt Compression.2 篇论文入选 AAAI 2024:PromptKD 与提示压缩。
- Dec 2023 5 papers accepted to EMNLP 2023.5 篇论文入选 EMNLP 2023。
- May 2023 3 papers accepted to ACL 2023.3 篇论文入选 ACL 2023。