ACL 2026

•

July 07, 2026

•

San Diego, United States

Please log in to leave a comment

Downloads

PaperTranscript English (automatic)

Next from ACL 2026

MixKVQ: Query-Aware Mixed-Precision KV Cache Quantization for Long-Context Reasoning
technical paper

MixKVQ: Query-Aware Mixed-Precision KV Cache Quantization for Long-Context Reasoning

ACL 2026

Ziqian Zeng
+2
Cen Chen and 4 other authors

07 July 2026

Similar lecture

Towards Efficient and Effective Diffusion Language Model Inference via Semantic-Aware Adaptive Denoising
technical paper

Towards Efficient and Effective Diffusion Language Model Inference via Semantic-Aware Adaptive Denoising

ACL 2026

Yu Gu
Fangling Leng
Fan Li
+3
Yu Gu and 5 other authors

06 July 2026