SlimInfer: Accelerating Long-Context LLM Inference via Dynamic Token Pruning

Content not yet available

This lecture has no active video or poster.

AAAI 2026

•

January 25, 2026

•

Singapore, Singapore

Please log in to leave a comment

Next from AAAI 2026

Ego-PMOVE: Prompt-aware Mixture of View Experts Network for Egocentric Gaze Prediction
poster

Ego-PMOVE: Prompt-aware Mixture of View Experts Network for Egocentric Gaze Prediction

AAAI 2026

+4
Hongliang Li and 6 other authors

25 January 2026

Similar lecture

Stop Looking for ``Important Tokens'' in Multimodal Language Models: Duplication Matters More
poster

Stop Looking for ``Important Tokens'' in Multimodal Language Models: Duplication Matters More

EMNLP 2025

Conghui He
Weijia Li
+5
Yifeng Gao and 7 other authors

06 November 2025