Explore
Discover
Solutions
AI Services
Our Products
Submit a tool
Search the catalog
⌘K
$
/
₹
Sign up free
Login
Back to papers
October 8, 2026
cs.LG
cs.CL
SparseDecoding: Decoding-Aware Pruning for Accurate and Efficient LLM Inference
Qitong Wang
,
Xinwei Niu
,
Mingluo Su
,
Shanwei Zhao
,
Shiai Zhu
,
Huan Wang
Original Abstract
Read on arXiv
Download PDF
HF Upvotes
21
Categories
cs.LG, cs.CL
SparseDecoding: Decoding-Aware Pruning for Accurate and Efficient LLM Inference - AI Research | One9Founders