arxiv:2601.14004
zhanghengyuan
hengyuanya
·
AI & ML interests
None yet
Recent Activity
upvoted a paper about 2 months ago
Massive Activations in Hybrid Linear Attention Large Language Models: Pre-Attention Spikes and Inter-Spike Plateaus upvoted a paper 4 months ago
OScaR: The Occam's Razor for Extreme KV Cache Quantization in LLMs and Beyond upvoted a paper 6 months ago
Attention Sink in Transformers: A Survey on Utilization, Interpretation, and MitigationOrganizations
None yet