Source: Pandaily
Xiaomi's LLM-Core team published HySparse2, a hybrid sparse-attention design with two-level KV sharing aimed at MiMo-V3-class agentic models, reporting about 5× lower 1M-token prefill FLOPs versus Hybrid SWA on an 80B-A3B MoE.
Read the full article at the source →
İlk yorumu siz yazın!
Yorum
Ad
E-posta
Save my name, email, and website in this browser for the next time I comment.
Gönder
Comments
İlk yorumu siz yazın!