4 stories tagged with #kv-sharing, in publish-time order across the WeSearch catalog. Tag pages update as new stories ingest.
⌘ RSS feed for this tag → or search "Kv Sharing"
R/MACHINELEARNING
Recent Developments in LLM Architectures: KV Sharing, mHC, and Compressed Attention [P]
R/LOCALLLAMA
Recent Developments in LLM Architectures: KV Sharing, mHC, and Compressed Attention
HACKER NEWS (AI / LLM)
Recent Developments in LLM Architectures: KV Sharing, MHC, Compressed Attention
From Gemma 4 to DeepSeek V4, How New Open-Weight LLMs Are Reducing Long-Context Costs…
SUBSTACK
Recent Developments in LLM Architectures: KV Sharing, MHC, Compressed Attention
From Gemma 4 to DeepSeek V4, How New Open-Weight LLMs Are Reducing Long-Context Costs…