Hi,
If you ever get to writing a blog post about kv cache quantisation, i'm interested in quantising K differently than V