[IEEE] Exploring KV Cache Quantization in Multimodal Large Language Model Inference

dnn___ Post time 11 hour(s) ago | Show all posts |Read mode
This post will be closed automatically in 2026-08-27 00:11
Reward30points

journalㄩIEEE Computer Architecture Letters

AuthorsㄩHyesung Ahn; Ranggi Hwang; Minsoo Rhu

Published dateㄩ2026-1-

DOIㄩ10.1109/lca.2025.3646170

PDF linkㄩhttps://ieeexplore.ieee.org/stampPDF/getPDF.jsp?arnumber=11304543

Article linkㄩhttps://doi.org/10.1109/lca.2025.3646170

Article SourceㄩInstitute of Electrical and Electronics Engineers (IEEE)


Remarkㄩ

Best Answer

Please approve

View Full Content

Reply

Use magic Donate Report

All Reply1 Show all posts
jakir_ete_ruet Post time 11 hour(s) ago | Show all posts

This post has been completed

Completed attachments will be deleted within 24 hours.
Reply

Use magic Donate Report

Junior Member
  • post

  • reply

  • points

    170

Same Category


Return to the list