QVLM
[NeurIPS'24]Efficient and accurate memory saving method towards W4A4 large multi-modal models.
// repository documentation
Was this content helpful?
(0 ratings)