QVLM

★ 102 Open GitHub ↗

[NeurIPS'24]Efficient and accurate memory saving method towards W4A4 large multi-modal models.

// repository documentation