Qwen2-VL-OCR-VQA
This project demonstrates how to use the Qwen2-VL model from Hugging Face for Optical Character Recognition (OCR) and Visual Question Answering (VQA). The model combines vision and language capabilities, enabling users to analyze images and generate context-based responses.
// repository documentation
Was this content helpful?
(0 ratings)