LLMOCR

Simple script that reads an image and dumps the text it reads using a vision model and KobolodCPP

// repository documentation