ocr-vqgan

(★ 86)

OCR-VQGAN, a discrete image encoder (tokenizer and detokenizer) for figure images in Paper2Fig100k dataset. Implementation of OCR Perceptual loss for clear text-within-image generation. Fork from VQGAN in CompVis/taming-transformers

  • .gitignore
  • environment.yaml
  • main.py
  • README.md
  • setup.py
// repository documentation