VLC

Research code for "Training Vision-Language Transformers from Captions Alone"

// repository documentation