VisionTransformer
A complete easy to follow implementation of Google's Vision Transformer proposed in "AN IMAGE IS WORTH 16X16 WORDS". This pytorch implementation has comments for better understanding.
파일 탐색기
최종 버전 다운로드 (.zip)- Google_ViT.py
- README.md
- ViT.png
// repository documentation
Was this content helpful?
(0 ratings)
