Disambiguation evidence for ViT via surface form
"An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale"
Triples (1)
Triples where some other subject referred to this entity
as "An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale".
ViT
→
introducedInPaper
→
"An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale"
ⓘ
↳ resolves to ViT