shreydan

shreydan/VisionGPT2

Combining ViT and GPT-2 for image captioning. Trained on MS-COCO. The model was implemented mostly from scratch.

⭐ 49 ⑂ 3 Jupyter Notebook · 2023-10-02推送
49
Watchers
0
贡献者
0
Commits
0
Releases
0
Open Issues
2023-10-02
最近推送
原文 中文
暂无 README