Like a Baby: Visually Situated Neural Language Acquisition
DGX agentarXiv:1805.11546v3 Announce Type: replace-cross Abstract: We examine the benefits of visual context in training neural language models to perform next-word prediction. A multi-modal neural architectur