遇见数据集

Learning Words from Images and Speech

收藏
Figshare2016-01-19 更新2026-04-08 收录
官方服务:

资源简介:

This paper explores the possibility to learn a semantically-relevant lexicon from images and speech only. For this, we train a multi-modal neural network working both on image fragments and on speech features, by learning an embedding in which images and content words that co-occur together are close. Making no assumption on the acoustic model, this paper shows promising results on how multi-modality could help word learning.

创建时间:
2014-12-23
二维码
社区交流群
二维码
科研交流群
商业服务