注册并分享邀请链接,可获得视频播放与邀请奖励。

AI at Meta
@AIatMeta
Together with the AI community, we are pushing the boundaries of what’s possible through open science to create a more connected world.
加入 August 2018
367 正在关注    856.1K 粉丝
📝 New from FAIR: An Introduction to Vision-Language Modeling. Vision-language models (VLMs) are an area of research that holds a lot of potential to change our interactions with technology, however there are many challenges in building these types of models. Together with a set of collaborators across academia, we’re releasing ‘An Introduction to Vision-Language Modeling’ — we hope that this new resource will help anyone who would like to enter this field to better understand the mechanics behind mapping vision to language. Full paper ➡️ This guide covers how VLMs work, how to train them and approaches to evaluation — and while it primarily covers mapping image to language, it also discusses how to extend VLMs to videos. We hope that releasing this guide will inspire and enable more work in this space.
显示更多
0
30
2K
476
转发到社区