๊ฐ€์ž… ํ›„ ์ดˆ๋Œ€ ๋งํฌ๋ฅผ ๊ณต์œ ํ•˜๋ฉด ๋™์˜์ƒ ์žฌ์ƒ ๋ฐ ์ดˆ๋Œ€ ๋ณด์ƒ์„ ๋ฐ›์„ ์ˆ˜ ์žˆ์Šต๋‹ˆ๋‹ค.

Tony Wu
@tonywu_71
Multimodal, RAG, Agents | ColPali co-first author | @centralesupelec ๐Ÿ‡ซ๐Ÿ‡ท x @Cambridge_Uni ๐Ÿ‡ฌ๐Ÿ‡ง | Core Researcher at @hcompany_ai ๐Ÿง‘๐Ÿปโ€๐Ÿ’ป
๊ฐ€์ž… February 2022
573 ํŒ”๋กœ์ž‰ ์ค‘    1.6K ํŒฌ
๐Ÿ‘€ Meet NeoMME: a family of 260M and 800M Multimodal-Native Multilingual efficient Encoders One bidirectional Transformer processes text tokens and raw image patches, with no pretrained vision tower, text encoder, or decoder. (1/N ๐Ÿงต)
๋” ๋ณด๊ธฐ