๊ฐ€์ž… ํ›„ ์ดˆ๋Œ€ ๋งํฌ๋ฅผ ๊ณต์œ ํ•˜๋ฉด ๋™์˜์ƒ ์žฌ์ƒ ๋ฐ ์ดˆ๋Œ€ ๋ณด์ƒ์„ ๋ฐ›์„ ์ˆ˜ ์žˆ์Šต๋‹ˆ๋‹ค.

Tencent Hy
@TencentHunyuan
Tencent's foundation model for text, image, video, and 3D generation.
๊ฐ€์ž… July 2024
8 ํŒ”๋กœ์ž‰ ์ค‘    46.5K ํŒฌ
๐ŸŽ‰ ๐ŸŽ‰ ๐ŸŽ‰ We're open-sourcing Chronicles-OCR, a visual perception benchmark evaluating VLLMs on ancient Chinese characters. The dataset spans 3,000 years of evolution. It covers 7 historical scripts from Oracle Bone to Cursive, featuring 2,800 balanced images across highly diverse physical media. We assess models on 4 core tasks: โ€ข Character Spotting โ€ข Fine-grained Recognition โ€ข Ancient Text Parsing โ€ข Script Classification The evaluation reveals how visual distribution shifts affect model perception over time. Explore the dataset and paper below. ๐Ÿ‘‡ ๐Ÿ“„ Paper: ๐Ÿ”— GitHub:
๋” ๋ณด๊ธฐ