๐ ๐ ๐ We're open-sourcing Chronicles-OCR, a visual perception benchmark evaluating VLLMs on ancient Chinese characters.
The dataset spans 3,000 years of evolution. It covers 7 historical scripts from Oracle Bone to Cursive, featuring 2,800 balanced images across highly diverse physical media.
We assess models on 4 core tasks:
โข Character Spotting
โข Fine-grained Recognition
โข Ancient Text Parsing
โข Script Classification
The evaluation reveals how visual distribution shifts affect model perception over time.
Explore the dataset and paper below. ๐
๐ Paper:
๐ GitHub: