🎉 🎉 🎉 We're open-sourcing Chronicles-OCR, a visual perception benchmark evaluating VLLMs on ancient Chinese characters.
The dataset spans 3,000 years of evolution. It covers 7 historical scripts from Oracle Bone to Cursive, featuring 2,800 balanced images across highly diverse physical media.
We assess models on 4 core tasks:
• Character Spotting
• Fine-grained Recognition
• Ancient Text Parsing
• Script Classification
The evaluation reveals how visual distribution shifts affect model perception over time.
Explore the dataset and paper below. 👇
📄 Paper:
🔗 GitHub: