167

BIRD: Bronze Inscription Restoration and Dating

Main:5 Pages
5 Figures
Bibliography:3 Pages
13 Tables
Appendix:5 Pages
Abstract

Bronze inscriptions from early China are fragmentary and difficult to date. We introduce BIRD(Bronze Inscription Restoration and Dating), a fully encoded dataset grounded in standard scholarly transcriptions and chronological labels. We further propose an allograph-aware masked language modeling framework that integrates domain- and task-adaptive pretraining with a Glyph Net (GN), which links graphemes and allographs. Experiments show that GN improves restoration, while glyph-biased sampling yields gains in dating.

View on arXiv
Comments on this paper