XunZi-MLLM: a multimodal large language model for ancient text and image recognition.

Photocopies of ancient works, as valuable cultural heritage of China, can be digitized through the integration of multimodal large language models (MLLMs). This approach allows for more vivid representations of these historical documents, fostering the preservation and advancement of traditional cul...

Full description

Bibliographic Details
Published in:Digital Scholarship in the Humanities Vol. 40; no. 2; pp. 709 - 723
Main Authors: Zhu, Dongmei, Liu, Chang, Zhao, Xue, Zhao, Zhixiao, Shen, Si, Wang, Dongbo
Format: Article
Published: Oxford University Press / USA Jun2025
Subjects:
Online Access:View this record in EBSCOhost