XunZi-MLLM: a multimodal large language model for ancient text and image recognition.

Photocopies of ancient works, as valuable cultural heritage of China, can be digitized through the integration of multimodal large language models (MLLMs). This approach allows for more vivid representations of these historical documents, fostering the preservation and advancement of traditional cul...

Descripción completa

Detalles Bibliográficos
Publicado en:Digital Scholarship in the Humanities Vol. 40; no. 2; pp. 709 - 723
Autores principales: Zhu, Dongmei, Liu, Chang, Zhao, Xue, Zhao, Zhixiao, Shen, Si, Wang, Dongbo
Formato: Artículo
Publicado: Oxford University Press / USA Jun2025
Materias:
Acceso en línea:Ver este registro en EBSCOhost