XunZi-MLLM: a multimodal large language model for ancient text and image recognition.
Photocopies of ancient works, as valuable cultural heritage of China, can be digitized through the integration of multimodal large language models (MLLMs). This approach allows for more vivid representations of these historical documents, fostering the preservation and advancement of traditional cul...
| Published in: | Digital Scholarship in the Humanities Vol. 40; no. 2; pp. 709 - 723 |
|---|---|
| Main Authors: | , , , , , |
| Format: | Article |
| Published: |
Oxford University Press / USA
Jun2025
|
| Subjects: | |
| Online Access: | View this record in EBSCOhost |