Do MLLMs really understand low-resource Khmer documents? A pilot study on Khmer document VQA
Read the original at arxiv.org→arXiv:2608.28635v1 Announce Type: new Abstract: Recent multimodal large language models (MLLMs) have advanced document understanding, visual question answering, and text extraction. However, their reliability in...
Original headline: "Do MLLMs Really Understand Low-Resource Khmer Documents? A Pilot Study on Khmer Document VQA"
Coverage timeline
- Sep 1, 04:00 UTC arXiv cs.CL lead source Do MLLMs Really Understand Low-Resource Khmer Documents? A Pilot Study on Khmer Document VQA