MedProb probes internal representations of vision-language models for medical question answering
Read the original at arxiv.org→arXiv:2609.04336v1 Announce Type: new Abstract: Medical visual question answering (Med-VQA) is often assumed to require medical fine-tuning, large models, or complex multi-agent pipelines. We revisit this assumption...
Original headline: "MedProb: Probing Internal Representations of Vision-Language Models for Medical Question Answering"
Coverage timeline
- Sep 7, 04:00 UTC arXiv cs.CL lead source MedProb: Probing Internal Representations of Vision-Language Models for Medical Question Answering