C3PO benchmark evaluates cross-modal composition and counterfactual conflict in omnimodal models
Read the original at arxiv.org→arXiv:2608.05381v1 Announce Type: new Abstract: Current Multimodal Large Language Models (MLLMs) can process diverse sensory inputs, yet their reasoning remains heavily biased toward a dominant modality, resulting...
Original headline: "C$^3$PO: Evaluating Cross-Modal Composition and Counterfactual Performance in Omnimodal Models"
Coverage timeline
- Aug 7, 04:00 UTC arXiv cs.AI lead source C$^3$PO: Evaluating Cross-Modal Composition and Counterfactual Performance in Omnimodal Models