multimodalcritiquebearishExisting disaster response datasets like Incidents1M and CrisisMMD suffer from either complete lack of text or severe text-image semantic misalignmentComputer Vision02 Aug 2026http://arxiv.org/abs/2607.28269v1