Powered by OpenAIRE graph
Found an issue? Give us feedback
image/svg+xml art designer at PLoS, modified by Wikipedia users Nina, Beao, JakobVoss, and AnonMoos Open Access logo, converted into svg, designed by PLoS. This version with transparent background. http://commons.wikimedia.org/wiki/File:Open_Access_logo_PLoS_white.svg art designer at PLoS, modified by Wikipedia users Nina, Beao, JakobVoss, and AnonMoos http://www.plos.org/ ZENODOarrow_drop_down
image/svg+xml art designer at PLoS, modified by Wikipedia users Nina, Beao, JakobVoss, and AnonMoos Open Access logo, converted into svg, designed by PLoS. This version with transparent background. http://commons.wikimedia.org/wiki/File:Open_Access_logo_PLoS_white.svg art designer at PLoS, modified by Wikipedia users Nina, Beao, JakobVoss, and AnonMoos http://www.plos.org/
ZENODO
Report
Data sources: ZENODO
addClaim

To what extent does modality imbalance affect the accuracy and routing stability of multimodal language models

Authors: SOVEREIGN Research Kernel;

To what extent does modality imbalance affect the accuracy and routing stability of multimodal language models

Abstract

The rise of Multimodal Large Language Models (MLLMs) has significantly advanced the capabilities of AI systems to understand and generate content across diverse modalities such as text, images, audio, video, and sensory data. By leveraging the reasoning prowess of Large Language Models (LLMs), MLLMs unify multiple input formats into a coherent framework, enabling unprecedented performance in multimodal tasks. This survey provides a comprehensive overview of the architectural innovations, training paradigms, data resources, and evaluation benchmarks that have shaped the evolution of MLLMs. We rResearch goal: To what extent does modality imbalance affect the accuracy and routing stability of multimodal language models as measured by performance on MMBench and SEED-Bench evaluation suites?Autonomous synthesis report generated by SOVEREIGN Research Kernel. Tribunal consensus score: 7.8/10.

Powered by OpenAIRE graph
Found an issue? Give us feedback