https://openreview.net/forum?id=4W95topbWX
Fairness Failure Modes of Multimodal LLMs | OpenReview
Although Multimodal Large Language Models (MLLMs) are increasingly deployed in high-stakes domains, the fairness of their outputs is under-explored. Building...
failure modesmultimodal llmsfairnessopenreview
https://collaborate.princeton.edu/en/publications/charxiv-charting-gaps-in-realistic-chart-understanding-in-multimo/fingerprints/
CharXiv: Charting Gaps in Realistic Chart Understanding in Multimodal LLMs - Fingerprint -...
multimodal llmscharxivchartinggapsrealistic
https://neuronad.com/agle-from-nvidia-unveiled-mastering-multimodal-llms-with-mixtures-of-vision-encoders/
AGLE from Nvidia Unveiled: Mastering Multimodal LLMs with Mixtures of Vision Encoders - Neuronad -...
Aug 29, 2024 - New Study Reveals Optimized Design Strategies for Enhanced Visual Perception in Multimodal Models. Streamlined Design Approach: The study shows that...
multimodal llmsnvidiaunveiledmasteringmixtures
https://arxiv.org/abs/2508.04175
[2508.04175] AD-FM: Multimodal LLMs for Anomaly Detection via Multi-Stage Reasoning and...
Abstract page for arXiv paper 2508.04175: AD-FM: Multimodal LLMs for Anomaly Detection via Multi-Stage Reasoning and Fine-Grained Reward Optimization
multimodal llmsanomaly detectionadfmvia
https://lrec.elra.info/lrec2026-main-743
I Came, I Saw, I Explained: Benchmarking Multimodal LLMs on Figurative Meaning in Memes - LREC 2026...
May 1, 2026 - Internet memes represent a popular form of multimodal online communication and often use figurative elements to convey layered meaning through the combination o
multimodal llmscamesawexplainedbenchmarking
https://www.analyticsvidhya.com/blog/2025/03/top-multimodal-llms/
Top 10 Multimodal LLMs to Explore in 2026 - Analytics Vidhya
Jan 5, 2026 - Top 10 multimodal LLMs of 2026, from OpenAI to Google DeepMind, transforming AI with text, image, audio, and video capabilities.
multimodal llmsanalytics vidhyatopexplore
https://resources.altium.com/p/how-multimodal-llms-address-supply-chain-challenges
How Multimodal LLMs Address Supply Chain Challenges | Octopart | The Pulse
Mar 13, 2026 - Modern supply chains face complex challenges that traditional methods struggle to address. Multimodal Large Language Models (LLMs) like GPT-4, Gemini, and...
multimodal llmssupply chainthe pulseaddresschallenges
https://tldr.takara.ai/p/2405.11273
Uni-MoE: Scaling Unified Multimodal LLMs with Mixture of Experts | Takara TLDR
Recent advancements in Multimodal Large Language Models (MLLMs) underscore the significance of scalable models and data to boost performance, yet this often ...
multimodal llmsunimoescalingmixture
https://huggingface.co/papers/2406.18521
Paper page - CharXiv: Charting Gaps in Realistic Chart Understanding in Multimodal LLMs
Join the discussion on this paper page
multimodal llmspapercharxivchartinggaps
https://jobs.upfront.com/companies/bland-ai-2/jobs/75775275-machine-learning-researcher-multimodal-llms
Machine Learning Researcher, Multimodal LLMs @ Bland AI | Upfront Ventures Job Board
Search job openings across the Upfront Ventures network.
machine learning researchermultimodal llmsbland aiupfront venturesjob board
https://scholarsarchive.byu.edu/studentpub_uht/430/
"SEEING THE STORM: LEVERAGING MULTIMODAL LLMS FOR DISASTER SOCIAL MEDIA" by Holden Clark
Emergency management relies on the rapid triage of information to respond appropriately to disaster events. Social media platforms can provide emergency...
the stormmultimodal llmssocial mediaseeingleveraging
https://scholars.duke.edu/publication/1689286
Scholars@Duke publication: Probing Multimodal LLMs as World Models for Driving
multimodal llmsworld modelsscholarsdukepublication
https://news.y0.exchange/article/last-framework-enhances-spatial-reasoning-in-multimodal-llms
LAST Framework Enhances Spatial Reasoning in Multimodal LLMs | y0 News
Apr 14, 2026 - LAST framework integrates specialized vision tools to improve multimodal LLM spatial reasoning by 20%, outperforming proprietary closed-source systems.
spatial reasoningmultimodal llmslastframeworknews
https://exchart.github.io/
Making Multimodal LLMs Reliable Chart Data Extractors: A Benchmark and Training Framework
Leverage sparse visual information and contextual knowledge to extract sequences of stroke techniques in racket sports videos for tactical analysis.
multimodal llmsmakingreliablechartdata
https://wandb.ai/wandb_fc/product-announcements-fc/reports/Weave-newsletter-Multimodal-LLMs-RAG-tutorial-and-OpenAI-DevDay--Vmlldzo5ODUwNjE0
Weave newsletter: Multimodal LLMs, RAG tutorial, and OpenAI DevDay
multimodal llmsweavenewsletterragtutorial
https://scholar.hit.edu.cn/en/publications/vision-enhancing-llms-empowering-multimodal-knowledge-storage-and/
Vision Enhancing LLMs: Empowering Multimodal Knowledge Storage and Sharing in LLMs - Harbin...
visionenhancingllmsempoweringmultimodal
https://www.aibase.com/tool/31091
Visual Sketchpad-A visual reasoning tool for multimodal large language models (LLMs)
Visual Sketchpad is a framework that provides a visual sketchpad and drawing tools for multimodal large language models (LLMs). It allows models to operate on v
large language modelsvisualsketchpadreasoningtool
https://aclanthology.org/2024.findings-eacl.105/
Unified Embeddings for Multimodal Retrieval via Frozen LLMs - ACL Anthology
Ziyang Wang, Heba Elfardy, Markus Dreyer, Kevin Small, Mohit Bansal. Findings of the Association for Computational Linguistics: EACL 2024. 2024.
unifiedembeddingsmultimodalretrievalvia