https://www.comfortdelgro.com/
Leading multi-modal transport operator - ComfortDelGro
Apr 2, 2026 - ComfortDelGro operates in 13 countries and has a global network of over 55,000 vehicles.
multi modal transportleadingoperatorcomfortdelgro
https://openreview.net/forum?id=pypsPHQCby
DPO-Finetuned Large Multi-Modal Planner with Retrieval-Augmented Generation @ EgoPlan Challenge...
This paper presents technical details for solving a multi-modal task, EgoPlan-Bench. Our model adopts Direct Preference Optimization (DPO), which is originally...
large multi
https://www.frontiersin.org/journals/medicine/articles/10.3389/fmed.2023.1342374/full
Frontiers | Editorial: Multi-modal learning and its application for biomedical data
With the rapid development of biomedical testing methods and the explosive growth of biomedical data, multimodal data can better meet the precise diagnosis o...
multi modalfrontierseditoriallearning
https://bookcreator.com/2015/09/using-book-creator-to-create-multi-modal-factual-texts/
Using Book Creator to create multi-modal factual texts - Book Creator app
Oct 13, 2022 - This primary teacher details his app smash adventure with no less than 6 apps used to make their ebooks, held together seamlessly with Book Creator.
book creatormulti modalusingcreatefactual
https://arxiv.org/html/2503.13940v1
Multi-Modal Self-Supervised Semantic Communication
multi modalself supervisedsemanticcommunication
https://www.ageotype.com/
Ageotype Portal (Beta) - Multi-Modal Health Data Integration
An AI-powered longevity data environment integrating MRI, metabolomics, multi-cancer screening, and brain-age reports into a single intelligent view.
multi modalhealth dataageotypeportalbeta
https://openreview.net/forum?id=HnNCSFuyRy
Windows Agent Arena: Evaluating Multi-Modal OS Agents at Scale | OpenReview
Large language models (LLMs) show remarkable potential to act as computer agents, enhancing human productivity and software accessibility in multi-modal tasks...
windows agent arenamulti modalat scaleevaluating
https://www.purdue.edu/innovativelearning/teaching/module/multi-modal-ai/
Multi-Modal AI - Teaching@Purdue
Aug 20, 2024 - Natural language processing of audio files has been used quite often in the last decade as the quality has continued to scale with computing power. In 2023,...
multi modalai teachingpurdue
https://deepai.org/publication/learning-multi-modal-similarity
Learning Multi-modal Similarity | DeepAI
Aug 30, 2010 - 08/30/10 - In many applications involving multi-media data, the definition of similarity between items is integral to several key tasks, e.g....
multi modallearningsimilaritydeepai
https://www.mdpi.com/2079-8954/12/1/7
Enhancing Multi-Modal Perception and Interaction: An Augmented Reality Visualization System for...
Visualization systems play a crucial role in industry, education, and research domains by offering valuable insights and enhancing decision making. These...
multi modal perception
https://openreview.net/forum?id=SulRfnEVK4&referrer=%5Bthe%20profile%20of%20Yuekai%20Sun%5D(%2Fprofile%3Fid%3D~Yuekai_Sun1)
LiveXiv - A Multi-Modal live benchmark based on Arxiv papers content | OpenReview
The large-scale training of multi-modal models on data scraped from the web has shown outstanding utility in infusing these models with the required world...
multi modal
https://arxiv.org/abs/2108.05603
[2108.05603] Multi-Modal MRI Reconstruction Assisted with Spatial Alignment Network
Abstract page for arXiv paper 2108.05603: Multi-Modal MRI Reconstruction Assisted with Spatial Alignment Network
multi modal2108mri
https://openreview.net/forum?id=2ZHKA9xo8V&referrer=%5Bthe%20profile%20of%20Matthias%20Fey%5D(%2Fprofile%3Fid%3D~Matthias_Fey2)
PyTorch Frame: A Modular Framework for Multi-Modal Tabular Learning | OpenReview
We present PyTorch Frame, a PyTorch-based framework for deep learning over multi-modal tabular data. PyTorch Frame makes tabular deep learning easy by...
modular frameworkmulti modalpytorch
https://arxiv.org/abs/2012.13755
[2012.13755] Probabilistic 3D Multi-Modal, Multi-Object Tracking for Autonomous Driving
Abstract page for arXiv paper 2012.13755: Probabilistic 3D Multi-Modal, Multi-Object Tracking for Autonomous Driving
multi modalobject tracking2012probabilistic3d
https://www.devdiscourse.com/news?tag=multi-modal+AI+traffic+prediction
multi-modal AI traffic prediction News | Devdiscourse
Latest News on multi-modal AI traffic prediction, Read more information on multi-modal AI traffic prediction
ai traffic predictionmulti modalnewsdevdiscourse
https://arxiv.org/abs/2604.10708
[2604.10708] Audio-Omni: Extending Multi-modal Understanding to Versatile Audio Generation and...
Abstract page for arXiv paper 2604.10708: Audio-Omni: Extending Multi-modal Understanding to Versatile Audio Generation and Editing
multi modal
https://openreview.net/forum?id=DgGF2LEBPS&referrer=%5Bthe%20profile%20of%20Huan%20Zhang%5D(%2Fprofile%3Fid%3D~Huan_Zhang1)
EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven...
Leveraging Multi-modal Large Language Models (MLLMs) to create embodied agents offers a promising avenue for tackling real-world tasks. While language-centric...
large language modelsmulti modalcomprehensivebenchmarking
https://www.gryfn.io/
GRYFN | Multi-Modal Sensing
GRYFN advanced sensing technology delivers data and analytics to drive field research.
multi modalsensing
https://arxiv.org/abs/2110.08080v1
[2110.08080v1] Multi-modal Aggregation Network for Fast MR Imaging
Abstract page for arXiv paper 2110.08080v1: Multi-modal Aggregation Network for Fast MR Imaging
multi modal2110aggregationnetworkfast
https://openreview.net/forum?id=ybw2U70q_Vd
End-to-end Multi-modal Video Temporal Grounding | OpenReview
We propose a multi-modal framework for text-guided video temporal grounding by learning complementary visual features from RGB, optical flow and depth...
multi modalendvideotemporalgrounding
https://www.bruker.com/ko/products-and-solutions/preclinical-imaging/nmi.html
Molecular Imaging | Multi-Modal Imaging | Bruker
Bruker brings you the World of Molecular Imaging. High resolution modular benchtop PET, SPECT and CT, and hybrid PET/MR, PET/CT, PET/SPECT/CT and PMOD...
molecular imagingmulti modalbruker
https://openreview.net/forum?id=Zg4Onwz46wI&referrer=%5Bthe%20profile%20of%20Kiran%20Premdat%20Kokilepersaud%5D(%2Fprofile%3Fid%3D~Kiran_Premdat_Kokilepersaud1)
Multi-Modal Learning Using Physicians Diagnostics for Optical Coherence Tomography Classification |...
In this paper, we propose a framework that incorporates experts diagnostics and insights into the analysis of Optical Coherence Tomography (OCT) using...
optical coherence tomographymulti modallearningusingphysicians
https://deepai.org/publication/ensemble-manifold-based-regularized-multi-modal-graph-convolutional-network-for-cognitive-ability-prediction
Ensemble manifold based regularized multi-modal graph convolutional network for cognitive ability...
Jan 20, 2021 - 01/20/21 - Objective: Multi-modal functional magnetic resonance imaging (fMRI) can be used to make predictions about individual behavioral an...
graph convolutional networkmulti modal
https://www.preprints.org/manuscript/202508.1271/v1
Autonomous Rescue Drone with Multi-Modal Vision and Cognitive Agentic Architecture[v1] |...
In post-disaster search and rescue (SAR) operations, unmanned aerial vehicles (UAVs) are essential tools, yet the large volume of raw visual data often...
multi modal
https://openreview.net/forum?id=Ev4iw23gdI
EMMA: Empowering Multi-modal Mamba with Structural and Hierarchical Alignment | OpenReview
Mamba-based architectures have shown to be a promising new direction for deep learning models owing to their competitive performance and sub-quadratic...
multi modalhierarchical alignmentemmaempoweringmamba
https://arxiv.org/abs/2401.16158v2
[2401.16158v2] Mobile-Agent: Autonomous Multi-Modal Mobile Device Agent with Visual Perception
Abstract page for arXiv paper 2401.16158v2: Mobile-Agent: Autonomous Multi-Modal Mobile Device Agent with Visual Perception
mobile agentmulti modal2401autonomous
https://aclanthology.org/2022.findings-emnlp.437/
Named Entity and Relation Extraction with Multi-Modal Retrieval - ACL Anthology
Xinyu Wang, Jiong Cai, Yong Jiang, Pengjun Xie, Kewei Tu, Wei Lu. Findings of the Association for Computational Linguistics: EMNLP 2022. 2022.
named entityrelation extractionmulti modal
https://www.preprints.org/manuscript/202507.2394
Federated Learning-Enabled Secure Multi-Modal Anomaly Detection for Wire Arc Additive...
This paper presents a federated learning (FL) architecture tailored for anomaly detection in wire arc additive manufacturing (WAAM) that preserves data privacy...
federated learningmulti modal
https://aclanthology.org/2022.emnlp-main.449/
META-GUI: Towards Multi-modal Conversational Agents on Mobile GUI - ACL Anthology
Liangtai Sun, Xingyu Chen, Lu Chen, Tianle Dai, Zichen Zhu, Kai Yu. Proceedings of the 2022 Conference on Empirical Methods in Natural Language Processing....
multi modalconversational agentsmetaguitowards
https://arxiv.org/abs/2509.25946
[2509.25946] Automated Model Discovery via Multi-modal & Multi-step Pipeline
multi modalautomatedmodeldiscoveryvia
https://arxiv.org/abs/2106.11473
[2106.11473] Sequential Late Fusion Technique for Multi-modal Sentiment Analysis
Abstract page for arXiv paper 2106.11473: Sequential Late Fusion Technique for Multi-modal Sentiment Analysis
multi modal2106sequentiallatefusion
https://www.biomodality.com/
Biomodality | Multi-Modal Biometric Verification
Biomodality offers face recognition, voice recognition, and a unique approach to liveness: Interactive Liveness Verification (ILV). All delivered through one...
multi modalbiometricverification
https://deepai.org/publication/omninet-a-unified-architecture-for-multi-modal-multi-task-learning
OmniNet: A unified architecture for multi-modal multi-task learning | DeepAI
Jul 17, 2019 - 07/17/19 - Transformer is a popularly used neural network architecture, especially for language understanding. We introduce an extended and u...
a unified architecturemulti modalomninettasklearning
https://www.amazon.science/publications/multi-modal-embeddings-using-multi-task-learning-for-emotion-recognition
Multi-modal embeddings using multi-task learning for emotion recognition - Amazon Science
General embeddings like word2vec, GloVe and ELMo have shown a lot of success in natural language tasks. The embeddings are typically extracted from models that...
multi modalemotion recognitionembeddingsusingtask
https://deepai.org/publication/beyond-first-impressions-integrating-joint-multi-modal-cues-for-comprehensive-3d-representation
Beyond First Impressions: Integrating Joint Multi-modal Cues for Comprehensive 3D Representation |...
Aug 6, 2023 - 08/06/23 - In recent years, 3D representation learning has turned to 2D vision-language pre-trained models to overcome data scarcity challeng...
first impressionsmulti modal
https://www.engadget.com/chinas-gigantic-multi-modal-ai-is-no-one-trick-pony-211414388.html?ref=alessiopomaro.it
China's gigantic multi-modal AI is no one-trick pony
Jun 2, 2021 - Researchers from the Beijing Academy of Artificial Intelligence announced on Tuesday the release of Wu Dao, a mammoth AI seemingly capable of doing everything...
china smulti modalno onegigantic
https://openreview.net/forum?id=JsBmdzLZqX
Multi-Modal Medical Image Augmentation for Controlled Heterogeneity and Fair Outcomes | OpenReview
Limited data in medical imaging exacerbate class imbalance and fairness gaps, undermining deep-learning across diverse patient subgroups. GAN- and...
multi modalmedical image
https://www.bu.edu/igs/2020/06/01/multi-modal-travel/
Multi-modal Travel | Institute for Global Sustainability
multi modalinstitute fortravelglobalsustainability
https://deepai.org/publication/social-vrnn-one-shot-multi-modal-trajectory-prediction-for-interacting-pedestrians
Social-VRNN: One-Shot Multi-modal Trajectory Prediction for Interacting Pedestrians | DeepAI
Oct 18, 2020 - 10/18/20 - Prediction of human motions is key for safe navigation of autonomous robots among humans. In cluttered environments, several motio...
one shotmulti modal
https://www.tcs.com/insights/blogs/multi-modal-product-portfolio-planning-cdlm-system
A Multi-modal Product Portfolio Planning and CDLM System
TCS' multi-modal planning product that supports agile frameworks can help organizations meet the rising demands of digital transformation. Learn more.
multi modalproduct portfolioplanningcdlmsystem
https://seedance-2.vip/
Seedance 2.0 VIP - Revolutionary Multi-Modal AI Video Generation
Create stunning AI-powered videos from images and text. Multi-modal input, precise motion control, and built-in audio generation.
seedance 2 0multi modalai videoviprevolutionary
https://aclanthology.org/2025.findings-acl.837/
Training Multi-Modal LLMs through Dialogue Planning for HRI - ACL Anthology
Claudiu Daniel Hromei, Federico Borazio, Andrea Sensi, Elisa Passone, Danilo Croce, Roberto Basili. Findings of the Association for Computational Linguistics:...
multi modalplanning fortrainingllms
https://research.google/pubs/alf-advertiser-large-foundation-model-for-multi-modal-advertiser-understanding/
ALF: Advertiser Large Foundation Model for Multi-Modal Advertiser Understanding
foundation modelmulti modalalfadvertiserlarge
https://grably.us/
High-quality multi-modal human interaction and conversational datasets | Grably
Grably is a multi-modal human interaction data research company trusted by leading AI labs and big-tech companies.
high qualitymulti modalhuman interactionconversationaldatasets
https://deepai.org/publication/dasc-robust-dense-descriptor-for-multi-modal-and-multi-spectral-correspondence-estimation
DASC: Robust Dense Descriptor for Multi-modal and Multi-spectral Correspondence Estimation | DeepAI
Apr 27, 2016 - 04/27/16 - Establishing dense correspondences between multiple images is a fundamental task in many applications. However, finding a reliable...
multi modal
https://easychair.org/publications/preprint/kfrT
Multi-Modal Co-Training for Fake News Identification Using Attention-Aware Fusion
multi modalco trainingfake news
https://openreview.net/forum?id=3MnMGLctKb
Multi-Modal and Multi-Attribute Generation of Single Cells with CFGen | OpenReview
Generative modeling of single-cell RNA-seq data is crucial for tasks like trajectory inference, batch effect removal, and simulation of realistic cellular...
multi modalsingle cellsattributegeneration
https://openreview.net/forum?id=8B3sAX889P&referrer=%5Bthe%20profile%20of%20Qiannan%20Zhang%5D(%2Fprofile%3Fid%3D~Qiannan_Zhang1)
Unified Insights: Harnessing Multi-modal Data for Phenotype Imputation via View Decoupling |...
Phenotype imputation plays a crucial role in improving comprehensive and accurate medical evaluation, which in turn can optimize patient treatment and bolster...
unified insightsmulti modaldata for
https://openreview.net/forum?id=rtdn6GHiLo
Leveraging Multi-Modal Saliency and Fusion for Gaze Target Detection | OpenReview
Gaze target detection (GTD) is the task of predicting where a person in an image is looking. This is a challenging task, as it requires the ability to...
multi modaltarget detectionleveragingsaliency
https://arxiv.org/abs/2207.02159v1
[2207.02159v1] Multi-modal Robustness Analysis Against Language and Visual Perturbations
Abstract page for arXiv paper 2207.02159v1: Multi-modal Robustness Analysis Against Language and Visual Perturbations
multi modal2207robustness
https://pmc.ncbi.nlm.nih.gov/articles/PMC11984272/
Multi-modal management of aggressive vertebral hemangioma: A single center experience - PMC
This study aims at spotlighting different lines of management of aggressive vertebral hemangioma (VH) through a retrospective analysis of single center...
multi modalvertebral hemangioma
https://openreview.net/forum?id=U4vaF1oQE9&referrer=%5Bthe%20profile%20of%20Jiechao%20Gao%5D(%2Fprofile%3Fid%3D~Jiechao_Gao4)
DRUM: Learning Demonstration Retriever for Large MUlti-modal Models | OpenReview
Recently, large language models (LLMs) have demonstrated impressive capabilities in dealing with new tasks with the help of in-context learning (ICL). In the...
large multidrumlearningdemonstrationretriever
https://www.analyticsvidhya.com/blog/2023/12/multi-modal-rag-pipeline-with-langchain/
Building a Multi-Modal RAG Pipeline with Langchain - Analytics Vidhya
Jan 7, 2025 - "Unlock the power of Multi-Modal RAG with Langchain! Learn to build pipelines using Gemini Pro Vision, Chroma vector database, and Langchain.
building amulti modalrag pipelinelangchainanalytics
https://aclanthology.org/2021.findings-acl.323/
Probing Multi-modal Machine Translation with Pre-trained Language Model - ACL Anthology
Kong Yawei, Kai Fan. Findings of the Association for Computational Linguistics: ACL-IJCNLP 2021. 2021.
multi modalmachine translation
https://www.databricks.com/dataaisummit/session/data-analytics-multi-modal-data
Data Analytics with Multi-Modal Data | Databricks
data analyticsmulti modaldatabricks
https://deepai.org/publication/scene-induced-multi-modal-trajectory-forecasting-via-planning
Scene Induced Multi-Modal Trajectory Forecasting via Planning | DeepAI
May 23, 2019 - 05/23/19 - We address multi-modal trajectory forecasting of agents in unknown scenes by formulating it as a planning problem. We present an a...
multi modalsceneinducedtrajectoryforecasting
https://arxiv.org/abs/2503.03509v2
[2503.03509v2] Sampling-Based Multi-Modal Multi-Robot Multi-Goal Path Planning
Abstract page for arXiv paper 2503.03509v2: Sampling-Based Multi-Modal Multi-Robot Multi-Goal Path Planning
multi modal2503samplingbasedrobot
https://www.mdpi.com/1420-3049/30/9/1983
Multi-Modal Design, Synthesis, and Biological Evaluation of Novel Fusidic Acid Derivatives
Fusidic acid (FA), a tetracyclic triterpenoid, has been approved to treat methicillin-resistant Staphylococcus aureus (MRSA) infections. However, there are few...
multi modalbiological evaluation
https://deepai.org/publication/commuting-conditional-gans-for-robust-multi-modal-fusion
Commuting Conditional GANs for Robust Multi-Modal Fusion | DeepAI
Jun 10, 2019 - 06/10/19 - This paper presents a data driven approach to multi-modal fusion, where optimal features for each sensor are selected from a commo...
conditional gansmulti modalcommutingrobustfusion
https://www.rca.ac.uk/research-innovation/research-centres/rca-robotics-laboratory/multi-modal-sensing/
Multi-modal Sensing | Royal College of Art
multi modalroyal collegesensingart
https://www.notion.so/Multi-Modal-Manipulation-via-Policy-Consensus-28645464339980d6a8a4ca588dec3905
Multi-Modal Manipulation via Policy Consensus | Notion
Why Feature Concatenation Fails for Robots That See and Feel?
multi modalmanipulationviapolicyconsensus
https://openreview.net/forum?id=Z1rbS7ZS32&referrer=%5Bthe%20profile%20of%20Corrado%20Pezzato%5D(%2Fprofile%3Fid%3D~Corrado_Pezzato1)
Multi-Modal MPPI and Active Inference for Reactive Task and Motion Planning | OpenReview
Task and Motion Planning (TAMP) has made strides in complex manipulation tasks, yet the execution robustness of the planned solutions remains overlooked. In...
multi modalactive inference
https://openreview.net/forum?id=jTh3rdEF3LH
HUM3DIL: Semi-supervised Multi-modal 3D HumanPose Estimation for Autonomous Driving | OpenReview
3D Human Pose estimation from RGB + LiDAR data, for autonomous driving.
multi modal
https://openreview.net/forum?id=fdV-GZ4LPfn
Multi-modal Self-supervised Pre-training for Large-scale Genome Data | OpenReview
Open genomic regions, being accessible to regulatory proteins, could act as the on/off switch or amplifier/attenuator of gene expression, and thus reflects the...
multi modalself supervisedpre training
https://pmc.ncbi.nlm.nih.gov/articles/PMC11071495/
Projection-TAGs enable multiplex projection tracing and multi-modal profiling of projection neurons...
Single-cell multiomic techniques have sparked immense interest in developing a comprehensive multi-modal map of diverse neuronal cell types and their...
multi modalprojectiontagsenablemultiplex
https://openreview.net/forum?id=IWWWulAX7g
Fine-grained Late-interaction Multi-modal Retrieval for Retrieval Augmented Visual Question...
Knowledge-based Visual Question Answering (KB-VQA) requires VQA systems to utilize knowledge from external knowledge bases to answer visually-grounded...
fine grainedmulti modallateinteraction
https://huggingface.co/collections/jeffreyschultz/multi-modal
Multi-modal - a jeffreyschultz Collection
Unlock the magic of AI with handpicked models, awesome datasets, papers, and mind-blowing Spaces from jeffreyschultz
multi modalcollection
https://arxiv.org/abs/1510.05318v1
[1510.05318v1] Latent Space Model for Multi-Modal Social Data
Abstract page for arXiv paper 1510.05318v1: Latent Space Model for Multi-Modal Social Data
latent spacemulti modal1510modelsocial
https://www.easychair.org/publications/keyword/NRdQ
Keyword: Multi-modal Safety Compliance Checking
multi modalsafety compliancekeywordchecking
https://aclanthology.org/volumes/2025.evalmg-1/
Proceedings of the First Workshop of Evaluation of Multi-Modal Generation - ACL Anthology
of thefirst workshopmulti modalproceedings
https://oecd.ai/en/catalogue/metric-use-cases/beyond-first-impressions-integrating-joint-multi-modal-cues-for-comprehensive-3d-representation
Beyond First Impressions: Integrating Joint Multi-modal Cues for Comprehensive 3D Representation -...
This work proposes a syntax-enhanced grammatical error correction (GEC) approach named SynGEC that effectively incorporates dependency syntactic information...
first impressionsmulti modal
https://aclanthology.org/2026.findings-eacl.7/
Plane Geometry Problem Solving with Multi-modal Reasoning: A Survey - ACL Anthology
Seunghyuk Cho, Zhenyue Qin, Yang Liu, Youngbin Choi, Seungbeom Lee, Dongwoo Kim. Findings of the Association for Computational Linguistics: EACL 2026. 2026.
plane geometryproblem solvingmulti modal
https://surgsync.github.io/
SurgSync: Time-Synchronized Multi-modal Data Collection
SurgSync: Time-Synchronized Multi-modal Data Collection Framework and Dataset for Surgical Robotics. ICRA 2026.
multi modaltimesynchronizeddatacollection
https://deepai.org/publication/mopa-multi-modal-prior-aided-domain-adaptation-for-3d-semantic-segmentation
MoPA: Multi-Modal Prior Aided Domain Adaptation for 3D Semantic Segmentation | DeepAI
Sep 21, 2023 - 09/21/23 - Multi-modal unsupervised domain adaptation (MM-UDA) for 3D semantic segmentation is a practical solution to embed semantic underst...
multi modaldomain adaptation
https://arxiv.org/abs/2506.05856v1
[2506.05856v1] Cross-View Multi-Modal Segmentation @ Ego-Exo4D Challenges 2025
Abstract page for arXiv paper 2506.05856v1: Cross-View Multi-Modal Segmentation @ Ego-Exo4D Challenges 2025
multi modalcrossview
https://easychair.org/publications/preprint/Xw9k
Enabling Fairness Across Multi-modal and Multi-agent Applications
multi modalenablingfairnessacrossagent
https://www.frontiersin.org/research-topics/69493/multi-modal-imaging-and-natural-products-nanomedicine-platform-for-advancing-imaging-guided-diagnosis-and-therapeutics/magazine
Multi-Modal Imaging and Natural Products Nanomedicine Platform for Advancing Imaging-Guided...
The evolution of medicine, from traditional herbal remedies to sophisticated artificial drugs, illustrates a dynamic journey in disease management. Recent...
multi modalnatural productsimaging
https://www.frontiersin.org/journals/neurorobotics/articles/10.3389/fnbot.2021.762252/full
Frontiers | Multi-Modal Image Fusion Based on Matrix Product State of Tensor
Multi-modal image fusion integrates different images of the same scene collected by different sensors into one image, making the fused image recognizable by ...
matrix product statemulti modalimage fusion
https://huggingface.co/papers/2307.08581
Paper page - BuboGPT: Enabling Visual Grounding in Multi-Modal LLMs
Join the discussion on this paper page
visual groundingmulti modalpaperenablingllms
https://www.osti.gov/pages/biblio/1823581-multi-modal-approach-understanding-degradation-organic-photovoltaic-materials
A Multi-modal Approach to Understanding Degradation of Organic Photovoltaic Materials (Journal...
The U.S. Department of Energy's Office of Scientific and Technical Information
multi modal
https://www.digitaljournal.com/pr/news/multi-modal-biometric-market-new-report-2023-current-industry-situation-analysis
Multi Modal Biometric Market (New Report 2023) Current industry Situation Analysis
multi modalmarket newindustry situationbiometric
https://openreview.net/forum?id=jqY3V5sM0u
MMCTAgent: Multi-modal Critical Thinking Agent Framework for Complex Visual Reasoning | OpenReview
Recent advancements in Multi-modal Large Language Models (MLLMs) have significantly improved their performance in tasks combining vision and language. However,...
multi modalcritical thinkingagent framework
https://www.snamuts.com/
Spatial Network Analysis for Multi-modal Urban Transport Systems (SNAMUTS) - Home
The official website for the Spatial Network Analysis for Multi-Modal Urban Transport Systems (SNAMUTS) Project.
spatial networkmulti modalurban transportanalysis
https://huggingface.co/MMInstruction
MMInstruction (Multi-modal Multilingual Instruction)
Org profile for Multi-modal Multilingual Instruction on Hugging Face, the AI community building the future.
multi modalmultilingualinstruction
https://www.seedance2pro.com/
Seedance 2.0 – Multi-Modal AI Video Generator with Audio
Seedance 2.0 is a next-generation multimodal AI video generator that transforms text, images, and audio into cinematic video content for professional creators.
seedance 2 0ai video generatormulti modal
https://www.msop-765k.org/
mSOP-765k: A Benchmark For Multi-Modal Structured Output Predictions | mSOP-765k
multi modalstructured outputmsopbenchmarkpredictions
https://arxiv.org/abs/2602.04405
[2602.04405] Interactive Spatial-Frequency Fusion Mamba for Multi-Modal Image Fusion
Abstract page for arXiv paper 2602.04405: Interactive Spatial-Frequency Fusion Mamba for Multi-Modal Image Fusion
spatial frequencymulti modalinteractive
https://www.abs.gov.au/statistics/detailed-methodology-information/concepts-sources-methods/australian-system-government-finance-statistics-concepts-sources-and-methods/2015/appendix-1-part-c-classification-functions-government-australia/classification-functions-government-26
Multi-modal urban transport (COFOG-A 116) | Australian Bureau of Statistics
multi modalurban transportbureau ofcofog
https://openreview.net/forum?id=UxmvCwuTMG&referrer=%5Bthe%20profile%20of%20Samir%20Awasthi%5D(%2Fprofile%3Fid%3D~Samir_Awasthi1)
ECG Representation Learning with Multi-Modal EHR Data | OpenReview
Electronic Health Records (EHRs) provide a rich source of medical information across different modalities such as electrocardiograms (ECG), structured EHRs...
representation learningmulti modalecgehrdata
https://arxiv.org/abs/2407.09209v2
[2407.09209v2] Pronunciation Assessment with Multi-modal Large Language Models
Abstract page for arXiv paper 2407.09209v2: Pronunciation Assessment with Multi-modal Large Language Models
pronunciation assessmentmulti modallarge language2407models
https://sheffield.ac.uk/insigneo/overview/events/insigneo-seminar-ai-powered-multi-modal-cardiac-data-analysis-towards-cardiac-digital-twins-pe
Insigneo Seminar: AI powered multi-modal cardiac data analysis: towards cardiac digital twins |...
We are pleased to welcome Dr Lei Li, Assistant Professor from the National University of Singapore to give a talk about cardiac digital twins for personalized...
ai poweredmulti modal
https://wandb.ai/geekyrakshit/poegan/reports/PoE-GAN-Generating-Images-from-Multi-Modal-Inputs--VmlldzoxNTA5MzUx
PoE-GAN: Generating Images from Multi-Modal Inputs
Feb 25, 2022 - PoE-GAN is a recent, fascinating paper where the authors generate images from multiple inputs like text, style, segmentation, and sketch. We dig into the...
multi modalpoegangeneratingimages
https://newatlas.com/robotics/limx-tron-1-biped-robot/
Watch: "World's first multi-modal biped robot" could soon be yours
Oct 17, 2024 - How would you like to have your own AT-ST walker from Star Wars: Return of the Jedi? Well, the just-announced Tron 1 biped robot is the next-best thing. It's...
world s firstmulti modal
https://www.zeiss.com/meditec/en/workflows/retina-workflow-idi-retina/clinical-case-library.html?f_item_overview=search_input%3Drheg
ZEISS multi-modal clinical cases library
Ever-expanding online clinical case library with multi-modal cases in retina.
multi modalclinical caseszeisslibrary
https://arxiv.org/abs/2505.03788
[2505.03788] Calibrating Uncertainty Quantification of Multi-Modal LLMs using Grounding
Abstract page for arXiv paper 2505.03788: Calibrating Uncertainty Quantification of Multi-Modal LLMs using Grounding
uncertainty quantificationmulti modalcalibrating
https://www.sintef.no/en/publications/publication/0198cc765af3-61e9ea76-0a57-4ca5-9afa-5fd51d341bbd/
Multi-Modal SLAM for Accurate Localisation in Self-similar Environments - SINTEF
multi modalself similarslamaccurate
https://elifesciences.org/articles/84122
Using multi-modal neuroimaging to characterise social brain specialisation in infants | eLife
The coupling between neural oscillatory activity, haemodynamics, and metabolism is localised to the temporo-parietal region in response to social stimuli,...
multi modal
https://wandb.ai/geekyrakshit/finance_multi_modal_rag/reports/Llama-3-2-Vision-for-multi-modal-RAG-in-financial-services--Vmlldzo5NTIyODkw
Llama 3.2-Vision for multi-modal RAG in financial services
llama 3 2multi modalvision
https://aclanthology.org/2023.acl-long.254/
Multi-modal Action Chain Abductive Reasoning - ACL Anthology
Mengze Li, Tianbao Wang, Jiahe Xu, Kairong Han, Shengyu Zhang, Zhou Zhao, Jiaxu Miao, Wenqiao Zhang, Shiliang Pu, Fei Wu. Proceedings of the 61st Annual...
multi modalabductive reasoningactionchainacl
https://www.uvm.edu/cems/trc/alternative-and-multi-modal-transportation
Alternative and Multi-Modal Transportation | Transportation Research Center | The University of...
The TRC conducts research that supports the advancement of sustainable and alternative transportation options in all communities.
multi modaltransportation researchthe universityalternativecenter