Back to search

Article

RobustVLM: Enhancing Large Vision-Language Models through Adaptive Visual Context Learning, Weak-to-Strong Robust Transfer, and Domain-Adaptive Feedback

2026-05-19

Abstract excerpt

<title>Abstract</title> <p>Large Vision-Language Models have demonstrated remarkable capabilities in multimodal understanding, yet they face critical challenges in visual in-context learning, weak-to-strong generalization, and domain-specific reasoning. We propose RobustVLM, a unified framework that addresses these challenges through three complementary modules. First, Adaptive Visual Context Learning enables eff...

Topics

Open a Topic to create a Post that cites this publication.

Identifiers and source

Literature Corpus work
dae977eb-6bb4-5d46-a5b2-dfbb4d48ef60
DOI
10.21203/rs.3.rs-9748489/v1
Open publication

Related research

Semantic proximity does not establish scientific evidence.

Click a neighbor to travelStep 1 · 12 closest
Interactive article relationship graphSelect a related publication card to move it into the centre and load its closest explainable connections. Solid lines are source-backed structured connections. Dashed lines are semantic discovery signals and are not scientific evidence.
RobustVLM: Enhancing Large Vision-Language Models through Adaptive Visual Context Learning, Weak-to-Strong Robust Transfer, and Domain-Adaptive FeedbackDOI 10.21203/rs.3.rs-9748489/v1
Select a neighboring publication to make it the new centre.