Back to papers
March 19, 2026cs.CVcs.LGAdvanced
PromptHub: Enhancing Multi-Prompt Visual In-Context Learning with Locality-Aware Fusion, Concentration and Alignment
AI-Generated Summary
PromptHub is a framework that improves visual in-context learning (learning vision tasks from image examples) by better combining multiple demonstration examples through spatial awareness and smarter training objectives. Unlike previous methods that treat image patches independently, PromptHub uses location information to understand context better and employs multiple complementary training goals to help the model learn more effectively. The approach shows strong performance across multiple vision tasks and works reliably even with different types of images.
Difficulty
Advanced
Categories
cs.CV, cs.LG
AI Tags
Computer VisionIn-Context LearningFew-Shot LearningMulti-Modal LearningPrompt EngineeringVision Tasks