Lv1
24 积分 2024-12-23 加入
CEPrompt: Cross-Modal Emotion-Aware Prompting for Facial Expression Recognition
1个月前
已完结
MARS: Paying More Attention to Visual Attributes for Text-Based Person Search
1个月前
已完结
Prototype-guided text-based person search on rich Chinese descriptions
2个月前
已完结
ROD-MLLM: Towards More Reliable Object Detection in Multimodal Large Language Models
2个月前
已完结
LMM-Det: Make Large Multimodal Models Excel in Object Detection
2个月前
已完结
Multimodal Foundation Models: From Specialists to General-Purpose Assistants
3个月前
已完结
From Structure to Synergy: A Survey of Vision-Language Perception Paradigm Evolution in Multimodal Large Language Models
3个月前
已关闭
Bottom-up color-independent alignment learning for text–image person re-identification
6个月前
已完结
Attribute-Centric Cross-Modal Alignment for Weakly Supervised Text-Based Person Re-ID
6个月前
已完结
UCPM: Uncertainty-Guided Cross-Modal Retrieval with Partially Mismatched Pairs
8个月前
已完结