DeepGlint-AI/UniME-R1-4B
Image-Text-to-Text • Updated
Artificial General Intelligence
Learning from Failures: Retrieval-Centric CoT via Hard Negatives for Unified Multimodal Retrieval
UniDoc-RL: Coarse-to-Fine Visual RAG with Hierarchical Actions and Dense Rewards