Point2RBox-v2: Rethinking Point-supervised Oriented Object Detection with Spatial Layout Among Instances Paper • 2502.04268 • Published Feb 6, 2025
Beyond Decision Boundaries: Relational Geometry Attacks on Contrastive Embedding Manifolds Paper • 2608.10237 • Published Aug 10
FIRM-Video: Check Before You Score for Reliable Text-to-Video Reward Modeling Paper • 2608.21839 • Published 27 days ago • 4
Video-IFBench: Evaluating Instruction Following of Multimodal LLMs in Video Understanding Scenarios Paper • 2608.25529 • Published 23 days ago • 17
Video-IFBench: Evaluating Instruction Following of Multimodal LLMs in Video Understanding Scenarios Paper • 2608.25529 • Published 23 days ago • 17
Video-IFBench: Evaluating Instruction Following of Multimodal LLMs in Video Understanding Scenarios Paper • 2608.25529 • Published 23 days ago • 17
FIRM-Video: Check Before You Score for Reliable Text-to-Video Reward Modeling Paper • 2608.21839 • Published 27 days ago • 4
FIRM-Video: Check Before You Score for Reliable Text-to-Video Reward Modeling Paper • 2608.21839 • Published 27 days ago • 4