-
ViQ: Text-Aligned Visual Quantized Representations at Any Resolution
Paper • 2606.27313 • Published • 38 -
Qwen3-VL-Embedding and Qwen3-VL-Reranker: A Unified Framework for State-of-the-Art Multimodal Retrieval and Ranking
Paper • 2601.04720 • Published • 59 -
Perception Encoder: The best visual embeddings are not at the output of the network
Paper • 2504.13181 • Published • 37
PEILIN XIONG
PANDATREE
·
AI & ML interests
None yet
Recent Activity
commentedon a paper 14 days ago
BRIDGE: Background Routing and Isolated Discrete Gating for Coarse-Mask Local Editing updated a collection about 1 month ago
Image feature