FlashRT: Agent Harness for Guiding Agents to Deploy Real-Time Multimodal Applications Paper • 2607.18171 • Published 1 day ago • 3
SWE-Pruner Pro: The Coder LLM Already Knows What to Prune Paper • 2607.18213 • Published 1 day ago • 64
Audio-Visual Flamingo: Open Audio-Visual Intelligence for Long and Complex Videos Paper • 2607.16107 • Published 5 days ago • 7
SearchOS-V1: Towards Robust Open-Domain Information-Seeking Agent Collaboration Paper • 2607.15257 • Published 6 days ago • 66
MeanFlowNFT: Bringing Forward-Process RL to Average-Velocity Generators Paper • 2607.15273 • Published 6 days ago • 15
Harness Handbook: Making Evolving Agent Harnesses Readable,Navigable, and Editable Paper • 2607.13285 • Published 8 days ago • 207
Motion4Motion: Motion Transfer Across Subjects at Inference Paper • 2607.11644 • Published 9 days ago • 13
LATO.2: Factorized 3D Mesh Generation with Vertex and Topology Flow Paper • 2607.10623 • Published 10 days ago • 14
Video Generation Models are General-Purpose Vision Learners Paper • 2607.09024 • Published 12 days ago • 82
A Sovereign, Open-Source Foundation Model for German and English Paper • 2607.09424 • Published 12 days ago • 11
Scalable Visual Pretraining for Language Intelligence Paper • 2607.09657 • Published 12 days ago • 56
ARDY: Autoregressive Diffusion with Hybrid Representation for Interactive Human Motion Generation Paper • 2607.08741 • Published 13 days ago • 11