Blind Men and the Elephant: Probing the Epistemic Myopia of LLMs under Long-Tail Divergent Knowledge Paper • 2608.28478 • Published 4 days ago • 15
ContextPilot: Teaching Agents for Proactive Context Management via Fine-grained RL Paper • 2608.28476 • Published 4 days ago • 24
CausalMix: Data Mixture as Causal Inference for Language Model Training Paper • 2607.01104 • Published Jul 1 • 21
CausalMix: Data Mixture as Causal Inference for Language Model Training Paper • 2607.01104 • Published Jul 1 • 21
CausalMix: Data Mixture as Causal Inference for Language Model Training Paper • 2607.01104 • Published Jul 1 • 21
BioMatrix: Towards a Comprehensive Biological Foundation Model Spanning the Modality Matrix of Sequences, Structures, and Language Paper • 2606.22138 • Published Jun 20 • 25
Distilling Long-CoT Reasoning through Collaborative Step-wise Multi-Teacher Decoding Paper • 2605.02290 • Published May 4 • 43
AstraFlow: Dataflow-Oriented Reinforcement Learning for Agentic LLMs Paper • 2605.15565 • Published May 15 • 18
The MiniMax-M2 Series: Mini Activations Unleashing Max Real-World Intelligence Paper • 2605.26494 • Published May 26 • 41
Nemotron 3 Ultra: Open, Efficient Mixture-of-Experts Hybrid Mamba-Transformer Model for Agentic Reasoning Paper • 2606.15007 • Published Jun 12 • 19
VibeThinker-3B: Exploring the Frontier of Verifiable Reasoning in Small Language Models Paper • 2606.16140 • Published Jun 15 • 126
Ling and Ring 2.6 Technical Report: Efficient and Instant Agentic Intelligence at Trillion-Parameter Scale Paper • 2606.15079 • Published Jun 13 • 87
Domain-Specific Data Synthesis for LLMs via Minimal Sufficient Representation Learning Paper • 2605.30039 • Published May 29 • 20 • 2
Domain-Specific Data Synthesis for LLMs via Minimal Sufficient Representation Learning Paper • 2605.30039 • Published May 29 • 20
MIRA: Mid-training Rubric Anchoring for Source-Aware Data Selection Paper • 2605.30288 • Published May 29 • 23
STRIDE: Training Data Attribution via Sparse Recovery from Subset Perturbations Paper • 2606.05165 • Published Jun 3 • 4
LLM Explainability with Counterfactual Chains and Causal Graphs Paper • 2606.05972 • Published Jun 4 • 18