-
The Art of Scaling Reinforcement Learning Compute for LLMs
Paper • 2510.13786 • Published • 34 -
Attention Is All You Need for KV Cache in Diffusion LLMs
Paper • 2510.14973 • Published • 43 -
BitNet Distillation
Paper • 2510.13998 • Published • 63 -
GigaBrain-0: A World Model-Powered Vision-Language-Action Model
Paper • 2510.19430 • Published • 54
🔄 In a Training Loop
Keylhan Paumard--André
keypa
AI & ML interests
Efficient deep learning, LLM fine-tuning, inference optimization, model compression, distributed training, GPU systems, open-source AI infrastructure
Recent Activity
upvoted an article about 12 hours ago
Kimi K3 Model Overview: 2.8T Parameters, MXFP4 Quantization, and What the Open Weights Mean for the Community liked a dataset about 12 hours ago
greghavens/kimi-k3-coding-and-debugging-traces liked a model 1 day ago
baseten/glm-52-debug