BDH-CQ: In-Context Learning with Recurrent Latent Reasoning Paper • 2608.09888 • Published 13 days ago • 708
Advancing Polish Language Modeling through Tokenizer Optimization in the Bielik v3 7B and 11B Series Paper • 2604.10799 • Published Apr 12 • 6
Bielik-Minitron-7B: Compressing Large Language Models via Structured Pruning and Knowledge Distillation for the Polish Language Paper • 2603.11881 • Published Mar 12 • 3
Bielik 11B v3: Multilingual Large Language Model for European Languages Paper • 2601.11579 • Published Dec 30, 2025 • 4
BDH-CQ: In-Context Learning with Recurrent Latent Reasoning Paper • 2608.09888 • Published 13 days ago • 708
Advancing Polish Language Modeling through Tokenizer Optimization in the Bielik v3 7B and 11B Series Paper • 2604.10799 • Published Apr 12 • 6
speakleash/Bielik-Minitron-7B-v3.0-Instruct-FP8-Dynamic Text Generation • 7B • Updated Mar 30 • 1.48k • 2
speakleash/Bielik-Minitron-7B-v3.0-Instruct-FP8-Dynamic Text Generation • 7B • Updated Mar 30 • 1.48k • 2
The Dragon Hatchling: The Missing Link between the Transformer and Models of the Brain Paper • 2509.26507 • Published Sep 30, 2025 • 553