One model for both halves of RAG retrieval; a strong default per size. Contact baa.ai for the optimal pick for your corpus.
AI & ML interests
Model Quantization
Recent Activity
View all activity
Organization Card
Smaller. Smarter. Sovereign.
Making frontier models run anywhere
We publish high-quality quantized models for Apple Silicon and GGUF. Our models use a proprietary optimisation method that delivers superior quality at your target memory budget.
Browse our models, or connect with us below.
π baa.aiβΒ·β π¬ Discord
models 78
baa-ai/paddock-reader-35b-gguf
35B β’ Updated β’ 90
baa-ai/paddock-reader-9b-gguf
9B β’ Updated β’ 92
baa-ai/GLM-5.2-RAM-333GB-MLX
96B β’ Updated β’ 594
baa-ai/Merino-XL-v2
Sentence Similarity β’ Updated β’ 41
baa-ai/Merino-XL
Sentence Similarity β’ Updated β’ 45
baa-ai/Merino-Pro-4bit
Sentence Similarity β’ Updated β’ 71
baa-ai/Merino-Pro
Sentence Similarity β’ Updated β’ 88
baa-ai/Merino-Large-v2
Sentence Similarity β’ Updated β’ 42
baa-ai/Merino-Large
Sentence Similarity β’ Updated β’ 41
baa-ai/Merino-Small
Sentence Similarity β’ Updated β’ 56 β’ 1
datasets 0
None public yet