One model for both halves of RAG retrieval; a strong default per size. Contact baa.ai for the optimal pick for your corpus.
AI & ML interests
Model Quantization
Recent Activity
View all activity
Organization Card
Smaller. Smarter. Sovereign.
Making frontier models run anywhere
We publish high-quality quantized models for Apple Silicon and GGUF. Our models use a proprietary optimisation method that delivers superior quality at your target memory budget.
Browse our models, or connect with us below.
π baa.aiβΒ·β π¬ Discord
models 79
baa-ai/Qwen3.8-27B-RAM-24GB-MLX
Image-Text-to-Text β’ 8B β’ Updated
baa-ai/paddock-reader-35b-gguf
35B β’ Updated β’ 116
baa-ai/paddock-reader-9b-gguf
9B β’ Updated β’ 128
baa-ai/GLM-5.2-RAM-333GB-MLX
96B β’ Updated β’ 1.77k
baa-ai/Merino-XL-v2
Sentence Similarity β’ Updated β’ 5
baa-ai/Merino-XL
Sentence Similarity β’ Updated β’ 6
baa-ai/Merino-Pro-4bit
Sentence Similarity β’ Updated β’ 7
baa-ai/Merino-Pro
Sentence Similarity β’ Updated β’ 29
baa-ai/Merino-Large-v2
Sentence Similarity β’ Updated β’ 6
baa-ai/Merino-Large
Sentence Similarity β’ Updated β’ 7
datasets 0
None public yet