Skip to main content
Product
Inference engine
Model conversion
macOS app
CLI tool
Inference for Android • Soon
Models
Metrics
Research
Docs
Careers
Company
About us
Contact us
Product
Models
Metrics
Research
Docs
Careers
Company
uzu on GitHub,
1811
stars
Talk to us
uzu on GitHub,
1811
stars
AI models library
optimized for on device.
Qwen models
from Alibaba
Qwen 3.8
27B
Qwen 3.6
27B
Qwen 3.5
9B
Qwen 3.5
4B
Qwen 3.5
2B
Qwen 3.5
0.8B
Qwen 3.8 27B
2 available checkpoints
Medium
Quantization
81
tok/s
Speed
Measured on Apple M5 Max 128GB with speculative decoding
15.5
GB
Size
Large
Quantization
58
tok/s
Speed
Measured on Apple M5 Max 128GB with speculative decoding
28.8
GB
Size
LFM models
from LiquidAI
LFM 2.5
2.6B
LFM 2.5
1.2B Thinking
LFM 2.5
1.2B Instruct
LFM 2.5
350M
LFM 2.5
230M
LFM 2.5 2.6B
2 available checkpoints
Medium
Quantization
284
tok/s
Speed
Measured on Apple M5 Max 128GB
1.5
GB
Size
Large
Quantization
172
tok/s
Speed
Measured on Apple M5 Max 128GB
2.8
GB
Size
Muse Glimmer models
from Meta
Muse-Glimmer 30B
2 available checkpoints
Medium
Quantization
85
tok/s
Speed
Measured on Apple M5 Max 128GB with speculative decoding
16.5
GB
Size
Large
Quantization
57
tok/s
Speed
Measured on Apple M5 Max 128GB with speculative decoding
30.2
GB
Size