Any-to-Any
Transformers
Safetensors
multilingual
minicpmo
feature-extraction
minicpm-o
omni
vision
ocr
multi-image
video
custom_code
audio
speech
voice cloning
live Streaming
realtime speech conversation
asr
tts
4-bit precision
gptq
Instructions to use openbmb/MiniCPM-o-2_6-int4 with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use openbmb/MiniCPM-o-2_6-int4 with Transformers:
# Load model directly from transformers import AutoModel model = AutoModel.from_pretrained("openbmb/MiniCPM-o-2_6-int4", trust_remote_code=True, device_map="auto") - Notebooks
- Google Colab
- Kaggle
Update README.md
Browse files
README.md
CHANGED
|
@@ -33,7 +33,7 @@ Running with int4 version would use lower GPU memory (about 9GB).
|
|
| 33 |
We are submitting PR to officially support minicpm-o 2.6 inference
|
| 34 |
|
| 35 |
```python
|
| 36 |
-
git clone https://github.com/
|
| 37 |
git checkout minicpmo
|
| 38 |
|
| 39 |
# install AutoGPTQ
|
|
|
|
| 33 |
We are submitting PR to officially support minicpm-o 2.6 inference
|
| 34 |
|
| 35 |
```python
|
| 36 |
+
git clone https://github.com/OpenBMB/AutoGPTQ.git && cd AutoGPTQ
|
| 37 |
git checkout minicpmo
|
| 38 |
|
| 39 |
# install AutoGPTQ
|