ysn-rfd commited on
Commit
885c4a2
Β·
verified Β·
1 Parent(s): b3613c8

Upload README.md with huggingface_hub

Browse files
Files changed (1) hide show
  1. README.md +138 -0
README.md ADDED
@@ -0,0 +1,138 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ ---
2
+ base_model: bigcode/starcoder2-3b
3
+ datasets:
4
+ - bigcode/the-stack-v2-train
5
+ library_name: transformers
6
+ license: bigcode-openrail-m
7
+ pipeline_tag: text-generation
8
+ tags:
9
+ - code
10
+ - llama-cpp
11
+ - matrixportal
12
+ inference: true
13
+ widget:
14
+ - text: 'def print_hello_world():'
15
+ example_title: Hello world
16
+ group: Python
17
+ model-index:
18
+ - name: starcoder2-3b
19
+ results:
20
+ - task:
21
+ type: text-generation
22
+ dataset:
23
+ name: CruxEval-I
24
+ type: cruxeval-i
25
+ metrics:
26
+ - type: pass@1
27
+ value: 32.7
28
+ - task:
29
+ type: text-generation
30
+ dataset:
31
+ name: DS-1000
32
+ type: ds-1000
33
+ metrics:
34
+ - type: pass@1
35
+ value: 25.0
36
+ - task:
37
+ type: text-generation
38
+ dataset:
39
+ name: GSM8K (PAL)
40
+ type: gsm8k-pal
41
+ metrics:
42
+ - type: accuracy
43
+ value: 27.7
44
+ - task:
45
+ type: text-generation
46
+ dataset:
47
+ name: HumanEval+
48
+ type: humanevalplus
49
+ metrics:
50
+ - type: pass@1
51
+ value: 27.4
52
+ - task:
53
+ type: text-generation
54
+ dataset:
55
+ name: HumanEval
56
+ type: humaneval
57
+ metrics:
58
+ - type: pass@1
59
+ value: 31.7
60
+ - task:
61
+ type: text-generation
62
+ dataset:
63
+ name: RepoBench-v1.1
64
+ type: repobench-v1.1
65
+ metrics:
66
+ - type: edit-smiliarity
67
+ value: 71.19
68
+ ---
69
+
70
+ # ysn-rfd/starcoder2-3b-GGUF
71
+ This model was converted to GGUF format from [`bigcode/starcoder2-3b`](https://huggingface.co/bigcode/starcoder2-3b) using llama.cpp via the ggml.ai's [all-gguf-same-where](https://huggingface.co/spaces/matrixportal/all-gguf-same-where) space.
72
+ Refer to the [original model card](https://huggingface.co/bigcode/starcoder2-3b) for more details on the model.
73
+
74
+ ## βœ… Quantized Models Download List
75
+
76
+ ### πŸ” Recommended Quantizations
77
+ - **✨ General CPU Use:** [`Q4_K_M`](https://huggingface.co/ysn-rfd/starcoder2-3b-GGUF/resolve/main/starcoder2-3b-q4_k_m.gguf) (Best balance of speed/quality)
78
+ - **πŸ“± ARM Devices:** [`Q4_0`](https://huggingface.co/ysn-rfd/starcoder2-3b-GGUF/resolve/main/starcoder2-3b-q4_0.gguf) (Optimized for ARM CPUs)
79
+ - **πŸ† Maximum Quality:** [`Q8_0`](https://huggingface.co/ysn-rfd/starcoder2-3b-GGUF/resolve/main/starcoder2-3b-q8_0.gguf) (Near-original quality)
80
+
81
+ ### πŸ“¦ Full Quantization Options
82
+ | πŸš€ Download | πŸ”’ Type | πŸ“ Notes |
83
+ |:---------|:-----|:------|
84
+ | [Download](https://huggingface.co/ysn-rfd/starcoder2-3b-GGUF/resolve/main/starcoder2-3b-q2_k.gguf) | ![Q2_K](https://img.shields.io/badge/Q2_K-1A73E8) | Basic quantization |
85
+ | [Download](https://huggingface.co/ysn-rfd/starcoder2-3b-GGUF/resolve/main/starcoder2-3b-q3_k_s.gguf) | ![Q3_K_S](https://img.shields.io/badge/Q3_K_S-34A853) | Small size |
86
+ | [Download](https://huggingface.co/ysn-rfd/starcoder2-3b-GGUF/resolve/main/starcoder2-3b-q3_k_m.gguf) | ![Q3_K_M](https://img.shields.io/badge/Q3_K_M-FBBC05) | Balanced quality |
87
+ | [Download](https://huggingface.co/ysn-rfd/starcoder2-3b-GGUF/resolve/main/starcoder2-3b-q3_k_l.gguf) | ![Q3_K_L](https://img.shields.io/badge/Q3_K_L-4285F4) | Better quality |
88
+ | [Download](https://huggingface.co/ysn-rfd/starcoder2-3b-GGUF/resolve/main/starcoder2-3b-q4_0.gguf) | ![Q4_0](https://img.shields.io/badge/Q4_0-EA4335) | Fast on ARM |
89
+ | [Download](https://huggingface.co/ysn-rfd/starcoder2-3b-GGUF/resolve/main/starcoder2-3b-q4_k_s.gguf) | ![Q4_K_S](https://img.shields.io/badge/Q4_K_S-673AB7) | Fast, recommended |
90
+ | [Download](https://huggingface.co/ysn-rfd/starcoder2-3b-GGUF/resolve/main/starcoder2-3b-q4_k_m.gguf) | ![Q4_K_M](https://img.shields.io/badge/Q4_K_M-673AB7) ⭐ | Best balance |
91
+ | [Download](https://huggingface.co/ysn-rfd/starcoder2-3b-GGUF/resolve/main/starcoder2-3b-q5_0.gguf) | ![Q5_0](https://img.shields.io/badge/Q5_0-FF6D01) | Good quality |
92
+ | [Download](https://huggingface.co/ysn-rfd/starcoder2-3b-GGUF/resolve/main/starcoder2-3b-q5_k_s.gguf) | ![Q5_K_S](https://img.shields.io/badge/Q5_K_S-0F9D58) | Balanced |
93
+ | [Download](https://huggingface.co/ysn-rfd/starcoder2-3b-GGUF/resolve/main/starcoder2-3b-q5_k_m.gguf) | ![Q5_K_M](https://img.shields.io/badge/Q5_K_M-0F9D58) | High quality |
94
+ | [Download](https://huggingface.co/ysn-rfd/starcoder2-3b-GGUF/resolve/main/starcoder2-3b-q6_k.gguf) | ![Q6_K](https://img.shields.io/badge/Q6_K-4285F4) πŸ† | Very good quality |
95
+ | [Download](https://huggingface.co/ysn-rfd/starcoder2-3b-GGUF/resolve/main/starcoder2-3b-q8_0.gguf) | ![Q8_0](https://img.shields.io/badge/Q8_0-EA4335) ⚑ | Fast, best quality |
96
+ | [Download](https://huggingface.co/ysn-rfd/starcoder2-3b-GGUF/resolve/main/starcoder2-3b-f16.gguf) | ![F16](https://img.shields.io/badge/F16-000000) | Maximum accuracy |
97
+
98
+ πŸ’‘ **Tip:** Use `F16` for maximum precision when quality is critical
99
+
100
+
101
+ ---
102
+ # πŸš€ Applications and Tools for Locally Quantized LLMs
103
+ ## πŸ–₯️ Desktop Applications
104
+
105
+ | Application | Description | Download Link |
106
+ |-----------------|----------------------------------------------------------------------------------------------|--------------------------------------------------------------------------------|
107
+ | **Llama.cpp** | A fast and efficient inference engine for GGUF models. | [GitHub Repository](https://github.com/ggml-org/llama.cpp) |
108
+ | **Ollama** | A streamlined solution for running LLMs locally. | [Website](https://ollama.com/) |
109
+ | **AnythingLLM** | An AI-powered knowledge management tool. | [GitHub Repository](https://github.com/Mintplex-Labs/anything-llm) |
110
+ | **Open WebUI** | A user-friendly web interface for running local LLMs. | [GitHub Repository](https://github.com/open-webui/open-webui) |
111
+ | **GPT4All** | A user-friendly desktop application supporting various LLMs, compatible with GGUF models. | [GitHub Repository](https://github.com/nomic-ai/gpt4all) |
112
+ | **LM Studio** | A desktop application designed to run and manage local LLMs, supporting GGUF format. | [Website](https://lmstudio.ai/) |
113
+ | **GPT4All Chat**| A chat application compatible with GGUF models for local, offline interactions. | [GitHub Repository](https://github.com/nomic-ai/gpt4all) |
114
+
115
+ ---
116
+
117
+ ## πŸ“± Mobile Applications
118
+
119
+ | Application | Description | Download Link |
120
+ |-------------------|----------------------------------------------------------------------------------------------|--------------------------------------------------------------------------------|
121
+ | **ChatterUI** | A simple and lightweight LLM app for mobile devices. | [GitHub Repository](https://github.com/Vali-98/ChatterUI) |
122
+ | **Maid** | Mobile Artificial Intelligence Distribution for running AI models on mobile devices. | [GitHub Repository](https://github.com/Mobile-Artificial-Intelligence/maid) |
123
+ | **PocketPal AI** | A mobile AI assistant powered by local models. | [GitHub Repository](https://github.com/a-ghorbani/pocketpal-ai) |
124
+ | **Layla** | A flexible platform for running various AI models on mobile devices. | [Website](https://www.layla-network.ai/) |
125
+
126
+ ---
127
+
128
+ ## 🎨 Image Generation Applications
129
+
130
+ | Application | Description | Download Link |
131
+ |-------------------------------------|----------------------------------------------------------------------------------------------|--------------------------------------------------------------------------------|
132
+ | **Stable Diffusion** | An open-source AI model for generating images from text. | [GitHub Repository](https://github.com/CompVis/stable-diffusion) |
133
+ | **Stable Diffusion WebUI** | A web application providing access to Stable Diffusion models via a browser interface. | [GitHub Repository](https://github.com/AUTOMATIC1111/stable-diffusion-webui) |
134
+ | **Local Dream** | Android Stable Diffusion with Snapdragon NPU acceleration. Also supports CPU inference. | [GitHub Repository](https://github.com/xororz/local-dream) |
135
+ | **Stable-Diffusion-Android (SDAI)** | An open-source AI art application for Android devices, enabling digital art creation. | [GitHub Repository](https://github.com/ShiftHackZ/Stable-Diffusion-Android) |
136
+
137
+ ---
138
+