lmg-anon/VNTL-v3.1-1k
Viewer β’ Updated β’ 14.5k β’ 22 β’ 2
How to use lmg-anon/vntl-gemma2-2b-gguf with PEFT:
Task type is invalid.
How to use lmg-anon/vntl-gemma2-2b-gguf with llama.cpp:
curl -LsSf https://llama.app/install.sh | sh # Start a local OpenAI-compatible server with a web UI: llama serve -hf lmg-anon/vntl-gemma2-2b-gguf:Q6_K # Run inference directly in the terminal: llama cli -hf lmg-anon/vntl-gemma2-2b-gguf:Q6_K
winget install llama.cpp # Start a local OpenAI-compatible server with a web UI: llama serve -hf lmg-anon/vntl-gemma2-2b-gguf:Q6_K # Run inference directly in the terminal: llama cli -hf lmg-anon/vntl-gemma2-2b-gguf:Q6_K
# Download pre-built binary from: # https://github.com/ggerganov/llama.cpp/releases # Start a local OpenAI-compatible server with a web UI: ./llama-server -hf lmg-anon/vntl-gemma2-2b-gguf:Q6_K # Run inference directly in the terminal: ./llama-cli -hf lmg-anon/vntl-gemma2-2b-gguf:Q6_K
git clone https://github.com/ggerganov/llama.cpp.git cd llama.cpp cmake -B build cmake --build build -j --target llama-server llama-cli # Start a local OpenAI-compatible server with a web UI: ./build/bin/llama-server -hf lmg-anon/vntl-gemma2-2b-gguf:Q6_K # Run inference directly in the terminal: ./build/bin/llama-cli -hf lmg-anon/vntl-gemma2-2b-gguf:Q6_K
docker model run hf.co/lmg-anon/vntl-gemma2-2b-gguf:Q6_K
How to use lmg-anon/vntl-gemma2-2b-gguf with Ollama:
ollama run hf.co/lmg-anon/vntl-gemma2-2b-gguf:Q6_K
How to use lmg-anon/vntl-gemma2-2b-gguf with Docker Model Runner:
docker model run hf.co/lmg-anon/vntl-gemma2-2b-gguf:Q6_K
How to use lmg-anon/vntl-gemma2-2b-gguf with Lemonade:
# Download Lemonade from https://lemonade-server.ai/ lemonade pull lmg-anon/vntl-gemma2-2b-gguf:Q6_K
lemonade run user.vntl-gemma2-2b-gguf-Q6_K
lemonade list
his repository contains some GGUF quantizations of the merged VNTL Gemma2 2B lora, created using the VNTL 3.1 dataset.
The purpose of this model is to improve Gemma2's performance at translating Japanese visual novels to English.
For more information about this model, please check out the original repository.
This is an prompt example for translation:
<<METADATA>>
[character] Name: Uryuu Shingo (ηη ζ°εΎ) | Gender: Male | Aliases: Onii-chan (γε
γ‘γγ)
[character] Name: Uryuu Sakuno (ηη ζ‘δΉ) | Gender: Female
<<TRANSLATE>>
<<JAPANESE>>
[ζ‘δΉ]: γβ¦β¦γγγγ
<<ENGLISH>>
[Sakuno]: γ... Sorry.γ<eos>
<<JAPANESE>>
[ζ°εΎ]: γγγγγγγθ¨γ£γ‘γγͺγγ γγ©γθΏ·εγ§γγγ£γγγζ‘δΉγ―ε―ζγγγγγγγγεΏι
γγ‘γγ£γ¦γγγ γδΏΊγ
<<ENGLISH>>
The generated translation for that prompt, with temperature 0, is:
[Shingo]: γNo, I'm glad you got lost. You were so cute that it made me worry.γ
6-bit
8-bit