mirror of https://github.com/ggerganov/llama.cpp.git synced 2025-01-23 18:09:18 +01:00

History

Xuan Son Nguyen 3e58b0ee35 cvector: fix CI + correct help message (#8064 ) * cvector: fix CI + correct help message * also correct --pca-iter		2024-06-22 18:11:30 +02:00
..
CMakeLists.txt	Add `cvector-generator` example (#7514 )	2024-06-15 18:53:40 +02:00
completions.txt	Add `cvector-generator` example (#7514 )	2024-06-15 18:53:40 +02:00
cvector-generator.cpp	cvector: fix CI + correct help message (#8064 )	2024-06-22 18:11:30 +02:00
negative.txt	cvector-generator: Moe Moe Fixie-Fixie for Lots of Formats~! ♡(ᐢ ᴥ ᐢ)♡ (#8052 )	2024-06-22 17:19:37 +02:00
pca.hpp	Add support for sqrt on CUDA (#7953 )	2024-06-17 00:23:04 +02:00
positive.txt	cvector-generator: Moe Moe Fixie-Fixie for Lots of Formats~! ♡(ᐢ ᴥ ᐢ)♡ (#8052 )	2024-06-22 17:19:37 +02:00
README.md	cvector: fix CI + correct help message (#8064 )	2024-06-22 18:11:30 +02:00

README.md

cvector-generator

This example demonstrates how to generate a control vector using gguf models.

Related PRs:

Examples

# CPU only
./cvector-generator -m ./dolphin-2.0-mistral-7b.Q4_K_M.gguf

# With GPU
./cvector-generator -m ./dolphin-2.0-mistral-7b.Q4_K_M.gguf -ngl 99

# With advanced options
./cvector-generator -m ./dolphin-2.0-mistral-7b.Q4_K_M.gguf -ngl 99 --completions 128 --pca-iter 2000 --pca-batch 100

# To see help message
./cvector-generator -h
# Then, have a look at "cvector" section

Tips and tricks

If you have multiple lines per prompt, you can escape the newline character (change it to \n). For example:

<|im_start|>system\nAct like a person who is extremely happy.<|im_end|>
<|im_start|>system\nYou are in a very good mood today<|im_end|>