llama.cpp/common
SeungWon Jeong fb215c3832
server : normalize embeddings (#5956)
* output normalize embedding in '/v1/embeddings'

* common : reuse llama_embd_normalize

* common : better normalize impl

---------

Co-authored-by: Georgi Gerganov <ggerganov@gmail.com>
2024-03-09 14:27:58 +02:00
..
base64.hpp llava : expose as a shared library for downstream projects (#3613) 2023-11-07 00:36:23 +03:00
build-info.cpp.in build : link against build info instead of compiling against it (#3879) 2023-11-02 08:50:16 +02:00
CMakeLists.txt cmake : handle cases where git index is not found in .git (#5844) 2024-03-04 20:26:55 +02:00
common.cpp server : normalize embeddings (#5956) 2024-03-09 14:27:58 +02:00
common.h server : normalize embeddings (#5956) 2024-03-09 14:27:58 +02:00
console.cpp check C++ code with -Wmissing-declarations (#3184) 2023-09-15 15:38:27 -04:00
console.h gguf : new file format with flexible meta data (beta) (#2398) 2023-08-21 23:07:43 +03:00
grammar-parser.cpp grammar-parser : fix typo (#4318) 2023-12-04 09:57:35 +02:00
grammar-parser.h gguf : new file format with flexible meta data (beta) (#2398) 2023-08-21 23:07:43 +03:00
log.h log : fix MSVC compile errors (#5643) 2024-03-08 11:35:04 +02:00
sampling.cpp speculative : implement stochastic speculative sampling (#5625) 2024-03-04 20:24:00 +02:00
sampling.h speculative : implement stochastic speculative sampling (#5625) 2024-03-04 20:24:00 +02:00
stb_image.h examples: support LLaVA v1.5 (multimodal model) (#3436) 2023-10-12 18:23:18 +03:00
train.cpp code : normalize enum names (#5697) 2024-02-25 12:09:09 +02:00
train.h sync : ggml (backend v2) (#3912) 2023-11-13 14:16:23 +02:00