llama-cpp-turboquant

History

SeungWon Jeong fb215c3832 server : normalize embeddings (#5956 ) * output normalize embedding in '/v1/embeddings' * common : reuse llama_embd_normalize * common : better normalize impl --------- Co-authored-by: Georgi Gerganov <ggerganov@gmail.com>		2024-03-09 14:27:58 +02:00
..
base64.hpp	llava : expose as a shared library for downstream projects (#3613 )	2023-11-07 00:36:23 +03:00
build-info.cpp.in	build : link against build info instead of compiling against it (#3879 )	2023-11-02 08:50:16 +02:00
CMakeLists.txt	cmake : handle cases where git index is not found in .git (#5844 )	2024-03-04 20:26:55 +02:00
common.cpp	server : normalize embeddings (#5956 )	2024-03-09 14:27:58 +02:00
common.h	server : normalize embeddings (#5956 )	2024-03-09 14:27:58 +02:00
console.cpp	check C++ code with -Wmissing-declarations (#3184 )	2023-09-15 15:38:27 -04:00
console.h	gguf : new file format with flexible meta data (beta) (#2398 )	2023-08-21 23:07:43 +03:00
grammar-parser.cpp	grammar-parser : fix typo (#4318 )	2023-12-04 09:57:35 +02:00
grammar-parser.h	gguf : new file format with flexible meta data (beta) (#2398 )	2023-08-21 23:07:43 +03:00
log.h	log : fix MSVC compile errors (#5643 )	2024-03-08 11:35:04 +02:00
sampling.cpp	speculative : implement stochastic speculative sampling (#5625 )	2024-03-04 20:24:00 +02:00
sampling.h	speculative : implement stochastic speculative sampling (#5625 )	2024-03-04 20:24:00 +02:00
stb_image.h	examples: support LLaVA v1.5 (multimodal model) (#3436 )	2023-10-12 18:23:18 +03:00
train.cpp	code : normalize enum names (#5697 )	2024-02-25 12:09:09 +02:00
train.h	sync : ggml (backend v2) (#3912 )	2023-11-13 14:16:23 +02:00