llama_cpp_for_radxa_dragon_wing_q6a

History

Radoslav Gerganov 2b6b55a59f server : include usage statistics only when user request them (#16052 ) * server : include usage statistics only when user request them When serving the OpenAI compatible API, we should check if {"stream_options": {"include_usage": true} is set in the request when deciding whether we should send usage statistics closes: #16048 * add unit test		2025-09-18 10:36:57 +00:00
..
batched-bench
cvector-generator
export-lora
gguf-split
imatrix
llama-bench
main
mtmd
perplexity
quantize
rpc
run
server
tokenize
tts
CMakeLists.txt