llama.cpp

Author	SHA1	Message	Date
ochafik	c0f972bb45	Use --reasoning-format, remove forced thinking for now	2025-02-08 17:58:33 +00:00
Olivier Chafik	e6d9b52480	align Command R7B w/ --think / reasoning_content behaviour	2025-02-05 15:47:37 +00:00
ochafik	9d7c3cc51b	--think to force any model to return reasoning_content (or just parse <think> for deepseek r1)	2025-02-05 12:16:37 +00:00
Olivier Chafik	c6214ee9d6	rm unneeded vocab	2025-02-03 19:59:50 +00:00
ochafik	87de852b7f	pass vocab to common_chat_params_init	2025-02-03 02:24:30 +00:00
Olivier Chafik	bfcce4d693	`tool-call`: support Command R7B (+ return tool_plan "thoughts" in API) (#11585 ) * `tool-call`: support Command R7B (w/ tool_plan return) * `tool-call`: cleaner preservation of tokens + warn when likely bad chat template override * `tool-call`: test cleanup / handle lazy grammar triggers	2025-02-02 09:25:38 +00:00
Olivier Chafik	8b576b6c55	Tool call support (generic + native for Llama, Functionary, Hermes, Mistral, Firefunction, DeepSeek) w/ lazy grammars (#9639 ) --------- Co-authored-by: Xuan Son Nguyen <thichthat@gmail.com> Co-authored-by: Georgi Gerganov <ggerganov@gmail.com> Co-authored-by: Xuan Son Nguyen <son@huggingface.co>	2025-01-30 19:13:58 +00:00