leonsarmiento commited on
Commit
904eac5
·
verified ·
1 Parent(s): 6762c94

Upload README.md with huggingface_hub

Browse files
Files changed (1) hide show
  1. README.md +4 -0
README.md CHANGED
@@ -68,6 +68,10 @@ BaseQuant_XL recipe — precision is allocated by layer importance, not applied
68
 
69
  > `preserve_thinking` is supported — set `chat_template_kwargs: {"preserve_thinking": true}` to retain thinking traces from historical messages (beneficial for agentic scenarios).
70
 
 
 
 
 
71
  ## Model Overview
72
 
73
  | Property | Value |
 
68
 
69
  > `preserve_thinking` is supported — set `chat_template_kwargs: {"preserve_thinking": true}` to retain thinking traces from historical messages (beneficial for agentic scenarios).
70
 
71
+ ## Chat Template
72
+
73
+ The `chat_template.jinja` and `tokenizer_config.json` were updated on **2026-07-30** to match the latest chat template published by Kwaipilot in [Kwaipilot/KAT-Coder-V2.5-Dev](https://huggingface.co/Kwaipilot/KAT-Coder-V2.5-Dev). This update enforces that system messages must appear at the beginning of the conversation — mid-conversation system messages now raise an exception instead of being silently rendered.
74
+
75
  ## Model Overview
76
 
77
  | Property | Value |