z0n3x commited on
Commit
f2b2cb2
·
verified ·
1 Parent(s): a0fdfd4

model card: document default <think> reasoning + LM Studio reasoning tags + text-only note

Browse files
Files changed (1) hide show
  1. README.md +12 -0
README.md CHANGED
@@ -104,6 +104,18 @@ prompt = tokenizer.apply_chat_template(messages, add_generation_prompt=True)
104
  print(generate(model, tokenizer, prompt=prompt, max_tokens=800))
105
  ```
106
 
 
 
 
 
 
 
 
 
 
 
 
 
107
  > **Note on loading:** these are brand-new Gemma 4 "unified" weights. With some `mlx-lm`
108
  > versions you may need a small load-time shim that (a) resolves `model_type: gemma4` and
109
  > (b) drops unused multimodal tensors. This is a **text-only** checkpoint.
 
104
  print(generate(model, tokenizer, prompt=prompt, max_tokens=800))
105
  ```
106
 
107
+ ### Thinking / reasoning
108
+
109
+ The chat template **defaults to a reasoning system prompt**, so the model produces
110
+ `<think> … </think>` then the answer **out of the box** — you don't need to pass a system
111
+ message (pass your own to override it). The reasoning markers are `<think>` / `</think>`
112
+ (this model was trained on those tags, not Gemma's native `<|channel>thought` format).
113
+
114
+ - **LM Studio:** set the reasoning / "thinking" section tags to `<think>` (start) and
115
+ `</think>` (end) to fold the chain-of-thought into a collapsible block.
116
+ - **Text-only:** the base Gemma 4 vision/audio weights were dropped during fine-tuning, so
117
+ this checkpoint is **not multimodal** — by design (Fin-R1 is a text recipe).
118
+
119
  > **Note on loading:** these are brand-new Gemma 4 "unified" weights. With some `mlx-lm`
120
  > versions you may need a small load-time shim that (a) resolves `model_type: gemma4` and
121
  > (b) drops unused multimodal tensors. This is a **text-only** checkpoint.