danielhanchen commited on
Commit
66205ea
·
verified ·
1 Parent(s): bcf2096

Update README.md

Browse files
Files changed (1) hide show
  1. README.md +4 -3
README.md CHANGED
@@ -12,12 +12,13 @@ pipeline_tag: text-generation
12
  ---
13
  # Read our How to [Run GLM-4.7-Flash Guide!](https://unsloth.ai/docs/models/glm-4.7-flash)
14
 
15
- # Jan 21 update: llama.cpp fixed a bug that caused looping and poor outputs. We updated the GGUFs - please re-download the model for much better outputs.
 
 
16
  You can now use Z.ai's recommended parameters and get great results:
17
  - For general use-case: `--temp 1.0 --top-p 0.95`
18
  - For tool-calling: `--temp 0.7 --top-p 1.0`
19
- - If using llama.cpp, set `--min-p 0.01` as llama.cpp's default is 0.1
20
- - **Repeat penalty: Disable it, or set `--repeat-penalty 1.0`**
21
 
22
  You can also fine-tune GLM-4.7-Flash with Unsloth via our [GLM free notebook](https://unsloth.ai/docs/models/glm-4.7-flash#fine-tuning-glm-4.7-flash).
23
 
 
12
  ---
13
  # Read our How to [Run GLM-4.7-Flash Guide!](https://unsloth.ai/docs/models/glm-4.7-flash)
14
 
15
+ ## Jan 21 update: llama.cpp fixed a bug that caused looping and poor outputs. We updated the GGUFs - please re-download the model for much better outputs.
16
+ - **Repeat penalty: Disable it, or set `--repeat-penalty 1.0`**
17
+
18
  You can now use Z.ai's recommended parameters and get great results:
19
  - For general use-case: `--temp 1.0 --top-p 0.95`
20
  - For tool-calling: `--temp 0.7 --top-p 1.0`
21
+ - If using llama.cpp, set `--min-p 0.01` as llama.cpp's default is 0.05
 
22
 
23
  You can also fine-tune GLM-4.7-Flash with Unsloth via our [GLM free notebook](https://unsloth.ai/docs/models/glm-4.7-flash#fine-tuning-glm-4.7-flash).
24
 
Free AI Image Generator No sign-up. Instant results. Open Now