ByteDance-Seed
/

Seed-OSS-36B-Instruct

Text Generation

Model card Files Files and versions Community

yo37 commited on 3 days ago

Commit

0193a51

·

verified ·

1 Parent(s): 5f4e324

Update README.md

Files changed (1) hide show

README.md +1 -1

README.md CHANGED Viewed

@@ -474,7 +474,7 @@ Incorporating synthetic instruction data into pretraining leads to improved perf
 Users can flexibly specify the model's thinking budget. The figure below shows the performance curves across different tasks as the thinking budget varies. For simpler tasks (such as IFEval), the model's chain of thought (CoT) is shorter, and the score exhibits fluctuations as the thinking budget increases. For more challenging tasks (such as AIME and LiveCodeBench), the model's CoT is longer, and the score improves with an increase in the thinking budget.
-![thinking_budget](./figures/thinking_budget.png)
 Here is an example with a thinking budget set to 512: during the reasoning process, the model periodically triggers self-reflection to estimate the consumed and remaining budget, and delivers the final response once the budget is exhausted or the reasoning concludes.
 ```

 Users can flexibly specify the model's thinking budget. The figure below shows the performance curves across different tasks as the thinking budget varies. For simpler tasks (such as IFEval), the model's chain of thought (CoT) is shorter, and the score exhibits fluctuations as the thinking budget increases. For more challenging tasks (such as AIME and LiveCodeBench), the model's CoT is longer, and the score improves with an increase in the thinking budget.
+![thinking_budget](./thinking_budget.png)
 Here is an example with a thinking budget set to 512: during the reasoning process, the model periodically triggers self-reflection to estimate the consumed and remaining budget, and delivers the final response once the budget is exhausted or the reasoning concludes.
 ```