L0-Luau-4B-Instruct / README.md
r-e1's picture
Update README.md
2a54433 verified
|
Raw
History Blame Contribute Delete
887 Bytes
---
license: apache-2.0
datasets:
- TorpedoSoftware/Roblox-Luau-Reasoning-v1.0
language:
- en
base_model:
- Qwen/Qwen3-4B-Instruct-2507
base_model_relation: finetune
tags:
- code
- luau
- roblox
- dora
- peft
- sft
library_name: transformers
---
## Training Data
This model was trained on a dataset derived from
[TorpedoSoftware/Roblox-Luau-Reasoning-v1.0](https://huggingface.co/datasets/TorpedoSoftware/Roblox-Luau-Reasoning-v1.0),
which is released under the MIT License.
The original authors are not affiliated with or responsible for this model.
## Base Model
Base model: [Qwen/Qwen3-4B-Instruct-2507](https://huggingface.co/Qwen/Qwen3-4B-Instruct-2507)
## Fine-tuning Method
- Adapter: DoRA
- Method: SFT
- Precision: trained with 4-bit base weights + BF16 compute, then merged to safetensors
## Training Details
- Training time: ~12 hours
- Hardware: 1x NVIDIA RTX 5060 Ti
Free AI Image Generator No sign-up. Instant results. Open Now