--- license: apache-2.0 datasets: - TorpedoSoftware/Roblox-Luau-Reasoning-v1.0 language: - en base_model: - Qwen/Qwen3-4B-Instruct-2507 base_model_relation: finetune tags: - code - luau - roblox - dora - peft - sft library_name: transformers --- ## Training Data This model was trained on a dataset derived from [TorpedoSoftware/Roblox-Luau-Reasoning-v1.0](https://huggingface.co/datasets/TorpedoSoftware/Roblox-Luau-Reasoning-v1.0), which is released under the MIT License. The original authors are not affiliated with or responsible for this model. ## Base Model Base model: [Qwen/Qwen3-4B-Instruct-2507](https://huggingface.co/Qwen/Qwen3-4B-Instruct-2507) ## Fine-tuning Method - Adapter: DoRA - Method: SFT - Precision: trained with 4-bit base weights + BF16 compute, then merged to safetensors ## Training Details - Training time: ~12 hours - Hardware: 1x NVIDIA RTX 5060 Ti