L0-Luau-4B-Instruct / README.md
r-e1's picture
Update README.md
2a54433 verified
|
Raw
History Blame Contribute Delete
887 Bytes
metadata
license: apache-2.0
datasets:
  - TorpedoSoftware/Roblox-Luau-Reasoning-v1.0
language:
  - en
base_model:
  - Qwen/Qwen3-4B-Instruct-2507
base_model_relation: finetune
tags:
  - code
  - luau
  - roblox
  - dora
  - peft
  - sft
library_name: transformers

Training Data

This model was trained on a dataset derived from TorpedoSoftware/Roblox-Luau-Reasoning-v1.0, which is released under the MIT License.

The original authors are not affiliated with or responsible for this model.

Base Model

Base model: Qwen/Qwen3-4B-Instruct-2507

Fine-tuning Method

  • Adapter: DoRA
  • Method: SFT
  • Precision: trained with 4-bit base weights + BF16 compute, then merged to safetensors

Training Details

  • Training time: ~12 hours
  • Hardware: 1x NVIDIA RTX 5060 Ti
Free AI Image Generator No sign-up. Instant results. Open Now