Benchgen

Training Method

Model & Dataset

Select a base model and training dataset from HuggingFace, BenchGen, or your saved assets

Saved to Knowledge API and reused as the default name when pushing the merged model to BenchGen.

LoRA Configuration

Low-Rank Adaptation parameters

8
464
16
4128
0.05
0.000.50

Training Summary

NameQwen2.5-0.5B-Instruct - r8
MethodSFT
ModelQwen2.5-0.5B-Instruct
DatasetCapybara
GPU SelectionAuto
LoRA Rank8
Alpha16
Dropout0.05
Learn Rate1e-4
Epochs2
Batch Size4
Max Steps50

Quick Tips

Rank 8-16 works well for most tasks

Alpha = 2xrank is the recommended ratio

50-100 steps for a quick validation run

Lower LR = more stable convergence

You'll be sent to login and brought right back to continue.