Efficient Fine-Tuning of DeepScaleR-1.5B Without Increasing Parameters

by HassanStar - opened 9 days ago

9 days ago

•

What are the best methods for fine-tuning DeepScaleR-1.5B without increasing the number of parameters during inference? Would LoRA or other PEFT methods be effective, and what settings are recommended?

HassanStar changed discussion title from Fine-Tuning DeepScaleR-1.5B Without Increasing Parameters to Efficient Fine-Tuning of DeepScaleR-1.5B Without Increasing Parameters 9 days ago

sijuntan

Agentica org 9 days ago

We don't have any specific recommendation here. For DeepScaleR's training we do full finetuning, but you can also try out LoRA to see if it works well.

HassanStar

4 days ago

We don't have any specific recommendation here. For DeepScaleR's training we do full finetuning, but you can also try out LoRA to see if it works well.

thanks will try it soon!

michaelzhiluo

Agentica org 4 days ago

Some people on Twitter have tried LoRA, apparently it does even better!

youyc22

3 days ago

Some people on Twitter have tried LoRA, apparently it does even better!

Could you provide a link？

Upload images, audio, and videos by dragging in the text input, pasting, or clicking here.

Tap or paste here to upload images

· Sign up or log in to comment