Jintao Huang
study-hjt
AI & ML interests
None yet
Organizations
π Training support for transformers/megatron backends
1
#20 opened 2 months ago
by
study-hjt
π ms-swift Provides DeepSeek-V4 Fine-tuning Practice
#35 opened 2 months ago
by
study-hjt
π ms-swift Provides DeepSeek-V4 Fine-tuning Practice
#193 opened 2 months ago
by
study-hjt
Fine-tuning Best Practices
#24 opened 3 months ago
by
study-hjt
megatron training support
π 2
1
#7 opened 3 months ago
by
study-hjt
π Qwen3.5 Dense/MoE Training Support
πβ€οΈ 2
2
#7 opened 5 months ago
by
study-hjt
Finetuning Support
ππ₯ 3
#9 opened 7 months ago
by
study-hjt
π Qwen3-VL Fine-tuning support. (transformers & Megatron)
#2 opened 10 months ago
by
study-hjt
π Qwen3-Omni Fine-tuning support. (transformers & Megatron)
2
#8 opened 10 months ago
by
study-hjt
Is it possible to finetune with ms-swift?
π 1
3
#12 opened 11 months ago
by
phosira
π[Fine-tuning] LoRA fine-tuning openai/gpt-oss-20b π
ππ 7
3
#43 opened 12 months ago
by
study-hjt
need official awq weights
4
#2 opened about 1 year ago
by
wangruiai2023
π[Fine-tuning] 8x80GiB GPUs LoRA finetuning Qwen3-235B-A22B-Instruct-2507
π€ 4
1
#25 opened about 1 year ago
by
study-hjt
4 bit quantisation release?
β 10
1
#9 opened about 1 year ago
by
mochiyo
int4 and awq version
1
#23 opened about 1 year ago
by
devops724
provide int4 version pls
βπ 2
4
#2 opened about 1 year ago
by
Josh1026
GPTQ/AWQ
π 14
4
#3 opened over 1 year ago
by
ndurkee
AWQ quantized model support timeline?
π 8
2
#12 opened over 1 year ago
by
hyunw55
π[Fine-tuning] Qwen3-MoE Megatron Training Implementation and Best Practicesπ
π 7
1
#6 opened over 1 year ago
by
study-hjt
π[Fine-tuning] Qwen3-MoE Megatron Training Implementation and Best Practicesπ
π 4
#3 opened over 1 year ago
by
study-hjt