jondurbin/truthy-dpo-v0.1
Viewer • Updated • 1.02k • 3.06k • 140
Shout-out to @huihui-ai for the abliterated model!
huihui-ai/Qwen2.5-VL-7B-Instruct-abliterated finetuned on:
QLoRA ORPO tuned with 1x RTX A6000 for 2 epochs.
Base model
Qwen/Qwen2.5-VL-7B-Instruct