Updated: Jan 7, 2026
styleDownload
1 variant available
Stats
740 1 2 3 4 5 6 7 8 90 1 2 3 4 5 6 7 8 9
100 1 2 3 4 5 6 7 8 90 1 2 3 4 5 6 7 8 9
40 1 2 3 4 5 6 7 8 9
Reviews
(6)
Published
Nov 30, 2025
Base Model
Hash
AutoV2
BA96415E6E
retraining pairwise head on different data mixture
different regularization during rl training


1.4K0 1 2 3 4 5 6 7 8 9.0 1 2 3 4 5 6 7 8 9K
3.3K0 1 2 3 4 5 6 7 8 9.0 1 2 3 4 5 6 7 8 9K
33K0 1 2 3 4 5 6 7 8 90 1 2 3 4 5 6 7 8 9K

License:
Illustrious LicenseLoRAs trained using DRaFT-like post-RL using Siglip/Dino
Don't expect sensible results on any other model than indexed v2
These are just experiments. I will often not avoid over-training and forgo stronger regularization as I am more so looking for what effects are most prominent for varying data regimes.
