2 ms·
How much data is the model trained on?
by leftstrokeviral 1y ago
How much data is the model trained on?
- dvrp 1y agoCopying and pasting Sangwu’s answer: We used two types of datasets for post-training. Supervised finetuning data and preference data used for RLHF stage. You can actually use less than < 1M samples to significantly boost the aesthetics. Quality matters A LOT. Quantity helps with generalisation and stability of the checkpoints though.
- lawlessone 1y agoHow is data acquired and curated?