4 ms·
Seems like this is already being answered: https://arxiv.org/abs/2407.10930 https://arxiv.org/abs/2407.10930 https://arxiv.org/abs/2006.04439 https://arxiv.org
by johnsutor 2y ago
Seems like this is already being answered:
https://arxiv.org/abs/2407.10930 https://arxiv.org/abs/2407.10930
https://arxiv.org/abs/2006.04439 https://arxiv.org/abs/2006.04439
- valine 2y agoNot really the first paper is just fine-tuning on synthetic data. The second paper doesn’t optimize the model weights.