4 ms·
Baidu's Improving Retrieval Augmented Language Model with Self-Reasoning
- kaspermarstal 2y agoCan anyone explain what is gained by training a model? Why not use the foundational LLM for the relevance, evidence, and trajectory processes?
- a-s-k-af 2y agoI assume you are referring to fine tuning a model here?
- Tostino 2y agoYou could also just continue pre-training of an existing foundation model. Would still be cheaper by not starting from zero.
- a-s-k-af 2y agoThe amount of accuracy while doing fine tuning or distillation is usually better than pre-training an existing model, not to mention the graph against the cost.
- deleted 2y ago[deleted]