8 ms·
Re continuous fine-tuning: how do you avoid catastrophic forgetting in your proposal?
by fittingopposite 6mo ago
Re continuous fine-tuning: how do you avoid catastrophic forgetting in your proposal?
- SphericalCowww 6mo agoMy understanding is that this is what the LoRAs are for; my belief is that they serve as "memory" to their live observations (a more NN-like cache, say), while the main LLM remains unchanged. These LoRAs are also weighted, so that LoRAs irrelevant to the current task will not be trained, while the relevant LoRAs will be reinforced. But I never built it, so I am not sure if such an emergent state will appear or not.
- SphericalCowww 6mo agoI put my ideas here in case you are interested: https://github.com/SphericalCowww/ML_LunaLoRA https://github.com/SphericalCowww/ML_LunaLoRA