3 ms·
Obligatory link from a sceptic: https://arxiv.org/pdf/2009.13807 https://arxiv.org/pdf/2009.13807 > However, they wanted to develop a technique that avoids fin
by beryilma 2y ago
Obligatory link from a sceptic: https://arxiv.org/pdf/2009.13807 https://arxiv.org/pdf/2009.13807
> However, they wanted to develop a technique that avoids fine-tuning, a process in which engineers retrain a general-purpose LLM on a small amount of task-specific data to make it an expert at one task.
This method does not avoid fine tuning. It just offloads the task to somebody else (i.e., to the LLM).
I'll buy the promise of the approach when the authors can show that they can vastly outperform an AR time series model or the simple techniques mentioned in the linked article.