3 ms·
Silly question time. Is this a fined tuned LLM, for example drop in replacement for Llama etc. Or is it some algorithm on top of an LLM, doing some chain of r
by nwnwhwje 2y ago
Silly question time.
Is this a fined tuned LLM, for example drop in replacement for Llama etc.
Or is it some algorithm on top of an LLM, doing some chain of reasoning?
- peakji 2y agoIt is an LLM fine-tuned using a new type of dataset and RL reward. It's good at reasoning, but I would not recommend to replace Llama for general tasks.