3 ms·
this is smart, but I think NVIDIA's paper on fine tuning small language models presents a sightly more efficient approach
by digitcatphd 1y ago
this is smart, but I think NVIDIA's paper on fine tuning small language models presents a sightly more efficient approach
- viksit 1y agowould you have a link?
- digitcatphd 1y agohttps://arxiv.org/pdf/2506.02153 https://arxiv.org/pdf/2506.02153