3 ms·
It's larger, but there are less parameters to train for your specific use case since you are training the small matrix only, while the original ones remain unal
by TuringTest 4y ago
It's larger, but there are less parameters to train for your specific use case since you are training the small matrix only, while the original ones remain unaltered.