4 ms·No, there are more training tokens than parameters in LLMs. They are in the classical first descent setting.by mxwsn 4mo agoNo, there are more training tokens than parameters in LLMs. They are in the classical first descent setting.