4 ms·
This is referred to as “online reinforcement learning” and is already something done by, for example Cursor for their tab prediction model. https://cursor.com/
by stevenpetryk 1y ago
This is referred to as “online reinforcement learning” and is already something done by, for example Cursor for their tab prediction model.
https://cursor.com/blog/tab-rl https://cursor.com/blog/tab-rl
- tinodb 1y agoNot sure that’s the same. They just very frequently retrain and “deploy a new model”.