Well to be precise the pretraining is prevented by this but the online post training does have this issue. The reinforcement learning trials show some sort of automation but I am not aware of any unsupervised methods that have been as successful as the latest batches of llms — haven’t been in this field for a while so honestly dunno.
Yes but that’s related to LLMs specifically, not backpropagation. There’s plenty of ML paradigms that use backprop and have continual learning setups.
Backprop is the process of how the weights are updated.
Well to be precise the pretraining is prevented by this but the online post training does have this issue. The reinforcement learning trials show some sort of automation but I am not aware of any unsupervised methods that have been as successful as the latest batches of llms — haven’t been in this field for a while so honestly dunno.