Why can't we just run local reinforcement learning?

Revolutionalredstone@alien.top · 2 years ago

Why can't we just run local reinforcement learning?

LuluViBritannia@alien.top · 2 years ago

Well, first of all, this is something you do while running the model. Sure, it’s the same model, but it’s still two different processes to run in parallel.

Then, from what I gather, it’s closer to model finetuning than it is to inference. And if you look up the figures, finetune requires a lot more power and VRAM. As I said, it’s rewriting the neural network, which is the definition of finetuning.

So in order to get a more specific answer, we should look up why finetuning requires more than inference.