TBPN
@tbpn
.@nathanbenaich explains why AI labs are shifting compute from pretraining to RL:
"The last 4-5 years have been trying to figure out what the recipe is for pretraining, what is the best data mixture, what are the best ingredients to this whole magical soup."
"Then over the
"The last 4-5 years have been trying to figure out what the recipe is for pretraining, what is the best data mixture, what are the best ingredients to this whole magical soup."
"Then over the
2 64