Arthur Allshire
@arthurallshire
On overfitting being helpful, I don’t think it’s unique to robotic BC
eg. from InstructGPT “we find that our SFT models overfit on validation loss after 1 epoch; however, we find that training for more epochs helps both the RM score and human preference ratings”.
I think the
eg. from InstructGPT “we find that our SFT models overfit on validation loss after 1 epoch; however, we find that training for more epochs helps both the RM score and human preference ratings”.
I think the
Seohong Park@seohong_park · Aug 24Behavioral cloning mystery
seohong.me/blog/behaviora…
I wrote a new blog post about "mysteries" in behavioral cloning that appear with real-world robot data (e.g., overfitting is "good"). I also tried to demystify them and shared my thoughts!
seohong.me/blog/behaviora…
I wrote a new blog post about "mysteries" in behavioral cloning that appear with real-world robot data (e.g., overfitting is "good"). I also tried to demystify them and shared my thoughts!
Open quoted post →
2 69