Audrey Tang

So, bridging rewards are often already in pre-training defaults for language models, because forum threads that end with someone mediating and pulling both sides together become good training data. Wikipedia, GitHub, Common Crawl — lots of datasets carry that capacity.

鍵盤快捷鍵Keyboard shortcuts

j 下一段next speechk 上一段previous speech