Dieser Bereich kann Inhalte enthalten, die nicht für alle Nutzer geeignet sind. Dazu können unter anderem Texte, Medien oder Diskussionen gehören, die als beleidigend, extremistisch, gewaltbezogen oder anderweitig belastend empfunden werden. Wenn du solche Inhalte nicht sehen möchtest, nutze bitte die jeweiligen Filter- und Meldeoptionen der Plattform oder meide entsprechende Threads/Communities.
A year ago they were similarly bad at writing code, often created unit tests that tested nothing, etc.
If the models are trained in what they’re doing wrong, that can accelerate their progress toward doing it right.
They would need to be trained for open ended creative tasks, which is just hard in the current reinforcement learning paradigm.
They still don’t get it right all the time, they just stacked a few together to filter out the obviously wrong stuff.
I think they’ll find infinite ways to fuck up. The guardrails will never be high enough, or strong enough.