• MangoCats@feddit.it
    link
    fedilink
    English
    arrow-up
    3
    arrow-down
    3
    ·
    19 hours ago

    A year ago they were similarly bad at writing code, often created unit tests that tested nothing, etc.

    If the models are trained in what they’re doing wrong, that can accelerate their progress toward doing it right.

    • richmondez@lemdro.id
      link
      fedilink
      English
      arrow-up
      2
      ·
      13 hours ago

      They still don’t get it right all the time, they just stacked a few together to filter out the obviously wrong stuff.

    • sourdough@lemmy.world
      link
      fedilink
      English
      arrow-up
      4
      ·
      18 hours ago

      They would need to be trained for open ended creative tasks, which is just hard in the current reinforcement learning paradigm.