• churrodestroyer@lemmy.zip
    link
    fedilink
    arrow-up
    38
    arrow-down
    1
    ·
    4 hours ago

    But just ten minutes into his presentation, a collaborator interjected telling that “something seemed off to them.” As it turns out, the way the new algorithm handled the edges of the surveys was “wrong,” Sutter admitted.

    “It wasn’t a typo, and it wasn’t a missing citation or a factor of two,” he wrote. “It was subtle, but it was very wrong, and everything downstream of it was also wrong, and I had shared the whole thing in a room full of people who trusted me.”

    “It was very wrong” HOW?!?!

    The article might as well have continued “It was so obviously wrong that to go into detail would waste our readers time as any reader of this publication is well aware of the kinds of things that go wrong with the edges of surveys.”

  • leadore@lemmy.world
    link
    fedilink
    arrow-up
    24
    arrow-down
    1
    ·
    4 hours ago

    If you read the scientist’s original article (linked to from this article), he seems to have learned very little from the experience, except to be more careful while continuing to use the LLM. He compares it to alchemy, which he dubs “pre-science”.

    He says:

    But the story doesn’t end with my humiliation. Alchemy was already on my mind, and the more I read about it, the more I realized the alchemists’ work wasn’t an embarrassment to science, it was the road that led to it. They never found the philosopher’s stone. But they found something else, something we need really badly right now.

    (though he doesn’t tell us what the ‘something else’ was). Then later,

    But dang it, LLMs are so undeniably useful. I have been writing code since I was 5 years old. I haven’t written a line of it since January. And for 20 years I did my own research, which is to say I read the papers, chased the citations down their rabbit holes, and built every argument myself. AI agents do all that now.

    These were not lazy trades. They were correct trades, and I’d make them again tomorrow. Somewhere in the past year an LLM became my thinking partner, and I say that without embarrassment. I’m not the only one. Hundreds of millions of people reach for these tools every day, and they are not idiots.

    • soratoyuki@piefed.zip
      link
      fedilink
      English
      arrow-up
      16
      ·
      2 hours ago

      Somewhere in the past year an LLM became my thinking partner, and I say that without embarrassment.

      You’re good, I have enough secondhand embarrassment for the both of us.

      • stoy@lemmy.zip
        link
        fedilink
        English
        arrow-up
        4
        ·
        2 hours ago

        Somewhere in the past year an LLM became my thinking partner,

        That statement terrifies me, I very seldom use LLMs, maybe one question in a month, and that question is usually when I have exhausted the other alternatives.

        Example, last week I attended an airshow and needed to program my radio scanner to a local radiostation where they rebroadcast the announcer to enable you to hear it clearly rather than through the annoying speakers.

        And I just could not get my Icom IC-R6 to work, it would randomly change radio modes as I was tuning, I also had trouble setting the tuning steps correctly.

        I read the manual, but I could not figure it out, so I caved and asked chatgpt, and got the help I needed in quick order.

        I am certain that I would have figured it out on my own, but it would have taken far longer.

        Anyway, I am always careful to limit my interaction with LLMs/AI as I start feeling my skills stagnating and even reducing as I use it.

    • gedfromgont@piefed.ca
      link
      fedilink
      English
      arrow-up
      9
      ·
      2 hours ago

      It really makes me sad to see actual people who are not under corporate pressure, write this out.

    • Arola@sh.itjust.works
      link
      fedilink
      arrow-up
      4
      ·
      3 hours ago

      The last couple of lines has me thinking this is satirical/sarcasm:

      “…Following his embarrassing slip-up in February, Sutter vowed that he works “differently now” and is ready to immediately “distrust” anything an AI says” and “…when we pasted his latest piece for Nautilus into AI detecting tool Pangram — which is far from perfect — it informed us that 56 percent of the text appeared to be written by an AI”.