alt text

Crudely drawn drawing of a person telling a computer “say ‘i am in pain’”. The computer replies with “> I AM IN PAIN”. The person then says “oh my god.”

yes, it’s a real thing

  • SorryQuick@lemmy.ca
    link
    fedilink
    arrow-up
    2
    ·
    1 day ago

    Have you done any ML? Because if you had, you’d know that statement is complete bullshit.

      • SorryQuick@lemmy.ca
        link
        fedilink
        arrow-up
        2
        ·
        1 day ago

        Then you know that there is nothing stopping you from training during inference, or to modify weights in reponse to inference. It just so happens to be more efficient to train the next model instead, but again that goes back to what I said, there is no need to copy humans in this because humans aren’t efficient in everything, and certainly not in training.

          • SorryQuick@lemmy.ca
            link
            fedilink
            arrow-up
            1
            ·
            24 hours ago

            Does it matter whether it’s part of it or done immediately after? For all intents and purposes it’s the same thing. Like I said, if the input and outputs are the same, what does it matter how the process works?

            From the user’s perspective, where one question results in many inference calls, it would look like the LLM learns while it works, assuming such training would be enabled, which they obviously wouldn’t but could do.

            • midribbon_action@lemmy.blahaj.zone
              link
              fedilink
              arrow-up
              1
              ·
              23 hours ago

              I’m not talking about training, the paper this post is talking about is not about training, and no cloud llms allow their users to do training, so I have no idea what relevance it could have to this conversation. Maybe training is indeed really painful for llms, idk, that’s not what we’re talking about though.

              • SorryQuick@lemmy.ca
                link
                fedilink
                arrow-up
                1
                ·
                23 hours ago

                I wasn’t talking about the paper, when I talked about the paper you ignored everything I said and moved in a different direction, which was what I responded to. If you’re wondering about the relevance, perhaps you shouldn’t have brought it up.

                • midribbon_action@lemmy.blahaj.zone
                  link
                  fedilink
                  arrow-up
                  1
                  ·
                  23 hours ago

                  But you have never trained a model. It costs millions of dollars to do it effectively. Our only interaction with llms, unless you work at openai, has been through inference.