alt text

Crudely drawn drawing of a person telling a computer “say ‘i am in pain’”. The computer replies with “> I AM IN PAIN”. The person then says “oh my god.”

yes, it’s a real thing

  • midribbon_action@lemmy.blahaj.zone
    link
    fedilink
    arrow-up
    1
    ·
    1 day ago

    modern LLMs has pretty much been around for two [years]

    What I said applies to all machine learning models ever created: inference does not make any impact on a model. Any question you ask it slides off of it just like, apparently, any attempt to explain things to you.

    • SorryQuick@lemmy.ca
      link
      fedilink
      arrow-up
      2
      ·
      1 day ago

      Have you done any ML? Because if you had, you’d know that statement is complete bullshit.

        • SorryQuick@lemmy.ca
          link
          fedilink
          arrow-up
          2
          ·
          24 hours ago

          Then you know that there is nothing stopping you from training during inference, or to modify weights in reponse to inference. It just so happens to be more efficient to train the next model instead, but again that goes back to what I said, there is no need to copy humans in this because humans aren’t efficient in everything, and certainly not in training.

            • SorryQuick@lemmy.ca
              link
              fedilink
              arrow-up
              1
              ·
              23 hours ago

              Does it matter whether it’s part of it or done immediately after? For all intents and purposes it’s the same thing. Like I said, if the input and outputs are the same, what does it matter how the process works?

              From the user’s perspective, where one question results in many inference calls, it would look like the LLM learns while it works, assuming such training would be enabled, which they obviously wouldn’t but could do.

              • midribbon_action@lemmy.blahaj.zone
                link
                fedilink
                arrow-up
                1
                ·
                23 hours ago

                I’m not talking about training, the paper this post is talking about is not about training, and no cloud llms allow their users to do training, so I have no idea what relevance it could have to this conversation. Maybe training is indeed really painful for llms, idk, that’s not what we’re talking about though.

                • SorryQuick@lemmy.ca
                  link
                  fedilink
                  arrow-up
                  1
                  ·
                  22 hours ago

                  I wasn’t talking about the paper, when I talked about the paper you ignored everything I said and moved in a different direction, which was what I responded to. If you’re wondering about the relevance, perhaps you shouldn’t have brought it up.

                  • midribbon_action@lemmy.blahaj.zone
                    link
                    fedilink
                    arrow-up
                    1
                    ·
                    22 hours ago

                    But you have never trained a model. It costs millions of dollars to do it effectively. Our only interaction with llms, unless you work at openai, has been through inference.