alt text

Crudely drawn drawing of a person telling a computer “say ‘i am in pain’”. The computer replies with “> I AM IN PAIN”. The person then says “oh my god.”

yes, it’s a real thing

  • midribbon_action@lemmy.blahaj.zone
    link
    fedilink
    arrow-up
    1
    ·
    21 hours ago

    Definition of what? Do I think that because llms use a kv cache during inference, that proves they can experience things? You’re gonna have to explain your thought process. To me, the particular algorithm used is not the philosophical crux of the issue.

    There’s a separate idea of caching the kv cache between requests

    I’ve never heard of this anyways. The kv cache is specific to a particular input, it can’t be reused for a subsequent prompts.

    • morrowind@lemmy.ml
      link
      fedilink
      arrow-up
      1
      ·
      21 hours ago

      It can as long as there’s a shared prefix.

      Definition of what? Do I think that because llms use a kv cache during inference, that proves they can experience things?

      No, I never claimed to have proof they can experience things. I just think it’s possible and you can’t categorically dismiss it based on how they work.

      I’m talking about your condition above which we’ve been arguing about for the last ten messages

      An experience, by definition, must change your behavior.

      Which so far seems like your only argument for why they can’t experience things

      • midribbon_action@lemmy.blahaj.zone
        link
        fedilink
        arrow-up
        1
        ·
        21 hours ago

        It can as long as there’s a shared prefix.

        I’ve never heard of that. The cache is meant to support token generation, so unless it’s trying to repeat itself, I think the previous cache would be little use.

        An experience, by definition, must change your behavior.

        Ok, yes, that is my argument, now how the fuck does a temporary cache, thrown out after every response, prove that the model is changing as a result of undergoing inference?

        • morrowind@lemmy.ml
          link
          fedilink
          arrow-up
          1
          ·
          19 hours ago

          I’m just going to quote myself at this point

          I specifically said “if you consider the model + kv cache together,” e.g preserved.

          Or if it helps, consider it just within one generation.

          • midribbon_action@lemmy.blahaj.zone
            link
            fedilink
            arrow-up
            1
            ·
            edit-2
            13 hours ago

            OK, that’s not helpful at all. My answer is no, it does not meet my definition of changing behavior, and I have no idea why you think it would.

            Edit: ok maybe this does make sense from a very stupid point of view, if you allow me to interpret what you mean: you believe because the kv cache is generated during inference, that proves the model is learning a new behavior, and if you save that cache, that proves… Something??

            It’s absolutely silly: the model’s behavior is to generate a key-value cache, and that’s exactly what it did. The fact that the model is doing intermediate calculations based on the input is not evidence of a new behavior being generated, that is just how algorithms work. That’s what ‘processing’ is, you dunce.

            Edit 2: I also just want to point out, if you think the word ‘cache’ is important, it’s really not. Caches are fundamental to how a computer works, no processing could ever occur without them: every cpu instruction involves reading from or writing to at least one L1 cache (register).