• captain_solanum@sh.itjust.works
    link
    fedilink
    English
    arrow-up
    1
    arrow-down
    1
    ·
    5 days ago

    I know what an LLM is and how it works. The model for them you currently use to understand them is really bad, I’m sorry to say. It just cannot explain how in context learning is possible, prediction of linebreaks and the model recalling what happened 200 tokens back (where is that information written on the “the” die?), etc. You almost certainly have a deeper understanding of how LLMs work that you have simplified away, if not watch this and then the thousand other more recent videos on how they actually work. You just need to switch from the equivalent model of “gravity makes things fall to the ground” to the equivalent of newtons gravitational laws. Otherwise you will be dumbfounded by completely reasonable things, and forced to reject them in favour of the flawed model you are using.

    • h0tbeef@lemmy.zip
      link
      fedilink
      English
      arrow-up
      1
      ·
      5 days ago

      Yeah, the point of my explanation was to be simplified.

      Both of your links ultimately back up what I’m saying tho.

      The computer isn’t thinking, it’s doing a math equation.