I was recently reading Ben Aaronvitch’s ‘Rivers of London’ and I was struggling to follow the plot. So I engaged in discussions with both ChatGPT and Gemini and both really struggled to not make lots of mistakes. Both AIs would regularly give inaccurate info about major plot points, and when I questioned them about those issues, they would apologize profusely and then try to continue on. I basically had to coach them through to get them to remember important points of the plot! They also struggled to not reveal spoilers past a certain chapter. It was interesting to see the AIs repeatedly fail.

  • hdsrob@lemmy.world
    link
    fedilink
    English
    arrow-up
    39
    arrow-down
    1
    ·
    1 day ago

    That’s literally how the work: They can’t reason or think. They don’t know or understand anything. They just output strings of words that are statistically likely to be the response to your question based on strings of words that they’ve been trained on.

    • fbn@slrpnk.net
      link
      fedilink
      English
      arrow-up
      10
      ·
      1 day ago

      I’ve heard it referred to as spicy autocomplete and also next word extruder

    • MooseInMaine@lemmy.worldOP
      link
      fedilink
      arrow-up
      5
      arrow-down
      2
      ·
      1 day ago

      I did have a much better discussion of ‘Moby Dick’ a while ago, but I imagine there is lots of info that AI has scraped from scholarly articles and the like about classics.

      • sudoMakeUser@sh.itjust.works
        link
        fedilink
        arrow-up
        6
        ·
        1 day ago

        They don’t do well with big sources like books unless they’re discussed. If there’s no online references or scholarly articles it won’t know the contents of the book. Same with movies or music.

  • e0qdk@reddthat.com
    link
    fedilink
    arrow-up
    9
    arrow-down
    2
    ·
    1 day ago

    LLMs don’t really know the subject matter they talk about in the same way a person does. Whatever you feed as input biases the responses towards and away from certain sequences of words. If you put in something about Star Wars, it’ll weigh its output in favor of saying things like “Luke Skywalker” and “Darth Vader” and away from things like “The best way to use a crockpot is…”. When you’re trying to get it to recall facts about things that aren’t extremely famous, it will jumble some parts or just make things up that sound plausible.

    As an example, I liked to prompt them with “Tell me about ‘A Valiant Effort (1989)’” as a test – it doesn’t actually exist, but they’d often tell me it was a painting or a movie or an obscure book with hallucinated details until very recently; newer models seem more likely to say “I don’t know what that is” or similar, but that’s not guaranteed.

    To get better results, try feeding it a passage you’d like it to work on directly in the input instead of assuming it just knows the story already, and you’ll get output weighted by the actual text (and whatever else you put in context).

  • terabyterex@lemmy.world
    link
    fedilink
    arrow-up
    1
    ·
    edit-2
    1 day ago

    i am guessing you arr using the free versions but i can imagine this type of thing is more open to hallucinations. so taking, replayjng text coupled with free model. wont be the best.

    are you able to attach the book to your chat inn the free version? do you have a digital copy?

      • Walters, Fern@bookwyr.me
        link
        fedilink
        English
        arrow-up
        2
        ·
        1 day ago

        Don’t.

        Next time just make a «[Discussion][Help] ’s Rivers of London — Ben Aaronvitch 9780575097568»🧵, and wait for everyone else interested in responding to do so.

        • MooseInMaine@lemmy.worldOP
          link
          fedilink
          arrow-up
          1
          ·
          22 hours ago

          I’m new here, so I don’t really know about this, but it looks like an attempt to create individual threads for specific books?

          • Walters, Fern@bookwyr.me
            link
            fedilink
            English
            arrow-up
            2
            arrow-down
            1
            ·
            edit-2
            21 hours ago

            it looks like an attempt to create individual threads for specific books?

            Correct. You can also make it a weekly. e.g.


            Weekly 📚 share discussion mega🧵

            Today I read upto ⸿8 of Rivers of London — Ben Aaronvitch 978057509756, and I am having issues comprehending

            spoiler

            XYZ plot point in ⸿7 [⚠️haven’t read ⸿9 yet, no spoilers past it!⚠️] Could a reader help me out?


            are you new new to threaded internet?

  • lIlIlIlIlIlIl@lemmy.world
    link
    fedilink
    arrow-up
    3
    arrow-down
    3
    ·
    1 day ago

    Did the AI “fail,” or did you use AI in a suboptimal way?

    If you had started the conversation by feeding in the actual text of the book, you would have very different results, because you would be grounding on context rather than asking the model to “figure it out somehow” based on training data.

    Not sure what you were expecting, using AI is mostly about context management, and if you don’t set up the context correctly you’re asking for results similar to this.

    • MooseInMaine@lemmy.worldOP
      link
      fedilink
      arrow-up
      1
      ·
      edit-2
      1 day ago

      It appeared it had access to the entire book - it linked me to a website that somehow had the text. I was able to coax better responses out of it when I corrected its mistakes about plot etc.

      • lIlIlIlIlIlIl@lemmy.world
        link
        fedilink
        arrow-up
        2
        ·
        1 day ago

        It appeared it had access

        Unless you provided it, the context window did not have the context. “Appeared to” is also not a valid way to evaluate technology.