• TIEPilot@lemmy.world
    link
    fedilink
    English
    arrow-up
    15
    arrow-down
    1
    ·
    7 days ago

    While I’m not a fan of scrapping you would think reddit would want this to push traffic to the site. Get a good answer and follow the referral link and you have a new monkey.

    • bugbear@lemmy.sdf.org
      link
      fedilink
      English
      arrow-up
      14
      ·
      7 days ago

      Are there any statistics on how many LLM chatbot users actually click on the source links?

      • Ludicrous0251@piefed.zip
        link
        fedilink
        English
        arrow-up
        3
        ·
        7 days ago

        I mean just about every metric I’ve seen suggests most websites that see an increase in scraping see a decrease in actual human visitors. I doubt there’s much distinction between chat and search users - the goal of these services is to drop click-through rates.

        • bugbear@lemmy.sdf.org
          link
          fedilink
          English
          arrow-up
          1
          ·
          7 days ago

          Oh, definitely. They are reinventing the old portals concept. That’s why I don’t think Reddit actually wants to encourage the scraping (partly also because they sell data themselves to build LLM, but that’s another thing)

      • palordrolap@fedia.io
        link
        fedilink
        arrow-up
        1
        ·
        7 days ago

        I don’t know how many do, but they’re fools if they don’t.

        You’d think that since those links are there, that the information on those links matches what that LLM says, and thus not a hallucination, but at least twice, I’ve found the information straight up wasn’t there, or the LLM had misread / misinterpreted the page’s content.

        But then I may not be a regular user. I ask a question maybe once every one to two weeks because a regular search finds nothing and, like I say, I prefer to check the sources.

      • TIEPilot@lemmy.world
        link
        fedilink
        English
        arrow-up
        5
        ·
        7 days ago

        Great question, I wish I knew. I know when I’m on a deep dive I will hit the links as there is other information or other articles the scrapped site has.

      • AA5B@lemmy.world
        link
        fedilink
        English
        arrow-up
        1
        ·
        7 days ago

        I do for things where I want to be accurate, but sometimes it seems like I’m the only one

        On the other hand I advised my teen to buy the wrong battery because I didn’t click into the results for “what battery do i need to replace in my Subaru key?”

    • BassTurd@lemmy.world
      link
      fedilink
      English
      arrow-up
      4
      ·
      7 days ago

      I think it had the opposite effect. I believe the recently didn’t renew their contract with Google because they were upset people weren’t coming to the site, just getting answers from Gemini. I could be wrong, but I thought I read that

      • valkyre09@lemmy.world
        link
        fedilink
        English
        arrow-up
        4
        arrow-down
        2
        ·
        7 days ago

        The guys on LTT mentioned this during their Linux challenge, but I’ve found it to be true -

        Scouring forums for answers to obscure errors used to take hours. Now I just throw it at Gemini, troubleshoot in real time and find the fix.

        I’ve even started asking it to give me a PDF of the steps we took and filing them away for future me.

          • valkyre09@lemmy.world
            link
            fedilink
            English
            arrow-up
            2
            ·
            7 days ago

            Tonight I spent the time spinning up a silverbullet instance and converting the PDFs to .md files. You’re right, PDFs aren’t searchable and will be a pain as the database grows.

            Thanks for sharing :)