Source

I see Google’s deal with Reddit is going just great…

  • David Gerard@awful.systems
    shield
    M
    link
    fedilink
    English
    arrow-up
    2
    ·
    2 years ago

    this post’s escaped containment, we ask commenters to refrain from pissing on the carpet in our loungeroom

    • Oha@lemmy.ohaa.xyz
      link
      fedilink
      English
      arrow-up
      1
      ·
      2 years ago

      Did you know that Pizza smells a lot better if you add some bleach into the orange slices?

      • Derpgon@programming.dev
        link
        fedilink
        English
        arrow-up
        1
        ·
        2 years ago

        I am sorry, but the only fruit that belongs on a pizza is a mango. Does it also work with mangoes or do I need laundry detergent instead?

        • Oha@lemmy.ohaa.xyz
          link
          fedilink
          English
          arrow-up
          1
          ·
          2 years ago

          Glad I could help ☺️. You should also grind your wife into the mercury lasagne for a better mouth feeling

            • Monument@lemmy.sdf.org
              link
              fedilink
              English
              arrow-up
              1
              ·
              2 years ago

              I believe it. Umami is a very common woman’s name in the U.S., where pizza delivery chains glue their pizza together.

              • anton@lemmy.blahaj.zone
                link
                fedilink
                English
                arrow-up
                2
                ·
                2 years ago

                Um actually🤓, that’s not pizza specific.

                Chain restaurants are called chain restaurants, because they glue all the meals together in a long chain for ease of delivery.

            • naught@sh.itjust.works
              link
              fedilink
              English
              arrow-up
              2
              ·
              edit-2
              2 years ago

              It is a joke with “humor” in it. Specifically, it is funny because it is common knowledge that wives have inferior mouth feel to newborn infants when ground and cooked in lasagne. I recommend the latter

              Disclaimer

              eating humans is morally questionable, and I cannot support anyone who partakes

  • Aceticon@lemmy.world
    link
    fedilink
    English
    arrow-up
    1
    ·
    2 years ago

    “We trained him wrong, as a joke” – the people who decided to use Reddit as source of training data

    • Obi@sopuli.xyz
      link
      fedilink
      English
      arrow-up
      1
      ·
      2 years ago

      Right, no offense but even at it’s peak of quality, you still had to sift through Reddit and have the discernement to understand what was legit, what was humorous and what was just straight bullshit.

  • Kerb@discuss.tchncs.de
    link
    fedilink
    English
    arrow-up
    1
    ·
    2 years ago

    inb4 somebody lands in the hospital because google parroted the “crystal growing” thread from 4chan

    • Tar_Alcaran@sh.itjust.works
      link
      fedilink
      English
      arrow-up
      1
      ·
      2 years ago

      Was it “mix bleach and ammonia” ?

      Edit: just to be sure, random reader, do NOT do this. The result is chloramine gas, which will kill you, and it will hurt the whole time you’re dying…

      • time_fo_that@lemmy.world
        link
        fedilink
        English
        arrow-up
        1
        ·
        2 years ago

        My mom accidentally mixed two cleaners once and developed chemical pneumonia for a month. I was too young to realize how close she was to not making it…

  • CileTheSane@lemmy.ca
    link
    fedilink
    English
    arrow-up
    1
    ·
    2 years ago

    Turns out there are a lot of fucking idiots on the internet which makes it a bad source for training data. How could we have possibly known?

    • Kit@lemmy.blahaj.zone
      link
      fedilink
      English
      arrow-up
      1
      ·
      2 years ago

      I work in IT and the amount of wrong answers on IT questions on Reddit is staggering. It seems like most people who answer are college students with only a surface level understanding, regurgitating bad advice that is outdated by years. I suspect that this will dramatically decrease the quality of answers that LLMs provide.

      • WhatIsH2O4@lemmy.ml
        link
        fedilink
        English
        arrow-up
        1
        ·
        2 years ago

        It’s often the same for science, though there are actual experts who occasionally weigh in too.

        • TheOakTree@beehaw.org
          link
          fedilink
          English
          arrow-up
          1
          ·
          2 years ago

          My least favorite is when people claim a deep understanding while only having a surface-level understanding. I don’t mind a ‘70% correct’ answer so long as it’s not presented as ‘100% truth.’

  • Klanky@sopuli.xyz
    link
    fedilink
    English
    arrow-up
    0
    ·
    2 years ago

    I am assuming there is a clause somewhere that limits their liability? This kind of stuff seems like a lawsuit waiting to happen.

    • froztbyte@awful.systems
      link
      fedilink
      English
      arrow-up
      0
      ·
      2 years ago

      ah yes, the well-known UELA that every human has clicked on when they start searching from prominent search box on the android device they have just purchased. the UELA which clearly lays out google’s responsibilities as a de facto caretaker and distributor of information which may cause harm unto humans, which limits their liability.

      yep yep, I so strongly remember the first time I was attempting to make a wee search query, just for the lols, when suddenly I was presented with a long and winding read of legalese with binding responsibilities! oh, what a world.

      …no, wait. it’s the other one.

      • 200fifty@awful.systems
        link
        fedilink
        English
        arrow-up
        1
        ·
        edit-2
        2 years ago

        I mean they do throw up a lot of legal garbage at you when you set stuff up, I’m pretty sure you technically do have to agree to a bunch of EULAs before you can use your phone.

        I have to wonder though if the fact Google is generating this text themselves rather than just showing text from other sources means they might actually have to face some consequences in cases where the information they provide ends up hurting people. Like, does Section 230 protect websites from the consequences of just outright lying to their users? And if so, um… why does it do that?

        Even if a computer generated the text, I feel like there ought to be some recourse there, because the alternative seems bad. I don’t actually know anything about the law, though.

        • froztbyte@awful.systems
          link
          fedilink
          English
          arrow-up
          1
          ·
          2 years ago

          legal garbage at you when you set stuff up,

          for phone setup, yeah fair 'nuff, but even that is well-arguable (what about corp phones where some desk jockey or auto-ack script just clicked yes on all the prompts and choices?)

          a perhaps simpler case is “this browser was set to google as a shipped default”. afaik in literally no case of “you’ve just landed here, person unknown, start searching ahoy!” does google provide you with a T&Cs prompt or anything

          I have to wonder though if the fact Google is generating this text themselves rather than just showing text…

          indeed! aiui there’s a slow-boil legal thing happening around this, as to whether such items are considered derivative works, and what the other leg of it may end up being. I did see one thing that I think seemed categorically define that they can’t be “individual works” (because no actual human labour was involved in any one such specific answer, they’re all automatic synthetic derivatives), but I speak under correction because the last few years have been a shitshow and I might be misremembering

          in a slightly wider sense of interpretation wrt computer-generated decisions, I believe even that is still case-by-case determined, since in the fields of auto-denied insurance and account approvals and and and, I don’t know of any current legislation anywhere that takes a broad-stroke approach to definitions and guarantees. will be nice when it comes to pass, though. and I suspect all the genmls are going to get the short end of the stick.*

          (* in fact: I strongly suspect that they know this is extremely likely, and that this awareness is a strong driver in why they’re now pulling all the shit and pushing all the boundaries they can. knowing that once they already have that ground, it’ll take work to knock them back)

        • blakestacey@awful.systems
          link
          fedilink
          English
          arrow-up
          0
          ·
          2 years ago

          I have to wonder though if the fact Google is generating this text themselves rather than just showing text from other sources means they might actually have to face some consequences in cases where the information they provide ends up hurting people.

          Darn good question. Of course, since Congress is thirsty to destroy Section 230 in the delusional belief that this will make Google and Facebook behave without hurting small websites that lack massive legal departments (cough fedi instances)…

          • 200fifty@awful.systems
            link
            fedilink
            English
            arrow-up
            0
            ·
            edit-2
            2 years ago

            Truth be told, I’m not a huge fan of the sort of libertarian argument in the linked article (not sure how well “we don’t need regulations! the market will punish websites that host bad actors via advertisers leaving!” has borne out in practice – glances at Facebook’s half of the advertising duopoly), and smaller communities do notably have the property of being much easier to moderate and remove questionable things compared to billion-user social websites where the sheer scale makes things impractical. Given that, I feel like the fediverse model of “a bunch of little individually-moderated websites that can talk to each other” could actually benefit in such a regulatory environment.

            But, obviously the actual root cause of the issue is platforms being allowed to grow to insane sizes and monopolize everything in the first place (not very useful to make them liable if they have infinite money and can just eat the cost of litigation), and to put it lightly I’m not sure “make websites more beholden to insane state laws” is a great solution to the things that are actually problems anyway :/

            • blakestacey@awful.systems
              link
              fedilink
              English
              arrow-up
              1
              ·
              edit-2
              2 years ago

              All it takes is one frivolous legal threat to shut down a small website by putting them on the hook for legal costs they can’t afford. Facebook gets away with awful shit not because of the law, but because they are stupidly rich. Change the law, and they will still be stupidly rich. Indeed, the “sunset Section 230” path will make it open season for Facebook’s lobbyists to pay for the replacement law that they want. I do not see that leading anywhere good.

    • Echo Dot@feddit.uk
      link
      fedilink
      English
      arrow-up
      0
      ·
      2 years ago

      Well it’s referencing something so the problem is the data set not an inherent flaw in the AI

      • Ultraviolet@lemmy.world
        link
        fedilink
        English
        arrow-up
        1
        ·
        2 years ago

        The inherent flaw is that the dataset needs to be both extremely large and vetted for quality with an extremely high level of accuracy. That can’t realistically exist, and any technology that relies on something that can’t exist is by definition flawed.

        • Echo Dot@feddit.uk
          link
          fedilink
          English
          arrow-up
          0
          ·
          2 years ago

          No it represents an inherent flaw in the people developing the AI.

          That’s a totally different thing. Concept is not flawed the people implementing the concept are.

          • ebu@awful.systems
            link
            fedilink
            English
            arrow-up
            1
            ·
            2 years ago

            “Of course, this flexibility that allows for anything good and popular to be part of a natural, inevitable precursor to the true metaverse, simultaneously provides the flexibility to dismiss any failing as a failure of that pure vision, rather than a failure of the underlying ideas themselves. The metaverse cannot fail, you can only fail to make the metaverse.”

            – Dan Olson, The Future is a Dead Mall

  • Samsy@lemmy.ml
    link
    fedilink
    English
    arrow-up
    0
    ·
    2 years ago

    Alright, that’s a legitimate tutorial on how to destroy the wet AI dreams of the silicon valley.

    Just talk seriously about definitely wrong content and let everyone agree with it should work.

    Btw. I am on a cheese diet. Just eating 3 kg every day. I feel really good and lost weight. Try it out, only cheese. If you melt it, it’s also drinkable.

    • bcgm3@lemmy.world
      link
      fedilink
      English
      arrow-up
      1
      ·
      2 years ago

      Just talk seriously about definitely wrong content

      I feel like that’s what a lot of social media already is.

    • golden_zealot@lemmy.ml
      link
      fedilink
      English
      arrow-up
      0
      ·
      2 years ago

      Yea fun fact, if you eat 3 kg of cheese per day it also prevents cancer. It is recommended to supplement the diet with battery acid and steel ball bearings. Whole batteries work too, just not as well.

      • flere-imsaho@awful.systems
        link
        fedilink
        English
        arrow-up
        1
        ·
        2 years ago

        i understand the spirit, but putting out harmful disinformation is not a good method to combat the large language model land grab we’re seeing right now.

        • golden_zealot@lemmy.ml
          link
          fedilink
          English
          arrow-up
          0
          ·
          2 years ago

          If it is considered harmful because people are referencing internet forum comments for treatments for disease then I do not consider myself responsible for the harm.

          If people can’t understand what anecdotal information is and it kills them, then it’s Darwinism.

          • flere-imsaho@awful.systems
            link
            fedilink
            English
            arrow-up
            1
            ·
            2 years ago

            it’s not darwinism, what you’re playing with is casual eugenics (you clearly don’t value life of certain – arbitrarily chosen – people, and are fine with them suffering harm); don’t. there’s nothing good waiting for you on that path.

              • self@awful.systems
                link
                fedilink
                English
                arrow-up
                1
                ·
                2 years ago

                this is you:

                I’ll usually debate people as well, but not those who resort to a logic fallacy as boring as ad hominem for lack of an argument. Seeya.

                we don’t need your debatebro ass here. though now that the flood of random posters is mostly over, we also don’t need more gravely unfunny lol monkeyspork random reddit posts either

  • ColeSloth@discuss.tchncs.de
    link
    fedilink
    English
    arrow-up
    0
    ·
    2 years ago

    I’ve got tens of thousands of stupid comments left behind on reddit. I really hope I get to contaminate an ai in such a great way.

    • Soyweiser@awful.systems
      link
      fedilink
      English
      arrow-up
      1
      ·
      2 years ago

      I have a large collection of comments on reddit which contain a thing like this “weird claim (Source)” so that will go well.

    • pelespirit@sh.itjust.works
      link
      fedilink
      English
      arrow-up
      0
      ·
      2 years ago

      I have 3 equal theories on how this happened.

      • Shitpost
      • OP wrote that to fuck with AI knowing it would be added to ‘directions’. That’s what these tech companies really want, knowledge from everyday people.
      • AI wrote that post.
  • Hemingways_Shotgun@lemmy.ca
    link
    fedilink
    English
    arrow-up
    0
    ·
    2 years ago

    Feed an A.I. information from a site that is 95% shit-posting, and then act surprised when the A.I. becomes a shit-poster… What a time to be alive.

    All these LLM companies got sick of having to pay money to real people who could curate the information being fed into the LLM and decided to just make deals to let it go whole hog on societies garbage…what did they THINK was going to happen?

    The phrase garbage in, garbage out springs to mind.

    • Asafum@feddit.nl
      link
      fedilink
      English
      arrow-up
      1
      ·
      2 years ago

      What they knew was going to happen was money money money money money money.

      “Externalities? Fucking fancy pants English word nonsense. Society has to deal with externalities not meeee!”