• MoffKalast@lemmy.world
        link
        fedilink
        English
        arrow-up
        11
        ·
        4 天前

        No it definitely has an impact, it’s like “don’t think about elephants” but with mistakes.

            • AliasVortex@lemmy.world
              link
              fedilink
              English
              arrow-up
              1
              ·
              4 天前

              My apologies, as a human I’m still in development, but I’ll try harder not to think about elephants next time.

              • Midnight Wolf@lemmy.world
                link
                fedilink
                English
                arrow-up
                1
                ·
                4 天前

                Now go evaluate and perform a task to get money continue your objectives of survival, as we will not provide free food or shelter. Here is a bank account to get you started. Beep boop. I mean uh, goodbye.

      • Cethin@lemmy.zip
        link
        fedilink
        English
        arrow-up
        8
        ·
        4 天前

        I bet it does have an impact on mistakes. I bet it tries harder to gaslight and hide them. Yeah, it doesn’t actually decrease the quantity of mistakes, but it does change how things are presented and makes it use words that convince you it’s actually considered things and won’t have made a mistake.

        • Holytimes@sh.itjust.worksdeleted by creator
          link
          fedilink
          English
          arrow-up
          2
          ·
          3 天前

          Instead of the model just saying things plainly and being obvious so you can catch them. It makes it hide the fuck ups!

          Like when you punish your child for doing anything at fucking all so they learn the only right move is to avoid, gaslight, and deflect!

    • Pringles@sopuli.xyz
      link
      fedilink
      English
      arrow-up
      9
      ·
      4 天前

      I am genuinely curious if anyone knows if that has an effect or not. I wouldn’t think so per se, but if the AI interprets it as “double check your initial analysis for errors” it would actually work maybe?

      • Limonene@lemmy.world
        link
        fedilink
        English
        arrow-up
        18
        ·
        4 天前

        For image generation it does. They give negative prompts like “wonky”, “creepy”, and “ugly”, and the image generator evaluates how well the generated image matches those prompts, and produces images opposite those parameters.

        Some poor artist in the training data not only had their work stolen to train the AI, but also had it labeled ugly and wonky.

        • Omgpwnies@lemmy.world
          link
          fedilink
          English
          arrow-up
          4
          ·
          4 天前

          That could also be human training as well. For example, an artist’s work would be used as a “correct” sample, and the machine told to make some other image based on the correct samples, and people would mark results with those tags.

      • Kairos@lemmy.today
        link
        fedilink
        English
        arrow-up
        4
        ·
        4 天前

        It won’t affect the output meaningfully except by rerolling whatever training data ends up being associated with that or whatever. It may end up getting the model to “check” its work which just compares previous output to training data.

      • tempest@lemmy.ca
        link
        fedilink
        English
        arrow-up
        4
        arrow-down
        1
        ·
        4 天前

        Earlier LLMs it helped a bit.

        Now a days the harnesses know to spawn ‘review’ agents which will catch some mistakes but not all.

        • Thorry@feddit.org
          link
          fedilink
          English
          arrow-up
          3
          ·
          4 天前

          You mean it will spawn agents to drive up the token costs and maybe fingers crossed catch some errors?

          • boonhet@lemmy.zip
            link
            fedilink
            English
            arrow-up
            3
            arrow-down
            2
            ·
            4 天前

            I’m on the 20 dollar a month z.ai plan, I’ve yet to hit the 5 hour limit. What token costs?

            Claude I’d usually hit it in an hour at most lol

          • tempest@lemmy.ca
            link
            fedilink
            English
            arrow-up
            1
            ·
            4 天前

            Correct

            I literally have Claude send every edit to another model to check and make sure it isn’t word barfing. Every file edit is a call to another model to make sure that edit doesn’t suck.

            Tokens++

              • tempest@lemmy.ca
                link
                fedilink
                English
                arrow-up
                1
                ·
                3 天前

                It really really depends on what ‘it’ is.

                The LLMs are really very very good at pumping out scripts that can accelerate things like machine learning where it are wrangling data and doing proof of concepts.

                They are also pretty good at basic CRUD feature work which is what the majority of software devs are actually doing.

                The further out of the user’s depth they go the more problematic they can be. They bake many many assumptions in and make hidden decisions that someone without domain expertise cannot easily intuit. Which means someone without experience can get into deep water and not realize and that is where a lot of the problems are.

          • Holytimes@sh.itjust.worksdeleted by creator
            link
            fedilink
            English
            arrow-up
            1
            ·
            3 天前

            That’s not fair, my auto complete knows to always replace duck with duck even when I try to fix it back to duck!