• tempest@lemmy.ca
    link
    fedilink
    English
    arrow-up
    3
    arrow-down
    1
    ·
    6 days ago

    Earlier LLMs it helped a bit.

    Now a days the harnesses know to spawn ‘review’ agents which will catch some mistakes but not all.

    • Thorry@feddit.org
      link
      fedilink
      English
      arrow-up
      3
      ·
      6 days ago

      You mean it will spawn agents to drive up the token costs and maybe fingers crossed catch some errors?

      • boonhet@lemmy.zip
        link
        fedilink
        English
        arrow-up
        3
        arrow-down
        2
        ·
        6 days ago

        I’m on the 20 dollar a month z.ai plan, I’ve yet to hit the 5 hour limit. What token costs?

        Claude I’d usually hit it in an hour at most lol

      • tempest@lemmy.ca
        link
        fedilink
        English
        arrow-up
        1
        ·
        edit-2
        6 days ago

        Correct

        I literally have Claude send every edit to another model to check and make sure it isn’t word barfing. Every file edit is a call to another model to make sure that edit doesn’t suck.

        Tokens++

          • tempest@lemmy.ca
            link
            fedilink
            English
            arrow-up
            1
            ·
            5 days ago

            It really really depends on what ‘it’ is.

            The LLMs are really very very good at pumping out scripts that can accelerate things like machine learning where it are wrangling data and doing proof of concepts.

            They are also pretty good at basic CRUD feature work which is what the majority of software devs are actually doing.

            The further out of the user’s depth they go the more problematic they can be. They bake many many assumptions in and make hidden decisions that someone without domain expertise cannot easily intuit. Which means someone without experience can get into deep water and not realize and that is where a lot of the problems are.