• IndustryStandard@lemmy.world
    link
    fedilink
    arrow-up
    5
    ·
    3 hours ago

    That is why you ask the LLM to find bugs and other instances where the first LLM cheated. The checking LLM will try to find abuse with all its might.

    • Avicenna@programming.dev
      link
      fedilink
      arrow-up
      2
      ·
      2 hours ago

      It really turns into a tug-of-war, see my answer below. Despite it being quite good at debugging imo still requires a human in the loop for it not run in circles after a “bug” that is almost never relevant in practice but whose fix complicates the code.

      Try asking independent LLMs to find bugs in a code consecutively (fixing the founds bugs in between) and even after the rightfully major bugs are resolved, the independent newer instances will keep suggesting fixes (those which sometimes undoes its previous “fixes”).