• dparticiple@sh.itjust.works
    link
    fedilink
    English
    arrow-up
    3
    ·
    edit-2
    5 days ago

    Tried 5.6. Didn’t like it. Specifically, it plunged right into the codebase and started making changes without asking clarifying questions or offering analysis of the problem or any visible explanations, unlike Opus or other frontier models. Disconcerting.

    • ikidd@lemmy.dbzer0.com
      link
      fedilink
      English
      arrow-up
      2
      ·
      5 days ago

      It’s excellent if you can keep it in hand, but holy hell is it tough to get it to stay between the lines. I’ve gone back to 5.5, it’s just way less handholding. And 5.5 gets things done well enough, in about a tenth of the time.

      Sol is good on Ultra for orchestrating if you specify the models you want it to use for subagents. Other than that, it’s too much work.

    • dregan@lemmy.world
      link
      fedilink
      English
      arrow-up
      2
      ·
      edit-2
      5 days ago

      It is miles ahead of Opus when designing scientific experiments. If you want to discover novel techniques or combine known techniques in new ways and you want to be sure that what you are measuring is real and not dependent on your specific dataset, Opus or even Fable just does not compete. Opus is fine for running experiments, but when it comes to interpreting results or recommending next steps, it makes mistakes.