6 comments

  • rdli 26 minutes ago
    It’s a really good model. Over the past few days, I give Opus some general directives to basically speed up our CI, and telling it I care both about billing minutes and wall clock time. I told it to create a plan after analyzing everything in our CI, run the plan by a Fable subagent, and then focus on low-risk, high-reward changes.

    9 hours later, I had 12 PRs ready to be merged, and the net result is CI time has dropped from ~10 minutes to ~4 minutes, and billing minutes have dropped around 60%. Less than an hour of my attention.

    • chewchewchew 21 minutes ago
      9 hours?!
      • rdli 16 minutes ago
        Yes. It spawned multiple subagents to run different experiments to benchmark a lot of different things, reviewed CI logs from past runs, etc. In the end, there were changes to what/how we cached, various code quality checks, speeding up test runners, and many other things.
  • alansaber 0 minutes ago
    "Don’t ask it to show its reasoning in the reply" “Explain why you chose this approach in three sentences” says it all really
  • danbrooks 10 minutes ago
    Agreed on Opus 5.5 being a great model. It's the first one that I trust for long running (>1 hour) tasks.
  • Handy-Man 36 minutes ago
    Phenomenal model, not sure what they did, but I have been able to do so much with my $20 plan!
    • amelius 8 minutes ago
      I think what makes it great is that they trained it to write harnesses for the code it writes, so it can test stuff even if the supplied code is not complete.
    • 233mhz 31 minutes ago
      If enough people keep saying it I'm sure they will nerf it
  • franze 35 minutes ago
    [flagged]
  • i_love_retros 31 minutes ago
    [flagged]