I am not convinced yet of letting agents completely unattended. Watching them work makes review easier for me. If I let the agent just produce some result it needed half an hour (or more) for, it’s very likely so convoluted that I can at best skim over it and then go „yeah yeah ok, it’s probably fine “.
aksdb
joined 2 years ago
I watched two colleagues this week and both had Opus 4.8 1M max thinking. No matter which task. It’s also slow as fuck. I work almost all day with GPT-5.4 low thinking and get good results… but faster and cheaper.
I guess good model selection and promoting will be what sets devs apart in the near future. Once that bubble bursts a bit more and prices increase further that will be an interesting reckoning. Also for companies who basically taunted their employees into tokenmaxxing.
In Germany it did.