Yet this is my problem, A tool that "waste" tokens, is slow (with some tasks taking 30~40 minutes) and generating a result, while a model running locally can do the 70% in 5min. It's a huge difference.
And don't forget, at the end of the day, I would still need to look into the code and fix some stuff myself.
And if compared with Deepseek (which had fit best on my use case), it's even faster like 1~3 minutes, to get on the 90% of the result, on a fraction of the cost. This is the points that I'm focused, and a enterprise should either understand that the market have changed, or pay the price. And looks like we will pay the price.
For my use case, didn't work. But I was in a more spec driven than anything else. So maybe with a shorter spec file would have a better result.