- You describe the breakdown in terms of time but it's more accurately a function of reasoning complexity.
- You seem to assume that no intermediate evaluation is possible.
- Often it is (e.g. the build breaks or tests start failing), allowing for course correction. There's definitely a cost to that but it can still be cost effective if the accuracy is "good enough" and the price difference significant.
- There are numerous tasks that don't require Fable or GPT5.6 level reasoning to improve efficiency by an order of magnitude.
- You describe the breakdown in terms of time but it's more accurately a function of reasoning complexity.
- You seem to assume that no intermediate evaluation is possible.
- Often it is (e.g. the build breaks or tests start failing), allowing for course correction. There's definitely a cost to that but it can still be cost effective if the accuracy is "good enough" and the price difference significant.
- There are numerous tasks that don't require Fable or GPT5.6 level reasoning to improve efficiency by an order of magnitude.