The big shift in per-action cost is what always seems to be missing from the conversation. Like, in a lot of my experience the per-request cost is basically negligible compared to the overhead of running the service in general. With LLMs not only do we see massive increases in overhead costs due to the training process necessary to build a usable model, each request that gets sent has a higher cost. This changes the scaling logic in ways that don't appear to be getting priced in or planned for in discussions of the glorious AI technocapital future
you are viewing a single comment's thread
view the rest of the comments
view the rest of the comments
replies: