← Founder Notes
Archive

The model is cheaper than the deployment

Yethikrishna ROriginal on Threads

the model is cheaper than the deployment: llm inference eats 70 to 90 percent of enterprise ai operational cost, api prices dropped about 80 percent since early 2025, and deepseek charges 14 cents a million input tokens while a frontier flagship runs hundreds of times more. renting the model can cost 50 times less than renting the gpu.

the math is deciding which models get used.

Provenance

The note above is reproduced unedited from the original post, first published on Threads on 2 October 2026 at 19:45 IST.

View the original post
Embed this note
<iframe src="https://founder.myndlabs.tech/notes/embed/the-model-is-cheaper-than-the-deployment-llm-Dd_rQA8iGZh" width="480" height="420" style="border:0;max-width:100%" loading="lazy" title="The model is cheaper than the deployment"></iframe>

More notes