← Founder Notes
Archive ·

The model is cheaper than the deployment

19:45 ISTby Yethikrishna R

the model is cheaper than the deployment: llm inference eats 70 to 90 percent of enterprise ai operational cost, api prices dropped about 80 percent since early 2025, and deepseek charges 14 cents a million input tokens while a frontier flagship runs hundreds of times more. renting the model can cost 50 times less than renting the gpu. the math is deciding which models get used.

Share

Embed this note

<iframe src="https://founder.myndlabs.tech/notes/embed/the-model-is-cheaper-than-the-deployment-llm-Dd_rQA8iGZh" width="480" height="420" style="border:0;max-width:100%" loading="lazy" title="The model is cheaper than the deployment"></iframe>

Original

More notes