> The fact that tokens aren't free will limit the number of unserious requests.
LLM costs are already pretty cheap (regardless of whether someone believes they'll continue falling).
For a recent example, see a relatively powerful LLM like DeepSeek Flash, where you can get a million tokens for $0.20. And if someone is OK with their task being batched (instead of executing it right now), that can lower prices too.
LLM costs are already pretty cheap (regardless of whether someone believes they'll continue falling).
For a recent example, see a relatively powerful LLM like DeepSeek Flash, where you can get a million tokens for $0.20. And if someone is OK with their task being batched (instead of executing it right now), that can lower prices too.
Inference price trends in previous years: https://epoch.ai/data-insights/llm-inference-price-trends