Like a printing service: cheaper to give them paper (input) than to buy printed pages (output).
LLM pricing: $/1M input tokens and $/1M output tokens. Input is always cheaper than output. Gemini Flash is the cheapest; GPT-4o and Claude Sonnet are premium.
> Cheapest: gemini-1.5-flash ($0.075 input) > Most expensive: gpt-4o ($5.00 input) > Ratio: 67× price difference > Output always > input in price
Like a printing service: cheaper to give them paper (input) than to buy printed pages (output).
LLM pricing: $/1M input tokens and $/1M output tokens. Input is always cheaper than output. Gemini Flash is the cheapest; GPT-4o and Claude Sonnet are premium.
> Cheapest: gemini-1.5-flash ($0.075 input) > Most expensive: gpt-4o ($5.00 input) > Ratio: 67× price difference > Output always > input in price
Sign in to cast your vote
Sign in to share your feedback and join the discussion.