☰
GPUs
Models
Markets
llama-2-7b-chat-hf Inference Pricing — Cheapest Provider by Token Count
Model
Estimate cost for
Input tokens
Output tokens
Price history
Search for a model to see price trends.
7D
30D
90D
Search for a model to see price trends.
Select a model
↻
Search for a model to see available routes.
Price history by provider
Search for a model to compare provider price history.
7D
30D
90D
Search for a model to compare provider price history.
GPU Requirements
✕
Context
In / 1M
Out / 1M
Price history (30d)