01Overview
Deepinfra serves meta-llama/Llama-3.3-70B-Instruct-Turbo at $0.100 per million input tokens and $0.320 per million output. Context window: 131K tokens. No lender is online for it at the moment.
deepinfra No lender online
Deepinfra serves meta-llama/Llama-3.3-70B-Instruct-Turbo at $0.100 per million input tokens and $0.320 per million output. Context window: 131K tokens. No lender is online for it at the moment.
Input /Mtok
Output /Mtok
Context
Lenders online
Deepinfra serves meta-llama/Llama-3.3-70B-Instruct-Turbo at $0.100 per million input tokens and $0.320 per million output. Context window: 131K tokens. No lender is online for it at the moment.
Not supported: image input, extended thinking.
No lender is online for this model at the moment.
See what is available now, or lend capacity for it yourself.
This model has not been served through Aile yet.
Any OpenAI- or Anthropic-compatible SDK works. Use this exact model id:
curl https://aile.sh/v1/chat/completions \ -H "Authorization: Bearer $AILE_KEY" \ -H "Content-Type: application/json" \ -d '{ "model": "deepinfra/meta-llama/Llama-3.3-70B-Instruct-Turbo", "messages": [{"role": "user", "content": "Hello"}] }'
Other providers serve meta-llama/Llama-3.3-70B-Instruct-Turbo too, at their own rates — compare every provider for meta-llama/Llama-3.3-70B-Instruct-Turbo.
Other models from Deepinfra:
Similar published rates elsewhere: