Llama 3.3 70B Instruct
Provider: meta · served via Vertex
Listed as deprecated on models.dev as of 2026-08-13.
Sticker prices & cost-per-token
List rates as published by Vertex, before any volume or enterprise discounts.
$0.72
$0.72
Get this data free
Want to stay updated on price changes on this model? Pull this data free (no API key required) here:
curl "https://coolhandlabs.com/api/v2/inference_apis?q%5Bmodel_eq%5D=llama-3.3-70b-instruct-maas&q%5Bsource_api_eq%5D=vertex"
Pricing History
| Effective | Price |
|---|---|
| Aug 16, 2026 – present | Input cost: $0.72 / 1M, Output cost: $0.72 / 1M, Cached input cost: $0.00 / 1M, Cache creation cost: $0.00 / 1M, 5-minute cache write cost: $0.00 / 1M, 1-hour cache write cost: $0.00 / 1M, Reasoning output cost: $0.00 / 1M, Cache storage cost: $0.00 / 1M/hr |
History
No automated scan changes recorded yet.
| Date | Changed by | What changed |
|---|---|---|
| Aug 16, 2026 | Unknown | Initial record created |
Source
https://models.dev/api.json (models.dev provider 'google-vertex', bulk import 2026-08-13)Comparable models
No comparable models identified yet.
Reliability
No reliability incidents recorded in the last 90 days.
Frequently asked questions
Llama 3.3 70B Instruct (meta via Vertex) costs $0.72 per 1M input tokens and $0.72 per 1M output tokens, as last published by Vertex.
Yes — meta deprecated Llama 3.3 70B Instruct on August 13, 2026. Listed as deprecated on models.dev as of 2026-08-13.
meta retired Llama 3.3 70B Instruct on August 13, 2026.
Since Llama 3.3 70B Instruct is deprecated, meta customers are moving to Llama-3.3-70B-Instruct, Llama 3.1 70B Instruct, and Llama 3.1 8B Instruct.
Call GET https://coolhandlabs.com/api/v2/inference_apis?q[source_api_eq]=vertex&q[model_eq]=llama-3.3-70b-instruct-maas — free, no API key required, and scoped to just this model since source_api and model together are unique. See https://coolhandlabs.com/docs for the full REST API reference.
Coolhand also tracks Llama-3.3-70B-Instruct, Llama 3.1 70B Instruct, Llama 3.1 8B Instruct, Llama 3.3 70B Instruct, Llama 4 Maverick 17B Instruct, and Llama 4 Scout 17B Instruct from meta. Some may already be deprecated, so check each model's status alongside its current pricing.
| Model | Input cost | Output cost | Status |
|---|---|---|---|
| Llama-3.3-70B-Instruct via Azure | $0.71/1M $0.0000007100 | $0.71/1M $0.0000007100 | Active |
| Llama 3.1 70B Instruct via Bedrock | $0.72/1M $0.0000007200 | $0.72/1M $0.0000007200 | Active |
| Llama 3.1 8B Instruct via Bedrock | $0.22/1M $0.0000002200 | $0.22/1M $0.0000002200 | Active |
| Llama 3.3 70B Instruct via Bedrock | $0.72/1M $0.0000007200 | $0.72/1M $0.0000007200 | Active |
| Llama 4 Maverick 17B Instruct via Bedrock | $0.24/1M $0.0000002400 | $0.97/1M $0.0000009700 | Active |
| Llama 4 Scout 17B Instruct via Bedrock | $0.17/1M $0.0000001700 | $0.66/1M $0.0000006600 | Active |
Track LLM pricing and deprecations automatically.
Start free