meta-llama/llama-3.2-1b-instruct API price comparison
Llama 3.2 1B is a 1-billion-parameter language model focused on efficiently performing natural language tasks, such as summarization, dialogue, and multilingual text analysis. Its smaller size allows it to operate efficiently in low-resource environments while maintaining strong task performance. Supporting eight core languages and fine-tunable for more, Llama 1.3B is ideal for businesses or developers seeking lightweight yet powerful AI solutions that can operate in diverse multilingual settings without the high computational demand of larger models.
Cheapest overall
Qianduoduo
0.80 USD/M tokens combined per 1M tokens
Lowest input
Qianduoduo
0.40 USD/M tokens input per 1M tokens
Lowest output
Qianduoduo
0.40 USD/M tokens output per 1M tokens
Last refreshed
Jul 29, 2026
1 listed platform
Best listed option
meta-llama/llama-3.2-1b-instruct is currently cheapest overall on Qianduoduo, with input at 0.40 USD/M tokens and output at 0.40 USD/M tokens per million tokens.
Pricing insight
Across 1 listed routes, meta-llama/llama-3.2-1b-instruct is cheapest overall on Qianduoduo at 0.80 USD/M tokens.
-
Lowest total cost
Qianduoduo
Best first check for balanced coding-agent sessions: 0.80 USD/M tokens.
-
Prompt-heavy work
Qianduoduo
Useful for large repositories and long context reads: 0.40 USD/M tokens input.
-
Response-heavy work
Qianduoduo
Useful for long answers, edits, and agent loops: 0.40 USD/M tokens output.
Median combined price
0.80 USD/M tokens
Baseline for judging whether the cheapest route is unusually low.
Cache price coverage
0%
0 of 1 routes publish cache read or write prices.
Fresh rows
0/1
Routes crawled in the last 72 hours.
Use case fit
meta-llama/llama-3.2-1b-instruct from Other can be compared by input, output, and combined per-million-token cost to select the most economical API route.
Platform price comparison
| Platform | Ratio | Input | Output | Total |
|---|---|---|---|---|
| Qianduoduo | 1.00x | 0.40 USD/M tokens | 0.40 USD/M tokens | 0.80 USD/M tokens |
FAQ
Where is meta-llama/llama-3.2-1b-instruct cheapest?
Qianduoduo has the lowest listed combined price at 0.80 USD/M tokens per 1M tokens.
What do input and output prices mean?
Input is charged for prompt tokens you send. Output is charged for tokens generated by the model.
How often are prices refreshed?
Prices are crawler-backed and this page was last refreshed Jul 29, 2026. Always confirm before large purchases.