Llama 3.2 1B is a 1-billion-parameter language model focused on efficiently performing natural language tasks, such as summarization, dialogue, and multilingual text analysis.
Prices update automatically — checked daily against provider list prices.
Independent benchmark scores for Llama 3.2 1B Instruct, measured by Artificial Analysis. Higher is better.
Llama 3.2 1B Instruct is an open-weight AI model by Meta. You can download and self-host it for free; the prices below are hosted-API list prices, tracked daily, for when you prefer convenience over self-hosting.
Llama 3.2 1B Instruct is an AI language model from Meta. It is open-weight: you can download it and run it on your own hardware, for free. It scores 1 on the Artificial Analysis intelligence index.
The weights are free and open — you can self-host Llama 3.2 1B Instruct and pay nothing per token. If you prefer a hosted API, list prices are $0.03 per million input tokens and $0.2 per million output tokens.
Independent benchmarks from Artificial Analysis give it GPQA 19.6%, MMLU-Pro 20%, Humanity's Last Exam 5.5%, Long Context Reasoning 6%, LiveCodeBench 1.9%, SciCode 1.7%, MATH-500 14%, AIME 0%, AIME 2025 0%, IFBench 22.8%, τ²-Bench 0%, Terminal-Bench Hard 0%.
Yes. Llama 3.2 1B Instruct has open weights, so you can download it and run it on your own GPU or server with tools like Ollama, vLLM or llama.cpp — with no per-token cost.