Compare LLM Models

Side-by-side comparison of 1 model. Best values highlighted.

Model ID llama-3.2-3b-instruct
Status ga
Knowledge cutoff 2023-12-31
Context window 80K tokens
Max output —
Parameters —
Open weights No
Pricing (per MTok)
Input $0.051
Output $0.34
Cache read —
Cache write —
Free tier No
Capabilities
Tool Use
Streaming
Prompt Caching
Batch API
Extended Thinking
Structured Outputs
Multi-turn Tool Calling
Agentic Workload Ready
Parallel Tool Calls
Vision Input
Audio Input
Audio Output
Modalities
Input text
Output text
Last updated May 9, 2026