Compare LLM Models

Side-by-side comparison of 1 model. Best values highlighted.

Model ID @cf/meta/llama-3.1-8b-instruct-awq
Status ga
Knowledge cutoff —
Context window 8K tokens
Max output —
Parameters —
Open weights No
Pricing (per MTok)
Input $0.12
Output $0.27
Cache read —
Cache write —
Free tier No
Capabilities
Tool Use
Streaming
Prompt Caching
Batch API
Extended Thinking
Structured Outputs
Multi-turn Tool Calling
Agentic Workload Ready
Parallel Tool Calls
Vision Input
Audio Input
Audio Output
Modalities
Input —
Output —
Last updated May 9, 2026