Llama 3 1 8B is a Cerebras model tracked in Studio. It supports a 33k token context window. Pricing starts at $0.1/1M input tokens and $0.1/1M output tokens. Key capabilities include Tool choice, Structured outputs. Best for cost-sensitive automations, background tasks, and high-volume workloads.
Best forBest for cost-sensitive automations, background tasks, and high-volume workloads.
TemperatureNot configurable
Reasoning effortNot supported
VerbosityNot supported
Thinking levelsNot supported
Structured outputsSupported
Tool choiceSupported
Computer useNot supported
Deep researchNot supported
Memory supportSupported
Max output tokensNot published
Frequently asked questions
Llama 3 1 8B is a Cerebras model available in Studio. Llama 3 1 8B is a Cerebras model tracked in Studio. It supports a 33k token context window. Pricing starts at $0.1/1M input tokens and $0.1/1M output tokens. Key capabilities include Tool choice, Structured outputs.
Llama 3 1 8B is listed at $0.1/1M input tokens, and $0.1/1M output tokens.
Llama 3 1 8B supports a context window of 33k tokens in Studio. In an agent, this determines how much conversation history, tool outputs, and retrieved documents the model can hold in a single call.
Llama 3 1 8B supports the following capabilities in Studio: Tool choice, Structured outputs.
Best for cost-sensitive automations, background tasks, and high-volume workloads. When used in a Studio workflow, it can be selected in any Agent block from the model picker.