qwen/qwen3.8-27b
groq/qwen/qwen3.8-27b
Context
131K
Max output
16K
Input / 1M
$0.80
Output / 1M
$4.00
Modality
Chat
Cutoff
unknown
About this model
Qwen3.8-27B served on GroqCloud (Preview) — Alibaba Cloud's Qwen3.8 27B-parameter model running at roughly 450 tokens/sec on Groq's LPU infrastructure. Preview models are for evaluation only and may be discontinued at short notice.
Best suited for
- General-purpose chat
- Reasoning tasks
- Tool use / agentic workflows
Capabilities
Files
Accepts file attachments — PDFs, transcripts, spreadsheets.
Tools
Native function calling, so agents can invoke your endpoints.
System prompt
Honours a dedicated system role, separate from the user turn.
Reasoning
Emits a separate thinking pass before the answer.
Supported parameters
max_tokensMax Tokens LimitSpecifies the maximum number of text units (tokens) allowed in a response, limiting its length.
toolsLists tool definitions or capabilities available to the model.
tool_choiceDecides whether to use tools or just the model for generating responses.
response_typeDefines the format or type of the generated response.
parallel_tool_callsEnables parallel execution of tools, allowing multiple tools to run simultaneously.
reasoningControls the level of reasoning used by the model.
streamSends the response in real-time as it's being generated.
service_tierGroqCloud preview-model serving tier (evaluation only, may be discontinued at short notice).