- Integrations
- /
- Cerebras
- /
- Actions
- /
- Cerebras Chat Completion
ActionCerebrasUpdated May 2026
How do I call Cerebras for inference?
Short answer: Drop the "Cerebras → Cerebras Chat Completion" action anywhere in your workflow, map the inputs from upstream nodes, and publish.
Inputs
The fields this action accepts.
Every field can be mapped from an upstream trigger, AI step, table row, or hard-coded literal.
| Field | Type | Required | Description |
|---|---|---|---|
Model model | options | Required | Which model to use |
User Message message | string | Required | Message sent to the model as user role |
System Prompt system_prompt | string | Optional | Optional system instructions |
Temperature temperature | string | Optional | 0-2, higher = more random |
Max Tokens max_tokens | string | Optional | Maximum tokens to generate |
Sample request
{"model": "{{trigger.model}}","message": "Your prompt","system_prompt": "e.g. You are a helpful assistant","temperature": "0.7","max_tokens": "1024"}
Returns
{"id": "chatcmpl_abc","model": "llama-3.3-70b","usage": {"total_tokens": 60,"prompt_tokens": 10,"completion_tokens": 50},"choices": [{"message": {"role": "assistant","content": "Sample"},"finish_reason": "stop"}]}
Use these fields in downstream nodes for routing, logging, or error handling.
Triggered by
Apps that pair well as the trigger for Cerebras Chat Completion.
Any of these apps can fire this action as part of a workflow.
FAQ
Questions about Cerebras Chat Completion.
What does the Cerebras Chat Completion action do in Cerebras?
Runs chat completion against Cerebras's custom-hardware-served Llama models. Extreme speed (1000-2000+ tokens/sec) is the value — for multi-step agent workflows, the latency win stacks meaningfully across calls.
What inputs does Cerebras Chat Completion require?
Required: Model, User Message. Every input accepts a static value or a variable from any upstream node in your workflow.
Can I use dynamic inputs from earlier workflow nodes?
Yes. Any field on this action can pull values from upstream nodes, whether that's a form response, a trigger payload, an AI output, or a lookup result.
What happens if Cerebras returns an error?
The workflow pauses on the failed node, the error message is captured in the run log, and you can retry the run with one click. Auto-retry policies are configurable per workflow with exponential backoff up to 5 attempts.
Does Cerebras Chat Completion support batch operations?
Yes. Run Cerebras Chat Completion inside a Loop node to process arrays. Tiny Command handles Cerebras's rate limits automatically so you don't have to throttle manually.
More actions
Other Cerebras actions.
Send cerebras chat completion from your workflows.
Triggered by anything in the catalog. Free tier available. No credit card.