High-Speed AI Chatbot with Cerebras GPT-OSS-120B

n8n workflow for ultra-fast AI chats using Cerebras Inference and OpenAI's GPT-OSS-120B model, delivering thousands of tokens/sec and <0.5s latency.

This n8n workflow enables seamless, high-performance AI chat integration via Cerebras' inference platform, powering OpenAI's open-source GPT-OSS-120B model. It processes incoming chat messages through a simple 4-node chain: trigger on message receipt, set API key, query the Cerebras endpoint with customizable parameters (temperature, max tokens, top-p, reasoning effort), and return the formatted response. Ideal for developers building responsive AI apps without infrastructure hassles. Key benef
Platform
n8n
Category
Internet of Things
Price
$12.99
Creator
BestWorkflows

How to import this workflow into n8n

  1. 1Purchase or download the workflow to get the n8n workflow JSON file.
  2. 2In your n8n instance, open Workflows and choose "Import from File" (or paste the JSON with Ctrl+V on the canvas).
  3. 3Open each node marked with a credential warning and connect your own accounts and API keys.
  4. 4Run the workflow once manually to verify the data flow, then toggle it to Active.

Related Internet of Things workflows

More from BestWorkflows

Need this deployed? We'll set it up for you.

Our automation experts deploy this workflow in your stack, connect your accounts, and verify it works — or build a custom solution from scratch.