High-Speed AI Chatbot with Cerebras GPT-OSS-120B
n8n workflow for ultra-fast AI chats using Cerebras Inference and OpenAI's GPT-OSS-120B model, delivering thousands of tokens/sec and <0.5s latency.
This n8n workflow enables seamless, high-performance AI chat integration via Cerebras' inference platform, powering OpenAI's open-source GPT-OSS-120B model. It processes incoming chat messages through a simple 4-node chain: trigger on message receipt, set API key, query the Cerebras endpoint with customizable parameters (temperature, max tokens, top-p, reasoning effort), and return the formatted response. Ideal for developers building responsive AI apps without infrastructure hassles.
Key benef
- Platform
- n8n
- Category
- Internet of Things
- Price
- $12.99
- Creator
- BestWorkflows
- AI
- Chatbot
- Cerebras
- GPT
- Inference
- OpenAI
- High-Speed
- Automation
- n8n
- Chat Completions
How to import this workflow into n8n
- 1Purchase or download the workflow to get the n8n workflow JSON file.
- 2In your n8n instance, open Workflows and choose "Import from File" (or paste the JSON with Ctrl+V on the canvas).
- 3Open each node marked with a credential warning and connect your own accounts and API keys.
- 4Run the workflow once manually to verify the data flow, then toggle it to Active.
Related Internet of Things workflows
- Automated Weather Alerts and Analysis Using AI and Home Assistant$14.99
- AI-Driven Handbook Generator with Multi-Agent Orchestration$24.99
- Vehicle Telematics Analyzer for Efficient Device Management$24.99
- Automate Solar Energy Monitoring and Alerts with Gmail, Google Sheets, and Slack$9.99
- Automate Network Disconnection Alerts with Omada, Gmail, and Pushover$14.99
- Birthday and Ephemeris Notification (Google Contact, Telegram & Home Assistant)$14.99
More from BestWorkflows
Need this deployed? We'll set it up for you.
Our automation experts deploy this workflow in your stack, connect your accounts, and verify it works — or build a custom solution from scratch.