Fast inference

Cerebras in AGNT

Wafer-scale inference for open models. Every AGNT workflow is provider-portable, so the same agents run here or anywhere else you point them.

Why teams pick it

Models

  • Llama and Qwen at speed

Strengths

  • thousands of tokens/sec
  • great for iterative agents
  • generous free tier

In AGNT

Agents, workflows and tools use Cerebras like any other provider — same nodes, same receipts, swap models per step if you want.

Connect Cerebras in two minutes

  1. Download AGNT Community Core — free, local-first, no account needed to run.
  2. Open Settings → Providers, choose Cerebras, and paste your API key. Connect by pasting an API key into AGNT’s vault — stored encrypted on your machine, never uploaded.
  3. Give an agent a job — or install one from the marketplace — and watch the first receipt come back.

Give AI a job. Get the proof.