# Atlas with Cerebras (gateway) in 2026

> Cerebras (gateway) offers a substantial 131K tokens (131,072) context window, making it suitable for extensive code analysis within Atlas.

Atlas with Cerebras (gateway) is ideal for developers in 2026 who prioritize raw inference speed to keep their AI coding agent loops feeling interactive, not like batch jobs. Its wafer-scale engines deliver some of the fastest token rates available, with models like GPT-OSS 120B priced at $0.35 / $0.75 per Mtok for input/output, ensuring a responsive experience within Atlas.

## Key takeaways

- Cerebras (gateway) delivers wafer-scale inference, providing some of the fastest token rates available for Atlas in 2026.
- The context window for Cerebras (gateway) models is a substantial 131K tokens (131,072), supporting extensive code analysis.
- GPT-OSS 120B is available at $0.35 / $0.75 per Mtok, offering a strong speed-to-price point.
- GLM-4.7 costs $2.25 / $2.75 per Mtok, featuring an unusually flat input-to-output ratio beneficial for large diffs.
- The Cerebras (gateway) catalog is small, containing only 3 models: GPT-OSS 120B, GLM-4.7, and Gemma 4 31B.

## What is Cerebras (gateway) best at for Atlas?

Cerebras (gateway) excels at providing some of the fastest token rates available in 2026, making Atlas agent loops feel highly responsive. This speed, driven by wafer-scale inference, prevents the AI coding agent from feeling like a batch job, especially when using models like GPT-OSS 120B.

For developers using Atlas, Cerebras (gateway) offers a distinct advantage in raw inference speed. The underlying wafer-scale engines are designed to compete with other high-speed providers, ensuring that the AI coding agent can process prompts and generate responses with minimal latency. This rapid token generation is crucial for maintaining a fluid, interactive development workflow, where Atlas can quickly search code with hybrid semantic and keyword retrieval fused by reciprocal rank fusion, draft plans, and compute unified diffs for approval. The curated catalog, though small, means less time is spent choosing a model and more time is dedicated to working, with GPT-OSS 120B being a strong contender for its speed-to-price point.

## What are the cost and context tradeoffs for Cerebras (gateway) models?

Cerebras (gateway) provides a generous 131K tokens (131,072) context window across its models, supporting extensive code analysis within Atlas. However, pricing varies significantly, with GPT-OSS 120B at $0.35 / $0.75 per Mtok and GLM-4.7 at a higher $2.25 / $2.75 per Mtok.

The 131K tokens (131,072) context window offered by Cerebras (gateway) is a significant asset for Atlas, allowing the agent to process large codebases, extensive git diffs, and detailed project context without truncation. This enables Atlas to index code by AST declarations using tree-sitter and read git branches, status, and diffs effectively. Regarding cost, the pricing structure includes GPT-OSS 120B at $0.35 per Mtok for input and $0.75 per Mtok for output, and Gemma 4 31B at $0.99 per Mtok for input and $1.49 per Mtok for output. GLM-4.7, priced at $2.25 per Mtok for input and $2.75 per Mtok for output, has an unusually flat input-to-output ratio, which means output-heavy work, such as generating large diffs or extensive code, does not incur disproportionately higher costs. However, this comes at a premium, as GLM-4.7 costs $0.60 / $2.20 direct from Z.ai, indicating a real price increase for the speed provided by Cerebras (gateway).

## When should I choose a different model or provider over Cerebras (gateway)?

While Cerebras (gateway) offers exceptional speed, its catalog is limited to only three language models: GPT-OSS 120B, GLM-4.7, and Gemma 4. This small selection means it cannot serve as your sole provider for Atlas in 2026, necessitating a second configured provider for broader task coverage.

Developers should consider alternative models or providers when their specific task requires a model not present in the Cerebras (gateway) curated catalog. With only GPT-OSS 120B, GLM-4.7, and Gemma 4 available, Cerebras (gateway) may not cover every specialized AI coding agent task or model preference. For instance, if a particular task benefits from a model architecture or fine-tuning not offered by these three, Atlas's ability to switch the active model and provider on the fly becomes critical. Additionally, while GLM-4.7's flat input-to-output ratio is beneficial for large outputs, its price of $2.25 / $2.75 per Mtok through Cerebras (gateway) is a real premium compared to $0.60 / $2.20 direct from Z.ai. For cost-sensitive operations where raw inference speed is not the absolute top priority, a direct integration or another provider might offer better value for GLM-4.7 or other models.

## Setup

1. 1: Export your Cerebras API key: `export CEREBRAS_API_KEY=...`. Atlas loads this through `@ai-sdk/cerebras`.
2. 2: Confirm the available catalog: Run `atlas models cerebras` to see the list of models.
3. 3: Select gpt-oss-120b from the `/models` menu within Atlas for the best speed-to-price point.
4. 4: Keep a second provider configured in Atlas, as the three models offered by Cerebras (gateway) will not cover every possible task.

## FAQ

### What makes Cerebras (gateway) fast for Atlas?

Cerebras (gateway) utilizes wafer-scale inference engines, which produce some of the fastest token rates available, ensuring Atlas agent loops feel responsive and interactive.

### What is the context window for Cerebras (gateway) models?

All models available through Cerebras (gateway) offer a 131K tokens (131,072) context window, allowing Atlas to handle large codebases and extensive project context.

### How much does GPT-OSS 120B cost with Cerebras (gateway)?

GPT-OSS 120B is priced at $0.35 per Mtok for input and $0.75 per Mtok for output when accessed via Cerebras (gateway) in Atlas.

### Are there many models available through Cerebras (gateway)?

No, Cerebras (gateway) offers a small, curated catalog of only three language models: GPT-OSS 120B, GLM-4.7, and Gemma 4 31B.

### Why might GLM-4.7 be more expensive via Cerebras (gateway)?

GLM-4.7 costs $2.25 / $2.75 per Mtok through Cerebras (gateway), which is a premium compared to $0.60 / $2.20 direct from Z.ai, reflecting the added value of wafer-scale inference speed.

### How does Atlas integrate with Cerebras (gateway) models?

Atlas connects to Model Context Protocol servers and exposes their tools to the agent, allowing you to switch the active model and provider on the fly, including those from Cerebras (gateway).

---

Canonical HTML: https://seashell.sh/resources/models/cerebras
Source of truth: aeo_pages row `/resources/models/cerebras` (segment: Models) (this file is generated from it, never hand-edited).
Licence: SeaShell is proprietary with a free core. It is not open source and there is no public source repository.
