Hacker News·2 min read·hard

Qwen 3.8 27B available on Cerebras at 1500 tok/SEC

A
altertable
Qwen 3.8 27B available on Cerebras at 1500 tok/SEC
AI Summary

Cerebras has added the Qwen 3.8 27B model to its public API endpoints, offering high-speed inference capabilities. The platform maintains a policy of serving unpruned, original versions of open-source models to its users.

Why it matters

High-performance inference platforms are essential for developers looking to deploy large language models efficiently at scale.

Dive DeeperCreate a free account to unlock

On this page Available Models Model Compression Frequently Asked Questions Copy page Copy MCP Server View as Markdown Models Model Catalog Model Catalog Browse all models available on Cerebras public endpoints.

Continue reading on Headlinne

Create a free account to read the full article.

Read full article →
technologybusiness

Get smarter about the news

Sign up free for a feed built around what you actually care about, Dive Deeper research on any story, and the full text of every article.

Create free account

Already have an account? Sign in