Article may be outdated

This article is 69 days old. Some details may have changed since publication.

Hacker News·3 min read·medium

Hetzner is working on LLM Inference

J
jonas_scholz
Hetzner is working on LLM Inference
✦AI Summary

Cloud provider Hetzner has launched an experimental LLM inference service that is compatible with the OpenAI API. The service currently offers access to the Qwen 35B model for testing purposes without production guarantees.

Why it matters

Hetzner's entry into the AI inference market provides a lower-cost alternative for developers to test and run large language models.

✦Dive DeeperCreate a free account to unlock

Hetzner is experimenting with LLM inference.

That is not a sentence I expected to write, but I think it is pretty interesting :)

Before anyone moves their production AI workloads to Hetzner: this is very much an experiment . There is no billing, no SLA, no production guarantee, and currently only one model. Hetzner says it wants to learn whether people actually want this, how the system scales, which features matter, and what kind of load it can handle.

So this is not a finished product launch. It is Hetzner putting something early in front of users and seeing what happens. I really like that approach.

Hetzner Inference is an OpenAI-compatible API running on Hetzner's own infrastructure. You create an API token in the Experiments dashboard, point an OpenAI client at Hetzner's base URL, and use it like most other inference APIs.

Continue reading on Headlinne

Create a free account to read the full article.

Read full article →
technologyai
✦

Get smarter about the news

Sign up free for a feed built around what you actually care about, Dive Deeper research on any story, and the full text of every article.

Create free account

Already have an account? Sign in