Get the latest tech news

Hetzner is working on LLM Inference


Hetzner has launched an experimental LLM inference API. I tested its Qwen model—and have a few guesses about where the product could go next.

None

Get the Android app

Or read this on Hacker News

Read more on:

Photo of Hetzner

Hetzner

Photo of LLM Inference

LLM Inference

Related news:

News photo

PostgreSQL Benchmark: AWS RDS vs. Self-Hosted on Hetzner (2026)

News photo

DeepSeek open sources DSpark, a new framework to speed up LLM inference by up to 85%

News photo

Real-time LLM Inference on Standard GPUs: 3k tokens/s per request