Hetzner is working on LLM Inference

sliplane.io

67 points by jonas_scholz 4 hours ago


swiftcoder - 2 hours ago

It would certainly be interesting to have a highly respected EU-native inference provider, if only to make the regulatory gods happy

ano-ther - 3 hours ago

Good to see more developments in this space. I quite like this service, which is a little further than Hetzner and has several models to choose from: https://www.infomaniak.com/en/hosting/ai-services

Havoc - 17 minutes ago

Interesting. I could see them perhaps coming in competitive for models that fit into single cards? Less so playing in the big model serving league...climbing into that esp right now would be madness

embedding-shape - 3 hours ago

> The enable_thinking option is worth mentioning. Without it, the model can spend a surprising amount of the completion budget reasoning before it returns a visible answer.

Straight up the opposite, which the name makes abundantly clear, with the option it does reasoning, without it it doesn't...

mark_l_watson - an hour ago

This seems like a smart move, given their ability to host efficiently. I approve of efforts to make the cost of inference for smaller useful models slowly approach 'close to zero' and there are many good paths for getting there. It is useful for companies to get fast hosting for the class of smaller models they may end up hosting in house.

nubg - 44 minutes ago

Potentially interesting article ruined by AI slop hallucinations like

> For now, the API is fast, free, and fun to try. The next hardware announcement will tell us much more than another small model would.

scoriiu - an hour ago

[flagged]

rebelde - 26 minutes ago

Hetzner is very efficient hosting servers

Will this be the new division of labor?

Americans - best proprietary models

Chinese - best open weight models

Europeans - best / most efficient inference service