logoalt Hacker News

Hetzner is working on LLM Inference

122 pointsby jonas_scholztoday at 9:24 AM50 commentsview on HN

Comments

NetOpWibbytoday at 2:55 PM

This is interesting because I thought Hetzner was anti-crypto? LLMs aren't the same but they're often lumped in with crypto as "things no one wants."

swiftcodertoday at 11:30 AM

It would certainly be interesting to have a highly respected EU-native inference provider, if only to make the regulatory gods happy

show 7 replies
ano-thertoday at 10:20 AM

Good to see more developments in this space. I quite like this service, which is a little further than Hetzner and has several models to choose from: https://www.infomaniak.com/en/hosting/ai-services

show 3 replies
perelintoday at 2:17 PM

Whats definitely missing: a solid (non Mistral) GDPR compliant coding plan / subscription. All offerings are either US or China based. With the newest open weights models this became really interesting imo.

mark_l_watsontoday at 11:35 AM

This seems like a smart move, given their ability to host efficiently. I approve of efforts to make the cost of inference for smaller useful models slowly approach 'close to zero' and there are many good paths for getting there. It is useful for companies to get fast hosting for the class of smaller models they may end up hosting in house.

show 2 replies
embedding-shapetoday at 10:18 AM

> The enable_thinking option is worth mentioning. Without it, the model can spend a surprising amount of the completion budget reasoning before it returns a visible answer.

Straight up the opposite, which the name makes abundantly clear, with the option it does reasoning, without it it doesn't...

show 1 reply
_pdp_today at 2:42 PM

I mean yah... host glm and kimi and I am game.

rebeldetoday at 12:35 PM

Hetzner is very efficient hosting servers

Will this be the new division of labor?

Americans - best proprietary models

Chinese - best open weight models

Europeans - best / most efficient inference service

show 2 replies
Havoctoday at 12:44 PM

Interesting. I could see them perhaps coming in competitive for models that fit into single cards? Less so playing in the big model serving league...climbing into that esp right now would be madness

show 1 reply
nubgtoday at 12:16 PM

Potentially interesting article ruined by AI slop hallucinations like

> For now, the API is fast, free, and fun to try. The next hardware announcement will tell us much more than another small model would.

show 1 reply
dk970today at 2:36 PM

[flagged]

scoriiutoday at 12:14 PM

[flagged]