Comments

Hacker News

Excited to see the price of every other product they offer triple for no reason.

by danlitt

Whats definitely missing: a solid (non Mistral) GDPR compliant coding plan / subscription. All offerings are either US or China based. With the newest open weights models this became really interesting imo.

by perelin

> The enable_thinking option is worth mentioning. Without it, the model can spend a surprising amount of the completion budget reasoning before it returns a visible answer.

Straight up the opposite, which the name makes abundantly clear, with the option it does reasoning, without it it doesn't...

by embedding-shape

Hetzner entering LLM inference is the cloud provider equivalent of your landlord also offering to cook you dinner. the margins on compute and the margins on food both rely on you not reading the invoice too carefully

by luciana1u

Very soon we will be having reseller programs for inference, this will be just like web hosting reseller. After big players, small players will also start entering in this field.

I'm waiting for that day so that inference will be affordable just like web hosting. 200$ per month is in no way affordable by everyone.

by pmg1991

This seems like a smart move, given their ability to host efficiently. I approve of efforts to make the cost of inference for smaller useful models slowly approach 'close to zero' and there are many good paths for getting there. It is useful for companies to get fast hosting for the class of smaller models they may end up hosting in house.

by mark_l_watson

It would certainly be interesting to have a highly respected EU-native inference provider, if only to make the regulatory gods happy

by swiftcoder

Join the discussion

Write your take first — we'll ask for email only when you're ready to publish.

  • Hacker News
  • Excited to see the price of every other product they offer triple for no reason.
  • Whats definitely missing: a solid (non Mistral) GDPR compliant coding plan / subscription. All offerings are either US or China based. With the newest open weights models this became really interesting imo.
  • > The enable_thinking option is worth mentioning. Without it, the model can spend a surprising amount of the completion budget reasoning before it returns a visible answer.

    Straight up the opposite, which the name makes abundantly clear, with the option it does reasoning, without it it doesn't...

  • Hetzner entering LLM inference is the cloud provider equivalent of your landlord also offering to cook you dinner. the margins on compute and the margins on food both rely on you not reading the invoice too carefully
  • Very soon we will be having reseller programs for inference, this will be just like web hosting reseller. After big players, small players will also start entering in this field.

    I'm waiting for that day so that inference will be affordable just like web hosting. 200$ per month is in no way affordable by everyone.

  • This seems like a smart move, given their ability to host efficiently. I approve of efforts to make the cost of inference for smaller useful models slowly approach 'close to zero' and there are many good paths for getting there. It is useful for companies to get fast hosting for the class of smaller models they may end up hosting in house.
  • Good to see more developments in this space. I quite like this service, which is a little further than Hetzner and has several models to choose from: https://www.infomaniak.com/en/hosting/ai-services
  • It would certainly be interesting to have a highly respected EU-native inference provider, if only to make the regulatory gods happy