Join the discussion

Write your take first — we'll ask for email only when you're ready to publish.

  • Hacker News
  • Can I get 32 GB of RAM for a sane price instead?
  • But can it run crysis?
  • How many arms and legs does one of these cost?
  • I'm just barely starting to wrap my head around mapping model sizes and quants to hardware components and constraints.

    I don't get what this device is for.

    I a million percent understand wanting 252GB-VRAM, that would get me a 280-320B model like glm-5.3 or deepseek-v4-flash, which would be a massive improvement over my gpt-oss:20b 16GB toy. I would gladly pay a grand for this, I would never pay ten grand for this, and it seems to be priced around a hundred grand.

    So obviously the customer is commercial not consumer.

    Can anyone planning a project around one of these at work share what their workload is shaped like and how they're modeling price/performance?

    For instance, I don't get the 512GB of system memory, I'd gladly drop that to 128 to save money. Am I missing something about commercial workloads? Is a 1T parameter model at 20 tok/s more important to your workload than a 300B one at 60? Is it simply a co-dependency of not the model but the other software you're running on the machine thats using/interacting-with/being-driven-by the model?

    Whats your napkin math to justify 100k? Actually thats not even really the question, its more like - whats your napkin math to determine between the "dual linked GB10" use case vs this product's use case vs an 8U supermicro with 4 cards use case.

  • 252GB of HBM3e at 100k vs. a multi A6000 96GB setup. GB300 seem expensive in comparison at ~$100k. Am I missing something?
  • Someone should edit the title to make it clear it only has 252GB of “AI” memory
  • > a Kensington slot

    I think you might need a bit more than that at this price point ...

Explore Birbla archives