Join the discussion

Write your take first — we'll ask for email only when you're ready to publish.

  • Hacker News
  • Works with GeForce RTX 20 series and newer, RTX PRO (Turing and newer), DGX Spark and Apple M4 or newer.

    It does not pool memory or split one inference request across machines. Adding machines increases parallel throughput but won't let you run bigger models.

Explore Birbla archives

Nvidia Personal AI Router (Pair) · Birbla