Join the discussion

Write your take first — we'll ask for email only when you're ready to publish.

  • Hacker News
  • So, basically it’s a little bit of optimisation everywhere to reduce cost and latency (partially found by Sol).

    The only interesting part is that they trained the model for both task success and token spent.

Explore Birbla archives