Join the discussion

Write your take first — we'll ask for email only when you're ready to publish.

  • Hacker News
  • If you change reasoning during a session, LLM needs to recompute KV cache. Maybe I get it wrong, but I don't think it can save money.

Explore Birbla archives