Join the discussion

Write your take first — we'll ask for email only when you're ready to publish.

  • Hacker News
  • I wonder if we have an AI LLM equivalent to Moore's Law. Like how often do we expect improvement in this technology and with what timing?
  • > Claude Haiku 5.5 is our fastest model to date at each model’s standard speed, although it runs less quickly than our Opus models in Fast Mode.

    Opus 5.5 runs 117 tps average on Openrouter, so it must be at least 10-20 tps slower for them to mention. IDK why they mention this as it does not help for marketing though. https://openrouter.ai/anthropic/claude-opus-5.5

  • Apparently quite a bit smarter than Luna, I wonder what use cases it can cover. I actually honestly don't need a Haiku level AI to be that smart, and looks like you pay for it in the per token cost, I need speed mainly. I might even rather have a dumber but much faster model for things like web searching and parsing to retrieve results for the app or other LLM to do things with.
  • The monthly API credits for Max plan seems fantastic, especially considering Haiku pricing. Being able to actually use my Claude plan for other harnesses and use-cases on top of regular CC usage is everything I wanted.

    Anthropic has really been doing all the right things in the past few weeks, while OpenAI continues to fumble the bag.

  • This is great! Been using GPT 6 Luna for decompiling my childhood favorite game (Age of Mythology) and this means I can throw Haiku into the mix as well. 17352/21965 functions matched so far...
    by bouk
  • Ran image -> html tests for this. I was curious if this smaller model was good enough for complex UI. It was not.

    Haiku 5.5: https://html.non.io/lcars-haiku-5.5/

    Opus 5.5 for comparison: https://html.non.io/lcars-opus-5.5

    Designs it was building from: https://diffui.ai/app/canvas/5093e689-1e74-4f26-b632-2a4500f...

    One interesting thing is it took a look at the job at hand, and immediately delegated it to Opus 5.5. It at least knows what it isn't good at. Very fast though, and likely best used for small subagent tasks / tightly scoped work.

    by jjcm
  • > Second, this week, we’ll roll out a new monthly API credit to all Max and Team subscribers for use on the Claude Platform. Max 5x users will get $100 in credits per month, Max 20x users will get $200, and Team subscribers will receive up to $500, pooled across their users

    This is a very big benefit for me. I can now ship actual ai enhanced features behind my subscription without paying extra or fully relying on on-device models. I do worry that this is to soften the blow for user-unfriendly changes

  • Pricing is...a bit weird.

        Input
        $0.10 / MTok for prompts up to 100,000 tokens
        $0.50 / MTok for prompts over 100,000 tokens
    
        Output 
        $0.50 / MTok for prompts up to 100,000 tokens
        $2.50 / MTok for prompts over 100,000 tokens
    
    100k tokens is an absurdly low cutoff and it is only applicable to Haiku and not Sonnet or Opus. It's a low enough cutoff that it will be quickly exceeded if you are doing anything with Agents; for typical generation or Jev-like classifiers, it's a good value and as noted in this article, that is apparently the vast majority of Haiku use.

    In both cases, still much cheaper than Haiku 4.5's $1 input / $5 output and these prices better compete with GPT-6 Luna. ($0.10 input / $0.50 output, but with no token threshold [EDIT: the threshold for Luna is apparently 272k])

Explore Birbla archives