Join the discussion

Write your take first — we'll ask for email only when you're ready to publish.

  • Hacker News
  • It would be very nice if artificialintelligence.ai actually, from a UX perspective, did the models in more than one thinking mode. I use claude, and I wanna build a feeling for what high, medium, etc. actually gives me. So far their comparisons, and having tried several different models for my work, has given me a feel of what 50 intelligence actually is. And I believe it would be be of even greater value to get a feel inside the single model I actually use, as most people do, because not many, I believe, switch heavily between models when working. I understand that the cost here is greater but the model provivders should obviously give you free access, because of the great work you are doing.
  • They probably figured they could increase their margins by releasing "6 Terra" as "6 Sol". After being surprised by the Opus 5.5 launch, they release the true "6 Sol" as "6.1 Sol" and with very aggressive pricing.
  • > the cache read discount rises from 90% to 95%. GPT-6.1 Sol’s overall blended price for agentic workloads is therefore slightly lower than GPT-6 Sol. This represents an additional price cut, following GPT-6 Sol’s original 50% discount from GPT-5.6 Sol

    Wow, just wow. We are racing to the bottom with these prices.

  • I suspect google has a model that is at the same level but is waiting strategically to release rather than this constant week to week dash.

    Still seems to me like google is in a good position with AI just based on their size and approach.

  • What do people get out of being so reductionist?

    Every time a new model comes out, people come out in droves "oh I don't notice anything different".

    People have been saying this about <currentModel-1> for 2 years now, and the entire state of AI has changed dramatically.

    It cannot be that the next AI model isn't better, but also suddenly what they are capable of is on an entirely different level.

  • This has got to be a panic move from OpenAI, right? They’ve had some bad press lately because from their billing changes, and Anthropic have finally released a fast, relatively cheap Opus with improved written English.
  • After 5.5, I basically don't notice a jump in model performance, other than my usage ending sooner.
  • I would just LOVE to see all the behind-the-scenes shithousery both companies are employing to one-up the other in this, largely, 2-horse AGI race. Someone should make a mockumentary when all is said and done!

Explore Birbla archives