Join the discussion

Write your take first — we'll ask for email only when you're ready to publish.

  • Hacker News
  • Crossing fingers for a 5.2 flash release - it’s been a while but I still feel like 4.7 flash is one of the strongest local coding models
  • Really? I had a terrible experience with 4.7-flash. Qwen-3.5 is still the best local model for me. (3.6 pushed VRAM usage just out of 24GB and then you're not using a consumer GPU any more)
  • Pretty sure I saw mention of no flash
  • I wish they would write a blog post about capabilities of this new model, what to expect from this model, is it cheaper, is it faster or does it have better quality in the outputs.

    But still, thank you for the release

  • maybe wait til monday guys
    by swyx
  • Is it a coincidence that both MiniMax and Z.ai are releasing frontier open weights models right as the USG is trying to impose a cap on model capability offered to the public?
  • No, not really. This has been telegraphed for a long time by everyone involved. HN denizens have been unashamedly anti-ai for years now, so what makes sense is the not knowing part of this audience. Chinese models are also not frontier models.
  • No, Dario became too tiresome and annoying that someone had to do something. Personally I hope they ban Opus too. It will only provide more support for open models development. Compare Dario horror posts with this from GLM release: “ Intelligence should be open, accessible, and ready to build with, empowering every developer, everywhere.”
  • I don’t think we will know. On the one hand, labs hold back until they have something competitive enough to release. So if Fable isn’t around, it removes that pressure. On the other hand, the Chinese labs have been moving fast anyways and are obviously behind, so it’s not any more of a problem to release a model that isn’t the very best.
  • I would say yes.

    You think they were sitting on a release waiting for the right marketing moment?

  • I think Z.ai rushed a bit for release, for example GLM 5.2 is only available under the coding plan right now and they didn't do a big write up. Not even some charts and graphs about its performance!

    This is around when people were predicting a new GLM to come out, so a couple corners clipped in order to catch the moment. I'm using it right now and it seems decent, but I haven't done heavy work with it yet. The expanded context window is great.

  • I don’t know if any open weight Chinese AI engineers are on HN, but thank you for everything you do for information freedom.
  • I'm interested in seeing how this changes folks' workflows.

    For me, at work I use opus to plan, brainstorm, grill, ask questions about my codebase, etc. It is pretty good about understanding the codebase holistically and providing architecturally clean solutions that actually work. Then I use sonnet as a plan executor and it does well. Follows instructions and runs tests and just overall does great.

    At home I make some toy projects using opencode go (I've standardized on deepseek 4 pro as my opus replacement) but it's pretty obvious from the amount of times I've had to fix or revert a change that broke something that it's no opus. I got similar results with kimi. Have not played too much with Qwen.

    So I'm wondering what I'd use to get a similar stack at work. Folks say that this version of glm is basically Jan 2026 opus pre me f. Big if true. So would I use GLM for plan and Deepseek v4 pro/flash for execution? Or maybe Kimi or Qwen? I know I'll probably never get as good quality code as I do at work but I'm just toying around here.

  • I tend to mix them. Write the thing with GLM and get DS or Opus to review the finished result for issues
  • I use glm for all code investigations and top level system design of all kinds, and then present finding to confirm and act upon to opus. everything that burns token goes there.

    the finding aren't always accurate, but it saves ton of opus token

    likewise I have google ai from my photo storage, so I give claude / opencode a skill that uses gemini (agy now) command line for web searches, using their flash model line.

  • I've found the prompting needs are drastically different from the latest frontier models to the latest open weight models. I can be much more vague and talk about an end goal with the frontier models vs needing to be more prescriptive + have a workflow on the open weight models. This gap continues to close, but the level of abstraction I'm working on with the latest models continues to move much higher.
  • This release was rushed to hang on the coattails of the Mythos drama (“hey, sorry you can’t use Fable, but try us while you wait this weekend!”) I think they planned to release next week, hence benchmarks not all being ready yet.
  • Could be, but AFAIK it was similar with other glm releases. Just a Twitter post with blog post coming later.
  • Released at the exact same time, 5:21 pm (Chinese time), as when Anthropic received the letter from the government banning Fable, and explicitly citing other models becoming unusable.
  • ... really? are you sure about the timezones? That's kind of odd, isn't it?

    Maybe the post was edited afterwards?

  • Given the US government’s latest stunt with Fable, this is looking more and more like the future.

    Can’t rely on strategic products if they’re gated by capricious actors.

    Open weight models are basically immune to that

  • You criticize the government, perhaps rightfully, but give Anthropic a pass. They are the ones fueling this bullshit. Downgrading your results without telling you. Refusing your requests in the name of “safety”. Even if the government didn’t make them pull the model for foreigners, we’d still be in a really shitty situation because Anthropic is really shitty.
  • It’s very likely the Chinese go dark too the second they have parity / lead
  • > Open weight models are basically immune to that

    Somewhat. The US Gov can make it illegal to transact with, download, use, etc. foreign open weight models.

    Of course, enforcement will be difficult for individuals (businesses will comply by default, and they would all be pulled off Github and other US based hosting locations if they went the sanctions route). But, we are also quickly going down the road of frightening levels of mass surveillance, which could aid enforcement.

    The Fable situation sets a very dangerous precedent, and I'm not looking forward the future here. We are losing the fight for information and computing freedom.

  • In the last few days, Chinese labs have given us MiniMaxM3, KimiK2.7 and now GLM5.2. Meanwhile US is censoring models. Reads like fiction.
  • I didn’t follow the news continuously enough to know what 5:21 or your comment meant.

    Background reading:

    https://www.anthropic.com/news/fable-mythos-access

    tl;dr: Anthropic supports government centralized government control over models, Amazon produced a probably bogus request to pull down Mythos and Fable, so Trump pulled it down.

    It’s probably bogus because no evidence of effective jailbreaks were provided, and also Fable/Mythos isn’t any more capable than OpenAI’s pre-jailbroken 5.5 offering, making it a moot point.

    Anthropic can put it back up once they institute citizenship checks for their customers and ban any foreign nationals they employ from using it.

    (All of the above according to Anthropic)

    I’ll editorialize and say that this is blatant illegal retaliation on the part of the admin, and also that anthropic brought it on themselves with their “this model will kill us all” Mythos marketing stunt.

    I guess in this story, Amazon is the useful pawn/idiot. Maybe it’ll go Shakespearean, and we’ll get some lowbrow comic relief from Bezos.

  • The Chinese models are censored (too?).

    > US is censoring models

    For the current Anthropic issue, I’d say that’s more likely to just be generic corruption, revenge, shakdeown, and/or incompetence from the Trump admin. ‘Censoring’ might be technically correct, but I think one of the aforementioned verbs is a better fit.

  • Any idea how kimi2.7 compares with GLM5.2?
  • I don’t understand how I grew up thinking USA is the gold standard is good and China just make cheap copies and is bad.

    But these news really changes my view on China and USA. I can’t believe it almost.

  • Okay so if this model is half a year behind, so let’s say January opus pre-nerf, this is it.

    Inference is actually quite cheap for token costs, the frontier labs burn most of their money on training new models, priced into their token costs ontop of some margins and paying record salaries. So if this goes open, distills are tried out, independent providers around the world host it with actual price competition, the house of cards for anthropic collapses pre-ipo. The floor is opus (open models caught up), the current ceiling is Mythos (self inflicted ban due to the safety bullshit theater), and no way out.

    It’s really comical I think it’s even the same guy that warned about gpt2 being too dangerous to release, well that mindset seems to now doing existential harm to anthropic, while the rest of the world essentially laughs and progresses anyway.