Join the discussion

Write your take first — we'll ask for email only when you're ready to publish.

  • Hacker News
  • How does Gemini have a million token context window?
  • any hardware recommendations? how much memory do we need to this?
  • by adt
  • Any comparison with existing models on common benchmarks? Text? Coding? MMLU?
  • Amazing how fast AI keeps improving, every new model feels like a big step forward
  • Everyone is worried about AI data centers destroying the planet with their extreme energy needs. Though it seems we have a big learning curve still to make AI inference and training more efficient.

    How likely are we to NOT see the AI data center apocalypse through better algorithms?

  • I switched from chatgpt to Perplexity; and now to Kimi K2, after reading an article here explaining that all the fear around some of the Chinese models spying and so on.. is simply not true. I have to say that in my experience Kimi K2 is way better than perplexity. I hope we can get our act together. Seems that building this Ai's requires a level of collaboration that is in opposition to greed.
  • For the uninitiated, what's a "hybrid linear attention architecture"?

Explore Birbla archives

Kimi Linear: An Expressive, Efficient Attention Architecture · Birbla