

Join the discussion
Write your take first — we'll ask for email only when you're ready to publish.
- Hacker News
- Open Teams originally wanted to rent out open source developers to sponsors with Oliphant controlling everything. Now they pivoted to installing local LLMs (on what hardware exactly?).
What will happen is that this will be the third consultancy with a lofty narrative after Enthought and Anaconda that Oliphant established. It is always bait-and-switch.
by asf1289 - > This is fine in most cases, but for open-weights models it can be a lot more expensive than what the exact same model can be rented for from third-party API providers.
Filter by quantization, and most providers will have the same price. There is some "base" price even for open-weight models. Anything cheaper means some tricks on the provider's side.
by akazantsev - I've been very impressed with GLM 5.3 Flash's performance even without considering cost, but once you factor that in, it's incomparable. Not surprised to see its position on the chart.by SturgeonsLaw
- Complaining about "bad charting" and posting a chart with y-axis that doesn't start at 0 is kinda weird.by vb-8448
- When I accessed the site, it showed the FBI badge says that the site was blocked and redirect to fbi.gov !! WTH?by sinuhe69
- The complaint about not being able to switch between linear and log is valid, which is what I did for making a 3D speed/cost/quality frontier application for a recent meetup talk: https://www.williamangel.net/apps/model_performance.html
Because speed is important, as the reasoning and hardware determine both cost and speed. it's a three dimensional tradeoff.
- There is an immense difference in cost between the state-of-the-art models from Anthropic and OpenAI and the much cheaper Chinese models ... How much extra intelligence emptying the wallet purchases obeys the law of diminishing returns: while a top-tier engineer or scientist is probably going to be able to appreciate how much better Fable 5.1 [is] ... most people will have a hard time doing so.
Nebari is officially listed as a JATIC product as part of the next-gen toolchain supporting DoD AI development.
Are we officially ~one degree of Kevin Bacon from the DoD endorsing running Chinese OSS models because they're self-hosted and we're all too dumb to tell the difference?
https://openteams.com/open-source-isnt-the-real-risk-in-nati...
by themgt - This looks great! I also think speed should be part of the metric (i.e. how long does the model take to actually solve a task). For me, I prefer to run expensive models such as Sol on light reasoning, which usually gives me good answers with quick responses.
For my style of coding (quick back-and-forths and corrections) it makes a big difference if a model comes back in 1-2 minutes compared to 5-10, and I am happy to pay a bit extra for that.
by oliwary