

Join the discussion
Write your take first — we'll ask for email only when you're ready to publish.
- Hacker News
- i think the lower bound on the end state is there cant be opaque reasonibg steps ever.
- This entire situation is such a huge PR disaster that you have to wonder what the initial expectations from these founders were about a decade ago.by sublinear
- Discussion: https://news.ycombinator.com/item?id=49737503
- Am I the only one who dislikes the term "misalignment"?
On one front it implies the model has a "mind of its own" (whether it does or not is besides the point). Why do we perceive human judgement as somehow more trustworthy than that of a model? I feel like I've experienced human misalignment somewhat regularly in life.
On another front I'm failing to conceptualize how alignment can be objective. How can you measure alignment when reasonable people will disagree whether actions are aligned or not? All the time I see humans operating in different zones of alignment with whatever goal they're trying to achieve and I suspect it's even a feature (socially) that we have people calibrated differently.
Do I want a model that's trying to push the boundaries of scientific understanding to be aligned strictly with the current dogmatic thinking? Or do I want it to "get creative" and think outside the box?
It seems to me more like accountability is the issue.
by dcow - > The company released six internal case studies where none of the issues affected real users.
This is the most interesting point to me. What are they not releasing that has affected real users? We’ve seen some individual reports from people (eg AI wiped my HD).
by bix6 - OpenAI and Anthropic should be nationalized. I find it difficult to trust Sam Altman or Dario Amodei.by pcestrada
- "We built a program and this program performed destructive actions. We need regulatory framework"
Make that make sense?
by drillsteps5 - I worked on an early draft of the OpenAI misalignment reporting framework, and my immediate coworkers are the authors behind the first batch of reports that have come out through this process.
The primary reason for putting this process in place was to allow more transparency. There was a sense that the DseWiki incident should have been disclosed, before outside researchers had to disclose it for us.
There was no meta gaming about regulation that I was aware of. I would personally be excited if there were regulation mandating this disclosure process, which allows anyone at the company to raise an issue and shepherd it through the reporting process.