Join the discussion
Write your take first — we'll ask for email only when you're ready to publish.
- Hacker News
- Finally some determinism in our high-temperature sampling world!by dvorkanton
- a few days ago I started an agent on gpt-5.6-terra to work on a project, and one of the website pages had a sentence to create GH issues. Agent read it and that was enough to derail and go creating issues with my contextby ildari
- I still remember the times when ai/ml security was about perturbing pixel gradients to misclassify a pandaby arseny_info
- aussi, si vous êtes à Montréal et vous aimerez apprendrez plus sur OpenAPPA, viens à l'evenement CNCF demain soir, le 29 - je ferai un p'tit discours sur l'integration kagent d'OpenAPPA.
https://www.meetup.com/kubernetes-montreal/events/316689391/
by joeyorlando - Guardrails with builtin remediation instead of simply blocking my agent is a mind blowing long awaited experience! Sooo good. Can't recommend more!by piercypixel
- Hi! One of the OpenAPPA authors here. Ask me anything!
My favorite part of APPA is “batteries”: you can run arbitrary programs as part of an authorization decision. For example, a battery could call the GitHub API to check whether a repository is public or private, then use that result to decide whether its contents can be posted to Slack.
by keshakon - Quick disclaimer, I work at Archestra.
I’ve had the chance to play with OpenAppa for a bit and if there’s one thing that I love with this project: it’s simple to get started with and easy to tweak. imo agentic security shouldn’t have to be painful to setup.
Give it a shot and hopefully ya’ll will find this project useful. It's also open source :)
by immafridge - Really interesting direction. What resonated with me is that you're treating agent security as an information-flow problem rather than a prompt-classification problem. It was not so obvious to me.
A key question I agree isn't just "is this tool call allowed?", but "given everything the agent has read so far, is this information now allowed to flow to this destination?" That feels like a much more fundamental abstraction.
The part I'm particularly curious about is how this will work with policy authoring at scale. What would be the main adoption challenge?
by vladimir_gor