Join the discussion

Write your take first — we'll ask for email only when you're ready to publish.

  • Hacker News
  • Finally some determinism in our high-temperature sampling world!
  • a few days ago I started an agent on gpt-5.6-terra to work on a project, and one of the website pages had a sentence to create GH issues. Agent read it and that was enough to derail and go creating issues with my context
  • I still remember the times when ai/ml security was about perturbing pixel gradients to misclassify a panda
  • aussi, si vous êtes à Montréal et vous aimerez apprendrez plus sur OpenAPPA, viens à l'evenement CNCF demain soir, le 29 - je ferai un p'tit discours sur l'integration kagent d'OpenAPPA.

    https://www.meetup.com/kubernetes-montreal/events/316689391/

  • Guardrails with builtin remediation instead of simply blocking my agent is a mind blowing long awaited experience! Sooo good. Can't recommend more!
  • Hi! One of the OpenAPPA authors here. Ask me anything!

    My favorite part of APPA is “batteries”: you can run arbitrary programs as part of an authorization decision. For example, a battery could call the GitHub API to check whether a repository is public or private, then use that result to decide whether its contents can be posted to Slack.

  • Quick disclaimer, I work at Archestra.

    I’ve had the chance to play with OpenAppa for a bit and if there’s one thing that I love with this project: it’s simple to get started with and easy to tweak. imo agentic security shouldn’t have to be painful to setup.

    Give it a shot and hopefully ya’ll will find this project useful. It's also open source :)

  • Really interesting direction. What resonated with me is that you're treating agent security as an information-flow problem rather than a prompt-classification problem. It was not so obvious to me.

    A key question I agree isn't just "is this tool call allowed?", but "given everything the agent has read so far, is this information now allowed to flow to this destination?" That feels like a much more fundamental abstraction.

    The part I'm particularly curious about is how this will work with policy authoring at scale. What would be the main adoption challenge?

Explore Birbla archives