微信内可能无法直接打开本站。请点右上角 ··· → 在浏览器打开,或复制链接。
Granite.Trust Policy Tools: Shareable, Actionable Policies for Generative AI Applications
RSS 官方收录 · 可信分层展示
关键摘要
Granite.Trust发布可共享、可执行的生成式AI安全策略工具包
- 推出YAML格式Actionable Policy规范,支持内容级约束与异常追踪
- 提供合成数据生成管线,产出策略对齐的训练与测试数据
- 开源工具支持策略定义、模型对齐及运行时全生命周期策略执行
AI 摘要 · 来源可核验
正文提要
arXiv:2608.23870v1 Announce Type: new Abstract: When it comes to safety policies for generative AI, one size does not fit all. Each organization and use case needs to mitigate different risks depending on the application context, regulatory environment, organizational values, and user personas. Yet, existing policy specification approaches are designed for traditional access control and fail to capture the nuances of GenAI application: the enforcement of content-based constraints. We present two contributions to address this gap: (1) the Actionable Policy schema, a YAML-based format for specifying what model responses can and cannot contain. The schema enables exception-based policy governance, proposing exceptions to track policy violations; (2) synthetic data generation pipeline that produces policy-aligned training data for model alignment and testing, and a set of tools to help define the schema and enforce policy. Together, these enable organizations to specify policies once and enforce them throughout the GenAI application lifecycle: from model alignment to runtime monitoring. The Actionable Policy schema, example policies, and tools are available as open source: https://github.com/ibm-granite/granite.trust.policy-tools We welcome new ideas, contributions and feedback.