人人都是产品经理发布于 08/25 14:33

AI 私聊场景的审核错位:为什么”不生成违规内容”被偷换成了”用户违规”

当你把一份包含社区舆情、用户投诉、行业争议事件的竞品调研材料丢给 AI 辅助分析,弹出来的是"内容违规,请遵守社区规范"——这一刻,产品逻辑和场景认知发生了根本错位。本文从一次具体的产品经理工作场景出发,拆解 AI 平台将"公域社区惩戒逻辑"套用到"私域工作场景"的结构性问题,并指出:平台完全有能力在不牺牲合规安全的前提下,做到不影响用户体验。

查看原文
Google DeepMind Blog发布于 07/17 23:00

Introducing Gemini 3.5 Flash Cyber

(翻译)推出 Gemini 3.5 Flash Cyber

Google introduces Gemini 3.5 Flash Cyber, a lightweight cybersecurity model to find and patch vulnerabilities.

查看原文
Google DeepMind Blog发布于 07/16 17:30

Our approach to bioresilience

(翻译)我们的生物弹性方法

Google DeepMind and Isomorphic Labs are sharing our joint approach to bioresilience and AI models.

查看原文
NVIDIA Technical Blog发布于 08/21 21:00

Where Security Fits in an AI Agent Stack

(翻译)安全在AI Agent技术栈中的位置

As AI agents become more capable and operate over longer horizons, building security and trust into the applications they power becomes increasingly important.

查看原文
Google DeepMind Blog发布于 06/16 23:46

Securing the future of AI agents

(翻译)保障AI代理的未来

Securing internal systems with an AI Control Roadmap, combining traditional safeguards and real-time monitoring.

查看原文
AWS Machine Learning Blog发布于 08/21 00:31

Authoring Dogwood policies from natural language in Amazon Bedrock AgentCore

(翻译)在 Amazon Bedrock AgentCore 中通过自然语言编写 Dogwood 策略

AI agents can take actions that do not match your organization's policies. Policy in Amazon Bedrock AgentCore lets teams enforce controls across agents, now including time-based constraints. This post shows how Policy Authoring turns natural-language policy documents into correct Dogwood policies, w

查看原文
arXiv cs.AI发布于 08/26 12:00

Auditing the Synthetic Memoir: Measuring Scene-Level Confabulation in LLM-Generated Autobiography Against the Documented Record of the Life It Describes

(翻译)审计合成回忆录:对照其所描述生活的文档记录度量LLM生成自传中的场景级虚构

Abstract page for arXiv paper 2608.23640: Auditing the Synthetic Memoir: Measuring Scene-Level Confabulation in LLM-Generated Autobiography Against the Documented Record of the Life It Describes

查看原文
arXiv cs.AI发布于 08/26 12:00

AI Agents Push Humans Out of the Loop

(翻译)AI智能体将人类挤出循环

Abstract page for arXiv paper 2608.23642: AI Agents Push Humans Out of the Loop

查看原文
arXiv cs.AI发布于 08/26 12:00

A Formal Methodological Framework for Auditing Robustness and Fidelity in Explainable AI: From Application to Trust Certification

(翻译)审计可解释人工智能鲁棒性与保真度的形式化方法论框架:从应用到信任认证

Abstract page for arXiv paper 2608.23817: A Formal Methodological Framework for Auditing Robustness and Fidelity in Explainable AI: From Application to Trust Certification

查看原文