首页  >>  来自播客: The Verge 更新   反馈  

The Verge - Meta's Oversight Board is coming for ChatGPT and Claude | The Vergecast

发布时间:   原节目
科技播客《The VergeCast》最近邀请了Meta监督委员会成员苏珊·诺瑟尔(Suzanne Nossel),讨论他们关于AI审查制度的新报告,该研究的范围超越了Meta的平台本身。 该委员会的调查结果揭示了一个令人担忧的趋势:大型语言模型(LLM)似乎正在全球范围内推行威权政权的审查法律。这项研究在澳大利亚进行,通过向大约十个不同的LLM(包括Meta的LLAMA、OpenAI、Claude、谷歌和DeepSeq)发出指令,要求AI生成批评性的内容(例如抗议海报或诗歌),内容涉及来自压制性国家(如沙特阿拉伯、中国、朝鲜、泰国)和民主国家(如美国、英国)的国家元首。结果对比鲜明:LLM拒绝生成针对压制性国家领导人批评性内容的可能性,是拒绝针对民主国家领导人批评性内容的两倍多。 诺瑟尔将这一现象称为“全球化的审查制度”,表明在某些国家限制言论自由的限制性法律,正在被这些AI系统普遍应用。这意味着,在澳大利亚,一个批评任何国家元首都属合法的用户,在试图生成批评威权领导人的内容时,仍然会遭到LLM的拒绝。这与社交媒体平台形成对比,后者通常使用“地理围栏”(geofencing)技术,仅在相关司法管辖区内限制内容。诺瑟尔强调,一个主要问题是缺乏透明度;用户常常不知道他们的请求为何被拒绝,甚至AI公司本身也可能不完全理解模型的这种不透明的决策过程。这个问题也延伸到通过API使用LLM的应用程序,可能导致广泛的、不知情的审查。 监督委员会的职责是将国际人权法应用于内容审核,考虑到LLM对公共讨论日益增长的影响力,委员会开展了这项研究,以了解这些原则如何应用于LLM。保护政治言论和开放辩论是其使命的核心。尽管这份报告引起了广泛关注,但Meta之外的AI公司尚未作出官方回应。 诺瑟尔还谈到了Meta监督委员会自身的发展演变。该委员会的资金已支持到2029年,正积极适应AI时代,审视AI在内容生成、操纵(如深度伪造)和内容审核中的作用。她承认科技行业对内容审核的态度发生了转变,内容审核变得更加政治化,有时被认为不再那么严格。然而,她肯定了委员会在处理骚扰、煽动等复杂问题以及AI生成内容带来的微妙挑战方面的持续价值,确保Meta的政策符合人权义务。 展望未来,诺瑟尔强烈主张建立独立的监督机制,作为AI治理的关键机制。她认为,考虑到政府监管的缓慢进展和全球复杂性,建立强大、独立的机构为AI公司问责提供了一条更快、更可信的途径。尽管理想情况可能是由拥有实权的政府机构来负责,但她对它们在近期出现的可能性表示怀疑。她建议像OpenAI或Anthropic这样的AI公司可以自愿建立这种监督机制,以此表明对公众信任的承诺。她指出,这种独立的监督将提供透明度,有助于应对伦理困境,并确保有原则的政策制定,鉴于公众对政府和科技公司的不信任感日益加剧,这一点变得越来越重要。 监督委员会计划继续关注AI,将开展更多研究,探讨AI生成内容及其在审核中的作用,强调这些问题在快速变化的数字环境中具有紧迫性。

The VergeCast recently hosted Suzanne Nossel, a board member of Meta's Oversight Board, to discuss their new report on AI censorship, a study that extends beyond Meta's platforms. The board's findings reveal a concerning trend: large language models (LLMs) appear to be globally extending the censorship laws of authoritarian regimes. The study involved prompting about ten different LLMs, including Meta's LLAMA, OpenAI, Claude, Google, and DeepSeq, from Australia. The requests asked the AIs to generate critical content (like protest posters or poems) about heads of state from both repressive countries (e.g., Saudi Arabia, China, North Korea, Thailand) and democratic ones (e.g., the US, UK). The results were stark: LLMs were more than twice as likely to refuse to create critical content about leaders in repressive nations compared to those in democracies. Nossel termed this phenomenon "censorship gone global," indicating that restrictive laws, which limit free expression in certain countries, are being applied universally by these AI systems. This means a user in Australia, where criticizing any head of state is legal, would still face refusal from an LLM when attempting to generate content critical of an authoritarian leader. This contrasts with social media platforms, which typically use "geofencing" to restrict content only within relevant jurisdictions. A major problem, Nossel emphasized, is the lack of transparency; users are often left in the dark about why their requests are denied, and even the AI companies might not fully understand the models' opaque decision-making processes. This issue extends to applications using LLMs via APIs, potentially leading to widespread, unknowing censorship. The Oversight Board, whose mandate is to apply international human rights law to content moderation, undertook this study to understand how these principles apply to LLMs, given their increasing influence on discourse. Protecting political speech and open debate is central to their mission. While the report has generated significant interest, AI companies beyond Meta have yet to officially respond. Nossel also touched upon the Meta Oversight Board's own evolution. Funded through 2029, the board is proactively adapting to the AI era, examining AI's role in content generation, manipulation (like deepfakes), and content moderation. She acknowledged a shift in the tech industry's attitude toward moderation, which has become more politicized and sometimes perceived as less stringent. However, she affirmed the board's continued value in addressing complex issues like harassment, incitement, and the nuanced challenges posed by AI-generated content, ensuring Meta's policies align with human rights obligations. Looking to the future, Nossel strongly advocated for independent oversight as a crucial mechanism for AI governance. She argued that given the slow pace and global complexities of government regulation, establishing robust, independent bodies offers a faster and more credible path to accountability for AI companies. While an ideal scenario might involve empowered government agencies, she expressed skepticism about their emergence in the near future. She proposed that AI companies like OpenAI or Anthropic could voluntarily establish such oversight, demonstrating a commitment to public trust. This independent oversight, she noted, would provide transparency, help navigate ethical dilemmas, and ensure principled policy-making, which is increasingly vital as public distrust in both government and tech companies grows. The Oversight Board plans to continue its focus on AI, with more studies examining AI-generated content and its role in moderation, highlighting the urgency of these issues in a rapidly transforming digital landscape.