美国国土安全部表示,将利用这项技术“彻底变革”该机构处理信息遮盖的方式。
The Department of Homeland Security says it will use the technology to “revolutionize” how the agency handles redactions.
信息遮盖是《信息自由法》(FOIA)申请人的心头大患。当政府最终交付公共记录申请时,那些黑框往往遮盖了大段的文字。尽管即使是最坚定的信息公开倡导者也会承认,某些信息(如社会安全号码)确实需要遮盖,但几乎所有的申请人都认为,政府利用信息遮盖隐藏了远超必要范围的内容。
Redactions are the Freedom of Information Act requester’s bane. Those black boxes that often hide wide swaths of words when the government finally delivers on a request for public records. While even the most ardent openness advocate will grant that some information — such as Social Security numbers — needs to be redacted, almost all requesters believe the government uses redactions to hide far more than it should.
如今,联邦机构正着手开展一项可能改变游戏规则的项目,以加快处理《信息自由法》申请的速度:利用人工智能来加速信息遮盖工作。
Now, federal agencies are embarking on a possibly game-changing project to speed up the handling of FOIA requests: using artificial intelligence to turbocharge redactions.
根据一月份发布在联邦人工智能用例清单上的一份说明,美国国土安全部将在响应《信息自由法》申请时,利用人工智能来“简化敏感内容的识别和遮盖”。该清单是根据唐纳德·特朗普总统于2020年签署的一项旨在鼓励透明度的行政命令所要求的,是一份联邦机构计划部署或已经部署的人工智能工具的公开列表。
The Department of Homeland Security will use AI to “streamline the identification and redaction of sensitive content” when responding to FOIA requests, according to a description posted in January to a federal AI Use Case Inventory. The Inventory, required by a 2020 executive order signed by President Donald Trump to encourage transparency, is a public list of AI tools that federal agencies plan to or have deployed.
具体而言,隶属于国土安全部的美国海关及边境保卫局计划使用谷歌的人工智能和机器学习工具,来建议其《信息自由法》官员应遮盖哪些信息。根据说明,“预计这种自动化将显著缩短处理时间”。
Specifically, U.S. Customs and Border Protection, which is part of DHS, plans to use Google’s AI and machine learning tools to recommend what information its FOIA officers should redact. “[T]his automation is expected to significantly reduce processing times,” according to the description.
人工智能的应用似乎有望推广。一位发言人告诉《华盛顿邮报》,到9月底,国土安全部将通过对占《信息自由法》申请总数10%以上(即“该部门收到的所有《信息自由法》申请中约10万至14万份”)的特定类型申请使用人工智能,从而“彻底变革”其《信息自由法》处理流程。
The AI use appears poised to spread. By the end of September, DHS will “revolutionize” its FOIA process by using AI for certain types of requests that represent more than 10 percent of FOIAs, “approximately 100,000-140,000, of all FOIA submissions received at the Department,” a spokesperson told The Washington Post.
该发言人表示,在人工智能提出建议后,“国土安全部的工作人员将审查案件,并确保申请人获得所请求的准确且完整的信息”。
After the AI makes a recommendation, “DHS employees will review the case and ensure that the requester is getting the accurate and complete information requested,” the spokesperson said.
当然,在大多数申请人看来,一份布满黑色涂改线的记录并不能被称为“完整”。
Of course, a record marred with black lines would not be considered “complete” by most requesters.
《信息自由法》(FOIA)要求机构必须公开所请求的任何记录,除非该记录可以根据法律规定的九项豁免条款进行部分涂改或完全不予公开,其中一些豁免属于酌情裁量。然而,其中一项豁免允许国会随时制定新法律以隐瞒信息。国会已经通过了250多项此类法规,禁止公众获知从西瓜处理方式到香烟添加剂等各类主题的相关信息。
The FOIA requires that any record requested from an agency be released unless it can be redacted or withheld in full under one of the nine exemptions specified in the law, some of which are discretionary. However, one of these exemptions permits Congress to create new laws to withhold information at any time. It has passed more than 250 such statutes, forbidding the public to know some information about topics from watermelon handling to cigarette additives.
在清单中,我还发现了其他几个机构正在开发或使用与涂改相关的AI工具的描述。内政部正在使用微软开发的技术来涂改潜在的律师-客户保密信息。司法部正在使用Veritone开发的工具aiWARE来涂改视频。卫生与公众服务部则正在与承包商合作开发一种名为“FRED”的工具,用于涂改未具体说明类型的信息。内政部和司法部均未回应置评请求。那么,这对政府透明度意味着什么呢?
In the Inventory, I also found descriptions of AI tools related to redaction that are in development or in use by several other agencies. The Interior Department is using technology developed by Microsoft to redact potential attorney-client information. The Justice Department is using a tool developed by Veritone, aiWARE, to redact video. And the Department of Health and Human Services is developing a tool called “FRED” with a contractor to redact unspecified types of information.
《信息自由法》的“AI化”将影响该法案的几乎所有使用者。去年,联邦机构收到了超过170万份《信息自由法》请求。这些请求者包括申请服役记录的退伍军人、揭露针对退伍军人权利法案(GI Bill)诈骗行为的记者,甚至包括试图获取监督行政部门所需记录的国会议员。
The departments of Interior and Justice did not respond to requests for comment. So, what would this mean for government transparency? The “AI-ification” of FOIA would affect almost all users of the law. Last year, over 1.7 million FOIA requests were filed with federal agencies. FOIA requesters include veterans requesting their service histories, reporters uncovering scams targeting the GI Bill program, and even members of Congress attempting to obtain records needed to oversee the executive branch.
当被问及国土安全部及其他机构计划采用AI一事时,一些专家警告称,使用该技术可能会导致公开的信息进一步减少。
The “AI-ification” of FOIA would affect almost all users of the law. Last year, over 1.7 million FOIA requests were filed with federal agencies. FOIA requesters include veterans requesting their service histories, reporters uncovering scams targeting the GI Bill program, and even members of Congress attempting to obtain records needed to oversee the executive branch.
另一些人则告诉我,目前利用AI进行涂改的效果可能还不够理想。
When asked about DHS’s and other agencies’ planned adoption of AI, some experts warned that use of the technology could lead to even less information being released.
如果人工智能能够比处理信息自由法(FOIA)申请通常所需的数年甚至数十年更快地、一致地删除允许范围内的最少信息量(而非最大信息量),它很可能会受到申请者的欢迎。
Others told me that the use of AI for redactions may not be very effective yet. If AI consistently redacted the minimum — rather than the maximum — allowable amount of information more quickly than the years or even decades it often takes to process FOIAs, it would likely be embraced by requesters.
然而,2025年9月,联邦信息自由法监察专员办公室对280个受信息自由法约束的部门和机构进行的一项调查发现,仅有18.6%的机构报告称使用人工智能和/或机器学习来协助处理申请。
But a September 2025 survey of 280 departments and agencies subject to FOIA — conducted by the federal FOIA ombuds office — found that just 18.6 percent reported using AI and/or machine learning to aid in processing requests.
巴伦表示,这一数字将广泛采用的技术,如关键词搜索、电子邮件链去重以及光学字符识别(将扫描的文本图像转换为数字文本),都算作“人工智能使用”。当然,自20世纪90年代以来,大多数机构在回应信息自由法申请时就已经在使用原始类型的人工智能,例如在发送批准或拒绝信函前使用拼写检查。迄今为止,关于人工智能能够成功用于推荐适当删改的证据有限。
Baron said that number counts widely adopted technologies such as key word searches, the de-duplication of email chains and optical character recognition (which turns scanned images of text into digital text) as “AI use.” Certainly, the bulk of agencies have been using primitive types of AI for FOIA responses since the 1990s, for instance when they used spell check before sending release or denial letters. So far, there is limited evidence that AI can be successfully used to recommend appropriate redactions.
2022年,巴伦与两位合著者在《计算与文化遗产杂志》上发表的一篇论文中,使用人工智能审查了克林顿总统图书馆电子邮件记录及附件中的3000多个段落。他们指示人工智能查找可根据信息自由法“审议过程特权”予以保留的信息,这是一项可以(但并非必须)用于删改政府官员之间讨论的豁免条款。随后,他们将人工智能推荐的删改内容与巴伦和另一位律师确定应当进行的删改内容进行了比较。他们发现,在查找和分离可依据该豁免条款合法保留的段落方面,人工智能方法的准确率通常为70%。
In a paper published in 2022 in the Journal on Computing and Cultural Heritage, Baron and two co-authors used AI to review over 3,000 paragraphs in email records and attachments from the Clinton Presidential Library. They told it to look for information that could be withheld under FOIA’s “deliberative process privilege,” an exemption that can — but is not required to — be used to redact discussions between government officials. Then they compared the AI’s recommended redactions with those that Baron and a fellow lawyer determined should be made. They found that AI methods generally performed at an accuracy rate of 70 percent in finding and segregating paragraphs that could be legally withheld under the exemption.
巴伦告诉我,在他的研究过程中,人工智能有时会错误地建议保留信息,有时又会错误地建议公开信息。
In a paper published in 2022 in the Journal on Computing and Cultural Heritage, Baron and two co-authors used AI to review over 3,000 paragraphs in email records and attachments from the Clinton Presidential Library. They told it to look for information that could be withheld under FOIA’s “deliberative process privilege,” an exemption that can — but is not required to — be used to redact discussions between government officials. Then they compared the AI’s recommended redactions with those that Baron and a fellow lawyer determined should be made. They found that AI methods generally performed at an accuracy rate of 70 percent in finding and segregating paragraphs that could be legally withheld under the exemption.
尽管各机构才刚刚开始采用人工智能来对信息进行遮盖,但我们已经看到了这项技术的发展速度之快。我最担心的是,这种转变会导致公众被剥夺更多的信息——即便处理速度变快了。
Baron told me that during his research, AI sometimes improperly recommended withholding information, and sometimes improperly recommended to release it. Even though agencies are just beginning to embrace AI to redact information, we’ve seen how quickly the technology is taking off. My great fear is that this transformation will lead to even more information being withheld from the public — even if it is done more quickly.
根据最新公布的统计数据,在2025财年,美国国土安全部处理了994,992份《信息自由法》(FOIA)请求,其中仅有2.3%的记录是未经遮盖发布的。在已完成的请求中,有40%的记录是以遮盖形式披露的——遮盖内容可能仅为一个词。其余574,075份请求则被完全拒绝,理由包括适用豁免条款(1.5%)、声称不存在相关记录(40.2%)或其他原因。
In fiscal year 2025, the Department of Homeland Security released records without redactions in just 2.3 percent of its 994,992 FOIA requests processed, according to the most recent published statistics. It disclosed records with redactions — which could be as minimal as one word withheld — in 40 percent of the requests completed. The remaining 574,075 requests were denied in full based on an exemption (1.5 percent), claim that no records existed (40.2 percent) or another reason.
防止人工智能遮盖过度使用的一道防线可能是《2016年信息自由法改进法案》中的一条规定,该规定指出,机构只有在“合理预见披露会损害”合法利益的情况下,才可以隐瞒信息。
One bulwark against overreach with AI redactions may be a line in the FOIA Improvement Act of 2016, which states that an agency may only withhold information if it “reasonably foresees that disclosure would harm” a legitimate interest.
《信息自由法》监察员调查证实了这一条款的重要性:“人工智能和机器学习……不能替代《信息自由法》专业人员在应用豁免和预见性损害判断方面的专业能力。”
The FOIA ombuds survey confirmed the power of that line today: “AI and machine learning … are not a substitute for a FOIA professional’s judgment on application of exemptions and foreseeable harm.”
奎利尔(Cuillier)表示:“关键问题在于政府机构是否会对整个过程保持透明,包括它们如何训练模型以及模型产生的结果,”特别是在使用“预见性损害”标准方面,该标准旨在强制要求尽可能多地发布信息。“因为任何导致文件被涂黑的‘黑箱’操作,都会让所有人产生极大的怀疑。”您有任何问题、意见或关于《信息自由法》的想法吗?请留下评论或发送电子邮件至RevealingRecords@washposm。
“The key question is whether government agencies are going to be transparent on the whole process, including how they teach the model and its outcomes,” notably on the use of the foreseeable harm standard, which is intended to force the release of as much information as possible, Cuillier said. “Because any sense of a ‘black box’ leading to blacked out documents is going to make everybody really skeptical.” Do you have a question, comment or FOIA idea? Leave a comment or email me atRevealingRecords@washposm.