6 matches found
CC_Watermark
用于无失真文本水印的最优耦合 用于无失真文本水印的最优耦合 的官方代码 摘要 大型语言模型(LLM)现在能够生成与人类内容无法区分的文本。 这推动了水印技术的发展,水印通过在 LLM 生成的文本中印记"信号",同时对 LLM 输出的扰动最小。 本文对一次性(one-shot)场景下的文本水印进行了分析。 通过带辅助信息的假设检验视角,我们系统地阐述并分析了水印检测能力与生成文本质量失真之间的基本权衡。 我们认为水印设计的一个关键组成部分是在与水印检测器共享的辅助信息与 LLM 词汇表的随机划分之间生成一种耦合。 我们的分析确定了在满足最小熵约束的最坏情况 LLM...
What Is Threat Hunting? A Complete Guide for Security Teams
What Is Threat Hunting? A Complete Guide for Security Teams Security tools catch a lot. They do not catch everything. Automated detection systems rely on known signatures, predefined rules, and behavioral baselines. Sophisticated adversaries know this and design their operations to slip through t...
Entropy Bounds Via Hypothesis Testing and Its Applications to Two-Way Key Distillation in Quantum Cryptography
Quantum key distribution QKD achieves information-theoretic security, without relying on computational assumptions, by distributing quantum states. To establish secret bits, two honest parties exploit key distillation protocols over measurement outcomes resulting after the the distribution of...
Centralized Dynamic State Estimation Algorithm for Detecting and Distinguishing Faults and Cyber Attacks in Power Systems
As power systems evolve with increased integration of renewable energy sources, they become more complex and vulnerable to both cyber and physical threats. This study validates a centralized Dynamic State Estimation DSE algorithm designed to enhance the protection of power systems, particularly...
Unifying Re-Identification, Attribute Inference, and Data Reconstruction Risks in Differential Privacy
Differentially private DP mechanisms are difficult to interpret and calibrate because existing methods for mapping standard privacy parameters to concrete privacy risks -- re-identification, attribute inference, and data reconstruction -- are both overly pessimistic and inconsistent. In this work...
Optimized Couplings for Watermarking Large Language Models
Large-language models LLMs are now able to produce text that is, in many cases, seemingly indistinguishable from human-generated content. This has fueled the development of watermarks that imprint a signal'' in LLM-generated text with minimal perturbation of an LLM's output. This paper provides a...