3 matches found
SecOPD
SecOPD: オンポリシー蒸留による適応型プロンプトインジェクションの軽減 Yibo Peng · Long Lian · David Wagner† · Sizhe Chen† † 共同指導。 論文 プロジェクトページ モデル このリリースは、論文の最終的なfull-response KL 定式化を実装したものであり、 no-parsing バリアントとも呼ばれます。生徒モデルは攻撃されたコンテキスト下で ロールアウトします。同じベースモデルから初期化されたクリーンコンテキストの教師モデルが、 ペアとなるクリーンコンテキスト下で生徒モデルのサンプリングされたトークンをスコアリングしま...
Telemetry and Concealment in Self-Adapting Generative AI: Logging Architecture, Adversarial Model Hiding, and the Limits of Detection
Model risk management MRM guidance assumes a static model lifecycle, in which models are developed, independently validated, and implemented without further autonomous modification. Continually self-adapting generative AI systems --- models that update their own weights during production deployme...
Optimizing Mouse Dynamics for User Authentication by Machine Learning: Addressing Data Sufficiency, Accuracy-Practicality Trade-Off, and Model Performance Challenges
User authentication is essential to ensure secure access to computer systems, yet traditional methods face limitations in usability, cost, and security. Mouse dynamics authentication, based on the analysis of users' natural interaction behaviors with mouse devices, offers a cost-effective,...