1 matches found
DINA: a Dual Defense Framework against Internal Noise and External Attacks in Natural Language Processing
As large language models LLMs and generative AI become increasingly integrated into customer service and moderation applications, adversarial threats emerge from both external manipulations and internal label corruption. In this work, we identify and systematically address these dual adversarial...
6.9AI score
SaveExploits0
20