Lucene search
+L

3 matches found

Packet Storm News
Packet Storm News
•added 2026/07/25 12:00 a.m.•45 views

Poster: Rethinking Security in LLM Code Generation through Real-World Risk Scenarios

Large Language Models LLMs are widely used for code generation, yet their security behavior in realistic development workflows remains underexplored. Existing benchmarks often rely on explicitly specified security requirements, failing to capture real-world scenarios where prompts are frequently...

5.9AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/06/20 12:00 a.m.•53 views

AgentRiskBOM: A Risk-Scoping Security Bill of Materials for Agentic AI Systems

Agentic AI systems retrieve private context, invoke tools, write files, call external services, coordinate with other agents, and may act without human approval. Existing bill of materials artifacts improve transparency for dependencies, model metadata, and training provenance, but leave an agent...

5.8AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2025/06/14 12:00 a.m.•17 views

SOSBENCH: Benchmarking Safety Alignment on Scientific Knowledge

Large language models LLMs exhibit advancing capabilities in complex tasks, such as reasoning and graduate-level question answering, yet their resilience against misuse, particularly involving scientifically sophisticated risks, remains underexplored. Existing safety benchmarks typically focus...

7.2AI score
SaveExploits0
Rows per page
Query Builder