Quantcast
Channel: cs.CR updates on arXiv.org
Browsing index pages (23481 articles)
↧

Intent-Hiding Jailbreaks: An Information-Theoretic Framework for...

arXiv:2610.02302v1 Announce Type: new Abstract: Recent work has shown that large language models (LLMs) can be vulnerable to jailbreak attacks in which harmful intent is obscured through composition...

View Article


MIRROR: Multipath Quorum Integrity for LLM Multi-Agent Communication

arXiv:2610.02349v1 Announce Type: new Abstract: Inter-agent communication is central to Large Language Model Multi-Agent Systems (LLM-MAS), but it introduces an underexplored vulnerability:...

View Article


Hop-Decayed Influence: New Vulnerabilities of Structural Auxiliary Indexing...

arXiv:2610.02373v1 Announce Type: new Abstract: GraphRAG pipelines construct auxiliary structures during offline indexing--semantic summaries, hierarchical edges, and pre-computed scores--that...

View Article

Unifying Privacy Accounting: Information Equivalence and Information Loss

arXiv:2610.02414v1 Announce Type: new Abstract: Differential privacy (DP) admits several notions, but the choice among them may affect both privacy analysis and utility. In this paper, we consider four...

View Article

Mitigating Private Data Leakage in LLMs with Whiteout

arXiv:2610.02418v1 Announce Type: new Abstract: Modern large language models (LLMs) are trained on massive, largely unfiltered datasets, including content scraped from nearly every accessible website...

View Article


Evaluating and Improving the Robustness of Large Language Models to Input...

arXiv:2610.02432v1 Announce Type: new Abstract: Large language models (LLMs) in production systems face prompt injections, trojans (backdoors), and manipulation of automatic quality metrics. This...

View Article

SoK: Stablecoins in the Quantum Era

arXiv:2610.02435v1 Announce Type: new Abstract: Stablecoins support payments, trading, collateral, and cross-chain settlement across the digital-asset ecosystem. They also concentrate value behind...

View Article

SideKernel: A Usable microVM Sandbox for AI Coding Agents on macOS

arXiv:2610.02456v1 Announce Type: new Abstract: AI coding agents are untrusted system components, yet they require autonomy on the developer machines they run on. This contradiction is a security...

View Article


CITADEL: CWE-Guided Insertion of Hardware Trojans via Analysis of DFG-Enabled...

arXiv:2610.02544v1 Announce Type: new Abstract: The increasing sophistication of Hardware Trojans (HTs) and system-level vulnerabilities poses significant risks to modern integrated circuits. However,...

View Article


Out of Sync, Out of Sight: Phantom State Attacks against IIoT Intrusion...

arXiv:2610.02552v1 Announce Type: new Abstract: Machine learning-based intrusion detection systems (IDS) are critical for securing Industrial Internet of Things (IIoT) environments. Most adversarial...

View Article

Pincer: Resource Authorization for Agents using a Digital Twin

arXiv:2610.02569v1 Announce Type: new Abstract: Coding agents have become increasingly long-horizon, autonomous, reliant on general-purpose shell and maintain their own persistent memory for...

View Article

From TS-SUF-2 to TS-SUF-4: Practical Security Enhancements for FROST2...

arXiv:2610.02805v1 Announce Type: new Abstract: Threshold signature schemes play a vital role in securing digital assets within blockchain and distributed systems. FROST2 stands out as a practical...

View Article

RMCW: A Deletion-Robust Watermark Based on Reed--Muller Codes for Language...

arXiv:2610.02817v1 Announce Type: new Abstract: Large Language Model (LLM) watermarking provides a lightweight mechanism for identifying text generated by a specific model, but its robustness remains...

View Article


Containing the Autonomous Operator: A Defense-in-Depth Framework and...

arXiv:2610.02861v1 Announce Type: new Abstract: Large language model (LLM) agents are moving from chat interfaces into infrastructure operations, where they read telemetry, call tools, generate and...

View Article

AgentTrap: Stateful Feedback Deception against Autonomous Penetration Testing...

arXiv:2610.02869v1 Announce Type: new Abstract: Autonomous penetration testing agents conduct multi-step attacks by continuously adapting their plans and actions to target responses. As a common...

View Article


Digital Twin-Assisted Mapping of ICS Telemetry to ATT&CK for ICS with...

arXiv:2610.02955v1 Announce Type: new Abstract: Reconstructing adversarial behavior from Industrial Control System (ICS) telemetry is difficult because process observations reveal physical changes more...

View Article

Beyond Predefined Sinks: Security-Aware Dependency Analysis for LLM Agents

arXiv:2610.03014v1 Announce Type: new Abstract: Large language model (LLM)-based agents increasingly connect model-generated decisions to security-sensitive software capabilities such as command...

View Article


SecJev: Bringing Security Expertise to System One Decision Models

arXiv:2610.03073v1 Announce Type: new Abstract: Security workflows need models that turn complex observations and explicit policies into decisions. System One models introduced by Jev return typed...

View Article

Securing Computer-Use Agents Against Branch Steering Attacks

arXiv:2610.03089v1 Announce Type: new Abstract: Modern Computer Use Agents (CUAs) directly interact with graphical user interfaces and execute third-party web tools, exposing them to indirect prompt...

View Article

The Fragility of Trigger-Tag Mechanisms for Misuse Detection in Open-Weight LLMs

arXiv:2610.03124v1 Announce Type: new Abstract: Open-weight language models can be downloaded, modified, and deployed beyond their developers' control, limiting the effectiveness of centrally enforced...

View Article
Browsing index pages (23481 articles)


Latest Images