惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

人人都是产品经理
人人都是产品经理
Stack Overflow Blog
Stack Overflow Blog
L
LINUX DO - 最新话题
Google Online Security Blog
Google Online Security Blog
Schneier on Security
Schneier on Security
Spread Privacy
Spread Privacy
www.infosecurity-magazine.com
www.infosecurity-magazine.com
雷峰网
雷峰网
Google DeepMind News
Google DeepMind News
Microsoft Azure Blog
Microsoft Azure Blog
IT之家
IT之家
V
Vulnerabilities – Threatpost
K
Kaspersky official blog
S
Schneier on Security
B
Blog
The Register - Security
The Register - Security
SecWiki News
SecWiki News
Hacker News: Ask HN
Hacker News: Ask HN
cs.AI updates on arXiv.org
cs.AI updates on arXiv.org
S
Security Affairs
T
The Blog of Author Tim Ferriss
G
Google Developers Blog
T
Tenable Blog
P
Proofpoint News Feed
Apple Machine Learning Research
Apple Machine Learning Research
D
DataBreaches.Net
S
Secure Thoughts
Security Latest
Security Latest
H
Heimdal Security Blog
The Hacker News
The Hacker News
O
OpenAI News
AWS News Blog
AWS News Blog
量子位
Exploit-DB.com RSS Feed
Exploit-DB.com RSS Feed
腾讯CDC
U
Unit 42
L
Lohrmann on Cybersecurity
cs.CL updates on arXiv.org
cs.CL updates on arXiv.org
L
LangChain Blog
阮一峰的网络日志
阮一峰的网络日志
T
The Exploit Database - CXSecurity.com
NISL@THU
NISL@THU
cs.CV updates on arXiv.org
cs.CV updates on arXiv.org
Application and Cybersecurity Blog
Application and Cybersecurity Blog
Hugging Face - Blog
Hugging Face - Blog
The Last Watchdog
The Last Watchdog
Recorded Future
Recorded Future
V2EX - 技术
V2EX - 技术
爱范儿
爱范儿
F
Full Disclosure

cs.CR updates on arXiv.org

CTFusion: A CTF-based Benchmark for LLM Agent Evaluation Large Language Models for Agentic NetOps and AIOps: Architectures, Evaluation, and Safety From Controlled to the Wild: Evaluation of Pentesting Agents for the Real-World Red-Teaming Agent Execution Contexts: Open-World Security Evaluation on OpenClaw Graph Representation Learning Augmented Model Manipulation on Federated Fine-Tuning of LLMs Containment Verification: AI Safety Guarantees Independent of Alignment Defense effectiveness across architectural layers: a mechanistic evaluation of persistent memory attacks on stateful LLM agents From Specification to Deployment: Empirical Evidence from a W3C VC + DID Trust Infrastructure for Autonomous Agents SafeHarbor: Hierarchical Memory-Augmented Guardrail for LLM Agent Safety Agentic Vulnerability Reasoning on Windows COM Binaries From Beats to Breaches:How Offensive AI Infers Sensitive User Information from Playlists Undetectable Backdoors in Model Parameters: Hiding Sparse Secrets in High Dimensions When Embedding-Based Defenses Fail: Rethinking Safety in LLM-Based Multi-Agent Systems When Alignment Isn't Enough: Response-Path Attacks on LLM Agents Block-wise Codeword Embedding for Reliable Multi-bit Text Watermarking FlexServe: A Fast and Secure LLM Serving System for Mobile Devices with Flexible Resource Isolation TwoHamsters: Benchmarking Multi-Concept Compositional Unsafety in Text-to-Image Models Symbolic Guardrails for Domain-Specific Agents: Stronger Safety and Security Guarantees Without Sacrificing Utility Hardening x402: PII-Safe Agentic Payments via Pre-Execution Metadata Filtering Hijacking Text Heritage: Hiding the Human Signature through Homoglyphic Substitution Like a Hammer, It Can Build, It Can Break: Large Language Model Uses, Perceptions, and Adoption in Cybersecurity Operations on Reddit Private Seeds, Public LLMs: Realistic and Privacy-Preserving Synthetic Data Generation Chimera: Neuro-Symbolic Attention Primitives for Trustworthy Dataplane Intelligence Sockpuppetting: Jailbreaking LLMs by Combining Prefilling with Optimization StegoStylo: Squelching Stylometric Scrutiny through Steganographic Stitching Learning-Based Automated Adversarial Red-Teaming for Robustness Evaluation of Large Language Models AutoGraphAD: Unsupervised network anomaly detection using Variational Graph Autoencoders CrossGuard: Safeguarding MLLMs against Joint-Modal Implicit Malicious Attacks Feedback Lunch: Learned Feedback Codes for Secure Communications Noise Aggregation Analysis Driven by Small-Noise Injection: Efficient Membership Inference for Diffusion Models A First Look at the Security Issues in the Model Context Protocol Ecosystem Formalizing the Safety, Security, and Functional Properties of Agentic AI Systems MEASER: Malware embedding attacks on open-source LLMs Fall into a Pit, Gain in a Wit: Cognitive-Guided Harmful Meme Detection via Misjudgment Risk Pattern Retrieval When Search Goes Wrong: Red-Teaming Web-Augmented Large Language Models Differentially Private Synthetic Text Generation for Retrieval-Augmented Generation (RAG) From surveillance to signalling: escalation channels as environmental controls for agentic AI STAC: When Innocent Tools Form Dangerous Chains to Jailbreak LLM Agents Federated Spatiotemporal Graph Learning for Passive Attack Detection in Smart Grids Guidance Watermarking for Diffusion Models SecureVibeBench: Benchmarking Secure Vibe Coding of AI Agents via Reconstructing Vulnerability-Introducing Scenarios xOffense: An Autonomous Multi-Agent Framework for Penetration Testing with Domain-Adapted Large Language Models Hammer and Anvil: Toward a Theory of Backdoors in Federated Learning Neuro-Symbolic AI for Cybersecurity: State of the Art, Challenges, and Opportunities Tell-Tale Watermarks for Explanatory Reasoning in Synthetic Media Forensics Between a Rock and a Hard Place: The Tension Between Ethical Reasoning and Safety Alignment in LLMs A Comprehensive Guide to Differential Privacy: From Theory to User Expectations Enabling Transparent Cyber Threat Intelligence Combining Large Language Models and Domain Ontologies Unveiling Unicode's Unseen Underpinnings in Undermining Authorship Attribution Searching for Privacy Risks in LLM Agents via Simulation SPRINT: Robust Model Attribution of Generated Images via Secret Pixel Reconstruction Majority Bit-Aware Watermarking For Large Language Models Coward: Collision-based OOD Watermarking for Practical Proactive Federated Backdoor Detection Prompt to Pwn: Automated Exploit Generation for Smart Contracts Activation-Guided Local Editing for Jailbreaking Attacks Random Walk Learning and the Pac-Man Attack ExCyTIn-Bench: Evaluating LLM agents on Cyber Threat Investigation White-Basilisk: A Hybrid Model for Code Vulnerability Detection Intrinsic Fingerprint of LLMs: Continue Training is NOT All You Need to Steal A Model! InvisibleInk: High-Utility and Low-Cost Text Generation with Differential Privacy Logit-Gap Steering: A Forward-Pass Diagnostic for Alignment Robustness Toward Principled LLM Safety Testing: Solving the Jailbreak Oracle Problem Exploring the Secondary Risks of Large Language Models Benchmarking Misuse Mitigation Against Covert Adversaries Efficient Preimage Approximation for Neural Network Certification Practical Adversarial Attacks on Stochastic Bandits via Fake Data Injection PARASITE: Conditional System Prompt Poisoning to Hijack LLMs Secure LLM Fine-Tuning via Safety-Aware Probing Can Large Language Models Really Recognize Your Name? PoLO: Proof-of-Learning and Proof-of-Ownership at Once with Chained Watermarking A Survey on the Safety and Security Threats of Computer-Using Agents: JARVIS or Ultron? AutoRAN: Automated Hijacking of Safety Reasoning in Large Reasoning Models Remote Rowhammer Attack using Adversarial Observations on Federated Learning Clients Open Challenges in Multi-Agent Security: Towards Secure Systems of Interacting AI Agents DiffMI: Breaking Face Recognition Privacy via Diffusion-Driven Training-Free Model Inversion Chronology of Multi-Agent Interactions for Provenance of Evolving Information Gungnir: Exploiting Stylistic Features in Images for Backdoor Attacks on Diffusion Models Detecting Malicious Concepts without Image Generation in AI-Generated Content (AIGC) How Vulnerable Is My Learned Policy? Universal Adversarial Perturbation Attacks On Modern Behavior Cloning Policies Imitation Game for Adversarial Disillusion with Chain-of-Thought Reasoning in Generative AI PromptGuard: Soft Prompt-Guided Unsafe Content Moderation for Text-to-Image Models A Multiparty Homomorphic Encryption Approach to Confidential Federated Kaplan Meier Survival Analysis Red-Teaming Text-to-Image Models via In-Context Experience Replay and Semantic-Preserving Prompt Rewriting DeTrigger: A Gradient-Centric Approach to Backdoor Attack Mitigation in Federated Learning Privacy Leakage via Output Label Space and Differentially Private Continual Learning ARQ: A Mixed-Precision Quantization Framework for Accurate and Certifiably Robust DNNs CoreGuard: Safeguarding Foundational Capabilities of LLMs Against Model Stealing in Edge Deployment Power-Softmax: Towards Secure LLM Inference over Encrypted Data Hypnopaedia-Aware Machine Unlearning via Psychometrics of Artificial Mental Imagery Anomaly Detection from a Tensor Train Perspective Survival of the Cheapest: Cost-Aware Hardware Adaptation for Adversarial Robustness Improving Clean Accuracy via a Tangent-Space Perspective on Adversarial Training The AI risk repository: A meta-review, database, and taxonomy of risks from artificial intelligence Towards Agentic Runtime Healing Verification of Machine Unlearning is Fragile Aggressive or Imperceptible, or Both: Network Pruning Assisted Hybrid Byzantines in Federated Learning Whispers in the Machine: Confidentiality in Agentic Systems MalPurifier: Enhancing Android Malware Detection with Adversarial Purification against Evasion Attacks Towards Adaptive, Learning-Based Security in Decentralized Applications Can Blockchains Reliably Train Machine Learning Models?
An AI Security Agent for Banking: Multi-Vector Fraud and AML Detection Across Retail and Corporate Accounts
[Submitted on 16 Jun 2026] · 2026-06-17 · via cs.CR updates on arXiv.org

View PDF HTML (experimental)

Abstract:Banks simultaneously face signature-based fraud (card-not-present attacks, account takeover, ATM cloning) and behavioural financial crime (structuring, layering, mule networks, business email compromise) -- two threat families with fundamentally different detection requirements. Static rule engines that reliably catch brute-force and high-velocity events are structurally blind to business-email-compromise (BEC) payment redirection, session hijacking, and money-laundering layering, which are engineered to appear indistinguishable from legitimate activity at the individual transaction or session level. This paper presents an AI security agent for retail and corporate banking that addresses this gap through a three-component fusion architecture operating on two parallel event streams: a transaction stream (card fraud, ACH/wire fraud, AML categories) and a session stream (account takeover, session hijacking, SIM-swap, insider abuse). Each stream combines an LSTM sequence model capturing per-account behavioural history, a statistical velocity/threshold monitor, and a graph/network module capturing account-counterparty relationship patterns (fan-in, fan-out, pass-through ratio) for money-laundering detection. Experiments on a synthetic event log of 237,669 transactions and 113,508 sessions across 13 threat categories and 3,470 simulated accounts demonstrate overall F1 of 0.787 (transaction stream) and 0.867 (session stream) for the proposed model, versus 0.562/0.733 for a rule-based baseline and 0.655/0.713 for an LSTM-only baseline. The agent includes a customer-facing transaction-verification chatbot (96.6% identity verification accuracy, 86.8% mass-reset attack detection) and an analyst case-summary assistant (99.3% action-recommendation F1), with Critical-tier automated response latency under 0.43 ms at the 95th percentile.

Submission history

From: Joseph Walusimbi [view email]
[v1] Tue, 16 Jun 2026 05:58:40 UTC (47 KB)