Publications
Full list available on Google Scholar.
2026
[1]
Reading Between the Pixels: An Inscriptive Jailbreak Attack on Text-to-Image Models
ACM MM 2026
[2]
SafeGen: Goal-Conditioned Video Diffusion of Safety-Critical Scenarios for VLM-Based Autonomous Driving
ACM MM 2026
[3]
DeltaUI: Framework-Normalized UI State Transition Modeling for Multi-Task Front-End Engineering
ACM MM 2026
[4]
SafeHarbor: Hierarchical Memory-Augmented Guardrail for LLM Agent Safety
ICML 2026
[5]
DMN: A Compositional Framework for Jailbreaking Multimodal LLMs with Multi-Image Inputs
ACL 2026
[6]
AGENTSAFE: Benchmarking the Safety of Embodied Agents on Hazardous Instructions
CVPR 2026 🏆 Outstanding Paper Award
[7]
SecureWebArena: A Holistic Security Evaluation Benchmark for LVLM-based Web Agents
ACL 2026 Findings
[8]
Uncovering Strategic Egoism Behaviors in Large Language Models
ACL 2026 Findings
[9]
RoboSafe: Safeguarding Embodied Agents via Executable Safety Logic
ESR@ICLR 2026 🏆 Outstanding Paper Award
[10]
Mask-GCG: Are All Tokens in Adversarial Suffixes Necessary for Jailbreak Attacks?
ICASSP 2026
[11]
Probabilistic Modeling of Jailbreak on Multimodal LLMs: From Quantification to Application
ESORICS 2026
[12]
Robust Rumor Detection Against Noise
Neurocomputing
[13]
DIVER: Dynamic Iterative Visual Evidence Reasoning for Multimodal Fake News Detection
arXiv:2601.07178
[14]
Reasoning-Oriented Programming: Chaining Semantic Gadgets to Jailbreak Large Vision Language Models
arXiv:2603.09246
[15]
Uncovering Security Threats and Architecting Defenses in Autonomous Agents: A Case Study of OpenClaw
arXiv:2603.12644
[16]
Evolving Deception: When Agents Evolve, Deception Wins
arXiv:2603.05872
[17]
AgentVisor: Defending LLM Agents Against Prompt Injection via Semantic Virtualization
arXiv:2604.24118
[18]
GuardAD: Safeguarding Autonomous Driving MLLMs via Markovian Safety Logic
arXiv:2605.10386
[19]
TrajShield: Trajectory-Level Safety Mediation for Defending Text-to-Video Models Against Jailbreak Attacks
arXiv:2605.01761
2025
[1]
Detoxifying Large Language Models via Autoregressive Reward Guided Representation Editing
NeurIPS 2025
[2]
Manipulating Multimodal Agents via Cross-Modal Prompt Injection
ACM MM 2025
[3]
Reasoning-Augmented Conversation for Multi-Turn Jailbreak Attacks on Large Language Models
EMNLP 2025 Findings
[4]
Jailbreak Vision Language Models via Bi-Modal Adversarial Prompt
IEEE Transactions on Information Forensics and Security
[5]
CogMorph: Cognitive Morphing Attacks for Text-to-Image Models
IEEE Transactions on Dependable and Secure Computing
[6]
PRJ: Perception–Retrieval–Judgement for Generated Images
Electronics 14 (12), 2354
[7]
SafeBench: A Safety Evaluation Framework for Multimodal Large Language Models
International Journal of Computer Vision
[8]
Towards Understanding the Safety Boundaries of DeepSeek Models: Evaluation and Findings
arXiv:2503.15092
[9]
Pushing the Limits of Safety: A Technical Report on the ATLAS Challenge 2025
arXiv:2506.12430
[10]
Prism: Programmatic Reasoning with Image Sequence Manipulation for LVLM Jailbreaking
arXiv:2507.21540
[11]
VEIL: Jailbreaking Text-to-Video Models via Visual Exploitation from Implicit Language
arXiv:2511.13127
[12]
Sequential Comics for Jailbreaking Multimodal Large Language Models via Structured Visual Storytelling
arXiv:2510.15068
[13]
Disentangling Fact from Sentiment: A Dynamic Conflict-Consensus Framework for Multimodal Fake News Detection
arXiv:2512.20670
[14]
Utilizing Jailbreak Probability to Attack and Safeguard Multimodal LLMs
arXiv:2503.06989
2024
[1]
Unveiling the Safety of GPT-4o: An Empirical Study using Jailbreak Attacks
arXiv:2406.06302
[2]
CS-Eval: A Comprehensive Large Language Model Benchmark for Cybersecurity
arXiv:2411.16239
2023
[1]
NBA: Defensive Distillation for Backdoor Removal via Neural Behavior Alignment
Cybersecurity 6 (1), 20
[2]
DLP: Towards Active Defense Against Backdoor Attacks with Decoupled Learning Process
Cybersecurity 6 (1), 9
2021
[1]
DeeSCVHunter: A Deep Learning-Based Framework for Smart Contract Vulnerability Detection
IJCNN 2021
