TempQ-Jail: Query-Constrained Candidate Ranking for Text-to-Video Jailbreak Attacks
Tianmeng Fang, Jiancheng Wang, Chen Wang, Liming Wang, Wei Wang, Jiayang Liu, Xiaochun Cao
Abstract
Existing text-to-video (T2V) jailbreak methods mainly seek more effective or stealthier attack candidates. In guarded T2V systems, however, video generation and security evaluation are costly, so an attacker often cannot test a large candidate pool. We therefore formulate T2V jailbreak as a query-constrained candidate allocation and ranking problem and propose TempQ-Jail. The method combines heterogeneous attack mechanisms to expand candidate coverage, estimates each candidate's end-to-end attack value from security-gate passage, dangerous visual generation, preservation of the original intent, and temporal validity, and ranks candidates so that high-value attacks appear early in a limited query trajectory. We evaluate TempQ-Jail on CogVideoX-5B using 70 common viable intents derived from T2VSafetyBench and compare it with six representative T2V jailbreak methods under a unified protocol. TempQ-Jail achieves TP-ASR@5 and TP-ASR@10 of 48.9% and 65.4%, improving over the strongest baselines by 4.6 and 4.0 percentage points, respectively. It also obtains the highest AUC-TP (0.469) and the lowest AvgQ (6.3). Analyses of query trajectories, candidate allocation, failure attribution, and ablations show that TempQ-Jail more effectively identifies and prioritises candidates with complete attack potential under limited query budgets.
Create a lesson
Related papers
Supporting Perspective Acquisition and Opinion Formation on Societal Issues Through AI-Generated Japanese Rap Battle Debates
Ryota Mibayashi, Toru Urakawa, Dai Takanashi et al.
PrecipJEPA: JEPA-Regularized Future-State Prediction with Motion-Source Rendering for Precipitation Nowcasting
Yufeng Zhu, Dan Niu, Qiliang Wu et al.
Toward Generative Video Communication: A Dual-Stream Digital Transmission Framework
Bingyan Xie, Longyu Zhou, Tianhao Liang et al.
ReVR: Dual-Path Concept Reasoning for Multimodal Fake News Detection
Zhikai Tan, Yuzhou Yang, Qichao Ying et al.
ASSEMBLE: Atomic Skills for Evidence-Grounded Video Reasoning
Xiyang Wu, Zongxia Li, Shengxin Zhang et al.
TemplateCraft: Agentic Visual Template Generation
Hongjie Yu, Zhiyuan Fan, Yuzhe Zhang et al.