Source · NVIDIA Research
lab NVIDIA Research · 23d ago · 23 views

ReasoningBomb induces pathologically long reasoning in large reasoning models

Researchers have formalized inference-time denial-of-service (PI-DoS) attacks that exploit the high computational cost of explicit multi-step reasoning in Large Reasoning Models (LRMs). They present ReasoningBomb, a reinforcement-learning-based framework that trains an attacker to generate short natural prompts that drive victim models into pathologically long and often non-terminating reasoning traces.