Back to feed
arXiv cs.AI·

Forced Deferral: Manipulating Routing Decisions in Multimodal LLM Cascades

Signal
72
Hype
25
In three linesResearchers identify a vulnerability in multimodal LLM cascades: an adversarial attack (Forced Deferral Attack) manipulates weak-model confidence to force routing to the strong model, increasing compute costs without targeting answer correctness.
Read source
Your take?
AI safetyVisionBenchmarks

Summary generated by Claude — human-verified