SOURCE-LINKED INTELLIGENCE
Fingerprinting Multimodal Large Language Models
While multimodal large language models (MLLMs) enable a wide range of image-text reasoning tasks, recent incidents indicate that they are vulnerable to illicit deployment and unauthorized distillation. Existing solutions for model provenance are typically confounded by shared language backbones in MLLMs and struggle to detect violations of distillation. To bridge this gap and safeguard model ownership, we present the first study on multimodal model fingerprinting. Inspired by recent findings that self-attention acts as a low-pass filter and that its low-frequency components are informative, we
Read original source ↗ Open in workspace
- recordType
- paper
- region
- Global
Evidence & attribution
- arXiv · AI, language, vision and robotics · 2026-09-17T14:21:39.000Z
- arXiv · Artificial Intelligence · 2026-09-17T14:21:39.000Z
First collected: 2026-09-19T20:26:32.566Z. This is not the publication date.