SOURCE-LINKED INTELLIGENCE
Form Over Content In Gradient-Based Data Attribution Methods
Data attribution methods using gradient similarity are widely used to analyze and select training data for large language models, but what gradient similarity actually measures is debated. Some interpret it as identifying task-relevant skills, while other work reports that surface form is the main factor. We resolve this debate for supervised fine-tuning examples by varying task and answer format independently. Specifically, we render benchmarks in different answer formats, such that datasets can share a task without a format or a format without a task. We find that gradient alignment follows
Read original source ↗ Open in workspace
- recordType
- paper
- region
- Global
Evidence & attribution
- arXiv · AI, language, vision and robotics · 2026-09-17T02:15:23.000Z
- arXiv · Artificial Intelligence · 2026-09-17T02:15:23.000Z
First collected: 2026-09-19T20:26:32.566Z. This is not the publication date.