ATTRIB 2026 @ NeurIPS

3rd Workshop on Attributing Model Behavior at Scale: Data Attribution and Provenance

Sydney, Australia

Contact info: attrib-neurips26 [at] googlegroups [dot] com

Submissions: OpenReview

How can we attribute model outputs and actions back to data? How can we design attributions that inform downstream uses like alignment, data use litigation, and regulatory audits?


As generative AI permeates rapidly across society, an increasingly relevant problem is attribution: how can we attribute model outputs and actions back to data? The appropriate notion of attribution depends on the specific applications. Contributive attribution aims to determine which training data causally influenced the generated output. A growing body of work has made this increasingly tractable, though challenges remain. In parallel, corroborative attribution methods leverage lexical techniques, conditional generation, and textual entailment to identify sources that textually entail, or otherwise semantically correspond to, a model output, i.e., citations.

However, contributive methods produce scores that are often difficult to interpret or verify and corroborative methods identify semantically supporting sources but make no claims about causal responsibility. In practice, however, stakeholders in law, journalism, and the arts need attribution that is interpretable, auditable, and grounded in identifiable sources, beyond computability. Meanwhile, the rapid growth of synthetic and model-generated data further complicates attribution by blurring the boundary between training data and model outputs. Bridging the gap between methods and applications will require not only technical innovation but also clearer problem formulations informed by the concrete constraints of real applications.

This workshop will bring together researchers from both the contributive and corroborative attribution communities with practitioners in law, music, journalism, and AI safety. Our goal is to identify where existing methods fall short of practical needs, surface shared technical challenges across application domains, and establish concrete research directions that can move attribution from a purely academic tool toward one that is useful in practice.

Call for Papers

Submissions open August 1st!
We are soliciting papers along two tracks: Along these tracks, we welcome submissions pertaining to any aspect of model behavior attribution. For example:

LLM Usage Policy

ATTRIB uses the same LLM usage policy as NeurIPS 2026. Please see here for details.

Submission Instructions

  1. Format submissions as follows:
    • 3-6 pages (main track) or 2-4 pages (idea track)
    • NeurIPS 2026 paper formatting (download from here)
    • Appendix included in the same PDF as the main body
    • No Appendix page limit

  2. ATTRIB uses a reciprocal reviewing model: for each submission, at least one author is expected to serve as a reviewer. Reviewers are assigned up to 2 papers, with reviews due September 22 (AOE). Please designate your submission's reciprocal reviewer on the Reviewer Registration Form linked in the OpenReview portal. We rely on authors to make the review process work, and submissions without a participating reviewer may be desk rejected.

  3. When ready, submit to OpenReview (the workshop is non-archival).

Important Dates


August 1: Submission portal opens

September 1 (AOE): Deadline for both idea and main track papers

September 22 (AOE): Deadline for Reviews

September 29: Decision notifications

Speakers

TBA