Skip to content
AI Atlas

Context-Dependent Affordance Reports in Vision-Language Models

Published 16 Sept 2026arXiv:2603.04419

data quality89

Updated 12 h ago · first seen 15 Sept 2026

paper_01M2JK0DNTSQJSWYTBHJ05379H

Abstract

-cross Abstract: Vision-language models produce different object and use descriptions under different persona prompts, but low overlap alone does not identify an affordance effect. We audit an earlier seven-prompt study and add matched-question controls. In the historical Qwen pilot, 363 of 3,213 parsed responses contain empty object lists. These affect 2,037 of 9,244 comparisons, with the implementation assigning zero lexical overlap to every affected pair. Conditioning on nonempty reports raises pooled word Jaccard from 0.095 to 0.121 and sentence cosine from 0.415 to 0.511. A previously named chef-specific Tucker factor loses its concentrated loading under missing-cell and complete-nonempty analyses. We withdraw the functional-manifold interpretation and the conversion of similarity scores into percentages of meaning. A new experiment uses 48 images absent from the original pilot, four personas, a shared three-object task, two wordings, and two requested seeds in each of two model configurations. Qwen3.5-9B's persona-minus-wording cosine-distance contrast is 0.0146 (95% CI [-0.0003, 0.0295]; 37 complete images). Ollama llava:13b's persona-minus-wording cosine-distance contrast is -0.0266 (95% CI [-0.0412, -0.0121]; 32 complete images). Persona-associated variation does not uniformly exceed wording or sampling variation. The study provides a reproducible analysis of context-conditioned reports while separating response availability, content similarity, and the limits of inference from linguistic outputs.

Authors

Authors 1

Murad Farzulla

Linked names open researcher pages (created from the paper's author list; name-only, no affiliation unless a source states it). Unlinked names have no researcher record yet.

Organizations

Organizations 0

No organization stated. arXiv metadata does not carry affiliations; an organization is linked only when a model card or lab page cites the paper.

Models

Models introduced or described 0

Inbound described_by relations from model cards and documentation.

No model links this paper yet

Model pages link papers through their model cards and documentation; the relation is written only when a source states it.

Datasets

Datasets used 0

No dataset relation recorded.

Benchmarks

Benchmarks used 0

No benchmark relation recorded.

Code

Repositories & frameworks 0

No repository linked.

Timeline

Timeline 2

Full timeline →

Sources

Sources 2

Source documents
SourceDocumentTypeTierLast observedSnapshots
arXiv (Atom API + RSS)rss.arxiv.org/rss/cs.AI feedT1· Official7 h ago5
arXiv (Atom API + RSS)rss.arxiv.org/rss/cs.LG feedT1· Official7 h ago4

Tier 1 = official/primary, 2 = quality secondary, 3 = community, 4 = unverified. Every snapshot is archived; see all sources and the methodology.