🛰️ Daily AI Frontier
‹ back to 2026-09-09

Record Grouping Controls Evidence Weight in Language Models

Research LLMs & Foundation Models

Ranking

Overall 81
Content 100
Popularity 37

Observed public metrics from 1 member.

Representative image for Record Grouping Controls Evidence Weight in Language Models

Merged summary

TL;DR - This paper shows that how retrieved records are grouped before being presented to a language model can substantially alter their evidential weight and the model’s decisions. It proposes a content-aware representation that deduplicates within groups, aggregates complementary information, and limits each group’s contribution.

  • Equal numbers of record groups can represent different evidence states depending on their content and partitioning.
  • Across 104,402 trials and six public checkpoints, false splits increased measured effects by 10.27–32.66 percentage points, while false merges reduced them by 9.13–31.79 points.
  • A matched six-slot control preserved the positive direction in all 16 tested cells, indicating that the split effect was not solely due to added presentation slots.
  • A 48-item controlled campaign panel found partition-induced decision shifts across all four models, with checkpoint-dependent behavior and substantial ordering interactions.

Sources (1)

Record Grouping Controls Evidence Weight in Language Models

arXiv cs.CL Zhongxuan Liu, Sicheng Zhou, Hongzhi Wang 2026-09-08 arXiv:2609.08698
Public signals Semantic Scholar citations 0 · Semantic Scholar influential citations 0
Providers: Hugging Face · N/A OpenAlex · N/A Publisher · N/A Semantic Scholar · Citations 0 · Influential citations 0 X · N/A Fetched 2026-09-13 14:07:52.496556 UTC

TL;DR - This paper shows that how retrieved records are grouped before being presented to a language model can substantially alter their evidential weight and the model’s decisions. It proposes a content-aware representation that deduplicates within groups, aggregates complementary information, and limits each group’s contribution.

  • Equal numbers of record groups can represent different evidence states depending on their content and partitioning.
  • Across 104,402 trials and six public checkpoints, false splits increased measured effects by 10.27–32.66 percentage points, while false merges reduced them by 9.13–31.79 points.
  • A matched six-slot control preserved the positive direction in all 16 tested cells, indicating that the split effect was not solely due to added presentation slots.
  • A 48-item controlled campaign panel found partition-induced decision shifts across all four models, with checkpoint-dependent behavior and substantial ordering interactions.
item →