Skip to content
#

position-bias

Here are 15 public repositories matching this topic...

Option-reversal control for paired binary vision–language benchmarks: estimates a model's answer-slot bias and perceptual accuracy from one extra inference pass per question, attributes each failure to slot or perception, and reports scores against the correct chance floors. Ongoing project; code only.

  • Updated Sep 19, 2026

SSIT: a label-free, gold-free test for whether LLM-judge position x verbosity bias corrections actually compose. No human labels, no model of the judge. Code, 7-judge/6-family pilot data, and pre-registered protocol for the NeurIPS 2026 JUDGe workshop paper.

  • Updated Sep 16, 2026
  • Python

Open-source implementation of the typed-decision pattern popularised by TypeSafe's Jev: read a decision out of a small language model's logits, in the browser. Library, benchmarks and paper. Not affiliated with TypeSafe.

  • Updated Sep 21, 2026
  • Python

Add this topic to your repo

To associate your repository with the position-bias topic, visit your repo's landing page and select "manage topics."

Learn more