bias-testing

Tag

Cards List
#bias-testing

LLM-Driven Autonomous Vehicles Inherit Human Driver Biases in Pedestrian Yielding: Results and Implications From A New Benchmark

arXiv cs.AI ↗ · 2026-09-02 Cached

This paper proposes new bias testing methodologies for LLMs and VLMs in autonomous vehicles, showing that these models inherit human biases in pedestrian yielding decisions based on attributes like gender, ethnicity, and age.

0 favorites 0 likes
#bias-testing

@no_stp_on_snek: Qwen3.8 lands in tomorrw, so I went back and finished the 3.6-27B card first. No point measuring a successor against a …

X AI KOLs Following ↗ · 2026-08-13 Cached

The author presents an off-label evaluation card for Qwen3.6-27B, covering quantization, reasoning mode effects, bias probes, and jailbreak resistance, and compares reasoning effects with Nemotron 3.5 Lightning, finding that thinking mode is net-negative for Qwen but positive for Nemotron.

0 favorites 0 likes
#bias-testing

Testing GLM 5.2 on Political Bias

Reddit r/ArtificialInteligence ↗ · 2026-07-10

Testing the GLM 5.2 language model for political bias to assess fairness and neutrality.

0 favorites 0 likes
#bias-testing

ai chatbots politically biased? here’s what the washington post found from testing:

Reddit r/singularity ↗ · 2026-06-24

The Washington Post tested major AI chatbots and found evidence of political bias in their responses, raising concerns about objectivity in AI systems.

0 favorites 0 likes
#bias-testing

Semantic Needles in Document Haystacks: Sensitivity Testing of LLM-as-a-Judge Similarity Scoring

arXiv cs.CL ↗ · 2026-04-22 Cached

Researchers from PNNL and Washington University introduce a systematic framework to test how five LLMs detect subtle semantic changes in documents, revealing positional bias, context coherence effects, and model-specific scoring fingerprints.

0 favorites 0 likes
← Back to home

Submit Feedback