crash-coding

Tag

Cards List
#crash-coding

Benchmarking Frontier Large Language Models Against Official Crash Database Coding Using Police Crash Narratives

arXiv cs.LG · 6d ago Cached

This paper benchmarks six frontier LLMs on coding crash attributes from police crash narratives against an official fatal-crash database, finding that while GPT-5.5 High leads among LLMs, simple baselines rival or beat LLM performance and attribute-specific differences outweigh model differences.

0 favorites 0 likes
← Back to home

Submit Feedback