table-qa

Tag

Cards List
#table-qa

ReliableTableQA:How Much Supervision Does Reliability Annotation Need?

arXiv cs.LG · 2026-07-24 Cached

Introduces ReliableTableQA, a framework for training LLMs to annotate statistical reliability of tabular QA results, showing that a small SFT set is sufficient and GRPO only helps when SFT is under-trained.

0 favorites 0 likes
#table-qa

Data Analysis in the Wild: Benchmarking Large Language Models Against Real-World Data Complexities

arXiv cs.CL · 2026-07-08 Cached

The article introduces DataGovBench, a benchmark derived from governmental open data, designed to evaluate LLMs on real-world data analysis tasks including table question answering and insight discovery. Experiments show current LLMs still underperform in complex data analytics scenarios.

0 favorites 0 likes
← Back to home

Submit Feedback