Tag
The article discusses collusion in the AAAI 2027 review process, particularly in reviewer assignment cycles, and critiques the lack of code publication in accepted papers at top AI conferences.
A tweet discussing a discovered quirk where renaming a paper PDF to a longer, positive title improves LLM judge scores, advising caution with score-based LLM evaluation and recommending binary labels instead.
AI-research-feedback is an academic paper review skill for Claude Code. It checks grammar, coherence, formulas, figures, and argument flaws through six parallel agents, supports specifying journals to simulate reviewers, and finally generates a structured review report.
The article recounts how PPO, as one of the core alignment algorithms of ChatGPT, was rejected by the top AI conference NIPS in 2017 on grounds of limited novelty and insufficient improvement, revealing the drawbacks of academic peer review.