prompting-techniques

Tag

Cards List
#prompting-techniques

LLM-as-judge anchored on one confidence value in 10 of 16 evals. Asking for a label fixed it.

Reddit r/AI_Agents · 5d ago

An LLM judge consistently returned a fixed confidence score of 0.72 in evaluations, but switching to categorical labels improved score distribution, showing that models are better at classification than numerical estimation for assessments.

0 favorites 0 likes
#prompting-techniques

Using Claude Code: The Unreasonable Effectiveness of HTML

Simon Willison's Blog · 2026-05-08 Cached

Simon Willison discusses the effectiveness of using HTML instead of Markdown as AI output format, highlighting benefits like SVG diagrams, interactive widgets, and rich explanations. Includes examples from Thariq Shihipar on Anthropic's Claude Code team and practical prompts for GPT-5.5.

0 favorites 0 likes
← Back to home

Submit Feedback