api-settings

Tag

Cards List
#api-settings

@OpenAI: A benchmark score reflects the model as well as the harness and settings used to run it. For long-running agents, retai…

X AI KOLs · 16h ago Cached

OpenAI reveals that enabling retained reasoning and context compaction tripled GPT-5.6 Sol's ARC-AGI-3 benchmark scores, highlighting how harness settings significantly impact measured model performance.

0 favorites 0 likes
#api-settings

@OpenAI: GPT-5.6 Sol has been used to solve open problems in mathematics. So why was it struggling with ARC-AGI-3, a benchmark o…

X AI KOLs · 16h ago Cached

GPT-5.6 Sol, a model that solved open math problems, initially struggled with the ARC-AGI-3 benchmark due to a harness memory limitation. Enabling two API settings tripled scores with 6x fewer output tokens.

0 favorites 0 likes
← Back to home

Submit Feedback