proprietary-models

Tag

Cards List
#proprietary-models

One Year Later...The Harms Persist, But So Do We!

arXiv cs.CL · 2026-06-24 Cached

This study evaluates six proprietary LLMs across 16 DSM-5 conditions using adversarial attacks, finding that safety safeguards are only reliable for suicide and self-harm, with failure rates up to 100% for other conditions like eating disorders and substance use disorder.

0 favorites 0 likes
#proprietary-models

There is minimal downside to switching to open models

Hacker News Top · 2026-06-21 Cached

The author argues that switching from proprietary AI models to open models is now much less of a professional sacrifice, citing improving open model quality and Claude's ID verification as a catalyst, similar to the historical shift from Windows to Linux.

0 favorites 0 likes
#proprietary-models

The new benchmarks like DeepSWE now show a very big gap in proprietary models and open source

Reddit r/singularity · 2026-05-31

New benchmarks like DeepSWE reveal a significant performance gap between proprietary and open-source AI models, causing disappointment in the open-source community.

0 favorites 0 likes
#proprietary-models

@Miles_Brundage: I am not sure I have seen a good analysis of how much distillation reduces this gap - people have very different views …

X AI KOLs Timeline · 2026-05-30 Cached

Miles Brundage comments on the lack of quantitative analysis on how distillation affects the capability gap between open-weight and proprietary AI models, referencing a claim by Epoch AI that open-weight models lag by four months.

0 favorites 0 likes
← Back to home

Submit Feedback