model-review

Tag

Cards List
#model-review

@CtrlAltDwayne: I don't know what @cognition did or how they did it, but SWE-2 is the real deal guys. The only downside is the lack of …

X AI KOLs Following ↗ · 2026-09-13

A user praises SWE-2 from Cognition for its strong coding capabilities, comparing it to Codex but noting the lack of computer use features.

0 favorites 0 likes
#model-review

@RayFernando1337: Grok 4.6 Ran All Night, Is It Good?

X AI KOLs Following ↗ · 2026-08-13 Cached

Ray Fernando hosts a live broadcast testing Grok 4.6 overnight and discusses whether the model is good.

0 favorites 0 likes
#model-review

V4-Flash-0731 - vibes after first weekend of use

Reddit r/LocalLLaMA ↗ · 2026-08-03

A user shares hands-on impressions of V4-Flash-0731 after a weekend of testing, noting that quantization heavily degrades performance, Q3 weights can replace Qwen3.6-27B in agentic workflows, and full precision approaches GLM 5.2-level capability at remarkably low cost, though it is weak in general knowledge.

0 favorites 0 likes
#model-review

Review testing on Ling 3.0 flash - From one prompt to a 3D world

Reddit r/LocalLLaMA ↗ · 2026-07-31

A review of the newly released Ling 3.0 flash model, which can generate 3D worlds from a single text prompt.

0 favorites 0 likes
#model-review

Honest take on Laguna S2.1 and its uses (from actual use)

Reddit r/LocalLLaMA ↗ · 2026-07-24

A user shares their experience with the Laguna S2.1 model, finding it effective for complex debugging due to its thorough reasoning style, but not suitable as a general planner. It successfully fixed bugs that other models like Qwen and Claude could not.

0 favorites 0 likes
#model-review

Quick thoughts on GLM-5.2 (Bonus: Censorship question answers)

Reddit r/LocalLLaMA ↗ · 2026-06-18

A detailed user review of GLM-5.2 accessed via API, praising its long-context coherence, adaptive reasoning, and frontier-level text performance comparable to GPT-5.5, while noting the lack of native vision and high local compute requirements.

0 favorites 0 likes
#model-review

@TheGeorgePu: I'm trying out DeepSeek V4 Pro, and really like it. Super underrated model. As good as Opus 4.8 from the few tests I ra…

X AI KOLs Timeline ↗ · 2026-06-17 Cached

User @TheGeorgePu praises DeepSeek V4 Pro, calling it underrated and comparing it favorably to Opus 4.8 based on initial tests.

0 favorites 0 likes
← Back to home

Submit Feedback