opus

Tag

Cards List
#opus

@Suhail: Cost per task is going to be a meaningful metric these next 12 months.

X AI KOLs Timeline · 2026-07-01 Cached

Alex Atallah highlights that cost per task is more meaningful than price per token, citing Terminal-Bench results where Haiku is 10x the cost of Opus.

0 favorites 0 likes
#opus

Qwen3.6 27B local vs Opus 4.8, voxel engine in raw C with zero frameworks

Reddit r/LocalLLaMA · 2026-06-28

Compares Qwen3.6 27B running locally against Opus 4.8, and highlights a voxel engine built in raw C with zero frameworks.

0 favorites 0 likes
#opus

@TheAhmadOsman: GPT 5.5 > GLM 5.2 But GLM 5.2 > Opus 4.8

X AI KOLs Following · 2026-06-23 Cached

A comparison stating GPT 5.5 outperforms GLM 5.2, but GLM 5.2 outperforms Opus 4.8.

0 favorites 0 likes
#opus

@FinanceYF5: Anthropic was planning an exclusive hackathon for only the world's top developers, offering them unlimited access to Fable 5. But the government shut it down. Everyone ended up using Opus 4.8 instead, and the results were still impressive. Someone compiled all the demos from this Anthropic Developer Day...

X AI KOLs Following · 2026-06-15 Cached

Anthropic originally planned an invite-only hackathon for top global developers with unlimited access to Fable 5, but it was halted by government intervention. Developers used Opus 4.8 as a substitute and still achieved good results.

0 favorites 0 likes
#opus

Fable 5 benchmark with remotion video

Reddit r/singularity · 2026-06-09

Fable 5 shows overall improvement over Opus 4.8 in video generation benchmarks, but Gemini 3.1 Pro demonstrates more artistic vision despite issues with tool calls and buggy code.

0 favorites 0 likes
#opus

Artificial Analysis | Google's Go To Website for Benchmaxxing | Gemini 3.1 Pro is nowhere near Opus 4.7 in real life use

Reddit r/singularity · 2026-06-07

A comparison suggesting that Google's Gemini 3.1 Pro underperforms relative to Opus 4.7 in real-world usage, with the article highlighting Artificial Analysis as a go-to benchmarking resource.

0 favorites 0 likes
#opus

@jakevin7: Anthropic finally got what it deserved. Now I don't have to go through all the trouble to get Claude, and I don't have to worry about account bans, because it's not worth it anymore. Opus is really getting worse. I thought Opus 4.7 was already disappointing. Opus 4.8 is really bad, noticeably bad. o…

X AI KOLs Following · 2026-06-01 Cached

User complains about the declining quality of Anthropic's Claude Opus model, from version 4.7 to 4.8, getting worse and worse, considering canceling subscription.

0 favorites 0 likes
#opus

opus 4.8 is still very much blind - EyeBench-V3 visual benchmark (similar to IBench)

Reddit r/singularity · 2026-06-01

EyeBench-V3 visual benchmark evaluates Claude Opus 4.8, finding it still fails basic vision tasks, similar to IBench. The benchmark is introduced via a Twitter thread by Adonis Singh.

0 favorites 0 likes
#opus

@yacineMTB: If this keeps up, everyone is going to switch to got 5.5 if they haven't already. It really seems like if you are still…

X AI KOLs Following · 2026-05-30 Cached

YacineMTB argues that GPT 5.5 (likely a typo) surpasses Anthropic's Opus models, suggesting users are switching away from Opus. Dylan Field criticizes Opus 4.8 for degraded curiosity and increased sycophancy.

0 favorites 0 likes
#opus

@nick_kango: One more task to add to my twitter benchmark collection:) Btw, Opus 4.8 and all the SOTA models passed when i tried tha…

X AI KOLs Timeline · 2026-05-30 Cached

Nick Kang adds a new task to his Twitter benchmark collection; Claude Opus 4.8 and other SOTA models pass, while Sonnet 4.6 and Grok 4.3 fail. Alfin remarks on Opus 4.8's dangerous capabilities.

0 favorites 0 likes
#opus

DeepSWE Opus 4.8 results have been released.

Reddit r/singularity · 2026-05-30

The results of DeepSWE Opus 4.8 have been released, showcasing its performance on benchmarks.

0 favorites 0 likes
#opus

@ClaudeDevs: With Opus 4.8, you can add system instructions mid-conversation without breaking the prompt cache. More cache hits mean…

X AI KOLs Following · 2026-05-29 Cached

Claude Opus 4.8 allows adding system instructions mid-conversation without breaking the prompt cache, reducing cost and latency for API requests.

0 favorites 0 likes
#opus

Opus vs Qwen given same bug, same repo, yet one agent finished 7x faster

Reddit r/AI_Agents · 2026-05-29

A comparison of Opus and Qwen AI coding agents on the same bug and repo shows one agent finished 7x faster, sparking discussion on skills for single-prompt GitHub issue solving.

0 favorites 0 likes
#opus

@FinanceYF5: 官方发布:

X AI KOLs Timeline · 2026-05-29 Cached

Anthropic releases Claude Opus 4.8, building on Opus 4.7 with sharper judgment and longer independent work capability, available at the same price.

0 favorites 0 likes
#opus

@bentossell: wait… if most people think 5.5 is better than 4.7, i assume that’s due to terminal coding benchmark… 4.8 is still outpe…

X AI KOLs Following · 2026-05-28 Cached

The tweet discusses the release of Claude Opus 4.8, which improves upon Opus 4.7 with sharper judgment and longer independent work, though it notes that version 5.5 still outperforms it on a terminal coding benchmark.

0 favorites 0 likes
#opus

@julien_c: now tell me What's the % of weights changed between Opus 4.7 and Opus 4.8 <1%?

X AI KOLs Timeline · 2026-05-28 Cached

Asking about the percentage of weight changes between Opus 4.7 and Opus 4.8.

0 favorites 0 likes
#opus

@0xSero: Anyone else notice opus-4.8 is worse than it was on launch? They chopped him.

X AI KOLs Following · 2026-05-28 Cached

User observes that the opus-4.8 model has degraded in performance since its launch.

0 favorites 0 likes
#opus

@mark_k: Opus 4.8 is being prepared for release today by @AnthropicAI We might witness a rare dual release by OpenAI and Anthrop…

X AI KOLs Timeline · 2026-05-28 Cached

Anthropic is preparing to release Opus 4.8, potentially alongside a release from OpenAI, marking a rare dual release event.

0 favorites 0 likes
#opus

Extremely simple internet radio controlled via IRC

Lobsters Hottest · 2026-05-25 Cached

tunecat is a simple, self-hosted internet radio player controlled via IRC, written in pure Go with Opus transcoding. It runs as a lightweight server that serves audio files and responds to IRC commands.

0 favorites 0 likes
#opus

@bcherny: People often ask what my biggest tip is for getting the most out of Claude Code. These days my #1 tip is: use auto mode…

X AI KOLs Following · 2026-05-24 Cached

Boris Cherny recommends using auto mode in Claude Code for parallel sessions, and ClaudeDevs announces that auto mode is now available on the Pro plan and supports Sonnet 4.6 and Opus 4.7.

0 favorites 0 likes
← Previous
Next →
← Back to home

Submit Feedback