@Saccc_c: Are Deepseek v4 flash and Qwen 3.8 max under that much pressure? GLM's official website shows GLM-5.3 is live
Summary
The tweet mentions that GLM-5.3 has been launched on GLM's official website, and marvels at the pressure Deepseek v4 flash and Qwen 3.8 max are facing.
View Cached Full Text
Cached at: 08/03/26, 09:39 AM
Are Deepseek v4 flash and Qwen 3.8 max really under that much pressure?
GLM’s official website shows GLM-5.3 has launched https://t.co/5GcfIFnDaL
Similar Articles
@Saccc_c: Want to add multimodal capabilities to DeepSeek v4 flash? I strongly recommend using it with qwen3.7-flash — currently the best value model combination. qwen3.7-flash is a lightweight multimodal model that is fast and well-suited to most image understanding tasks. Key point: the price is low enough, and new registrations get 100…
Recommends pairing DeepSeek v4 flash with qwen3.7-flash to add multimodal image understanding capabilities to DeepSeek at low cost, and provides simple configuration steps using Alibaba Cloud Bailian and Codex.
@zhixianio: Finished testing, feeling quite surprised, not sure if I'm using it wrong. Feel free to provide counterexamples. Here are my results: On M5 Max, pitting this community fine-tuned gemma-4-12B-coder (llama.cpp) against my daily driver Qwen3.6-35B-…
The user tested the community fine-tuned gemma-4-12B-coder against Qwen3.6-35B-A3B MoE on three programming tasks, finding that gemma performed poorly on complex stateful programs, while Qwen 35B remained robust.
@TheAhmadOsman: DeepSeek V4 Flash is ~70% smaller in size than GLM 5.2 It also beats GLM 5.2 which was the SoTA model just about a mont…
DeepSeek V4 Flash is about 70% smaller than GLM 5.2 yet outperforms it, implying state-of-the-art-level AI could run on consumer hardware like an RTX 5090 much sooner than expected.
Qwen3.8-Max matches Kimi K3 and DeepSeek V4 Flash
Qwen3.8-Max, a 2.4T parameter open-weight model, matches Kimi K3 and DeepSeek V4 Flash on benchmarks, excelling in coding and software tasks. Weights release next week, with pricing of $2/M input and $6/M output tokens.
@wquguru: https://x.com/wquguru/status/2057852569054278045
Performed source code analysis and multi-model testing on the pi-goal tool, finding that DeepSeek V4 Pro is 31x cheaper and higher quality than Gemini 3.5 Flash on long-horizon tasks, and that higher thinking mode actually increases hallucination.