Tag
Multiple leading Chinese AI labs have released new flagship models within the past month, including Kimi K3, Qwen3.8, DeepSeek-V4-Pro, and GLM-5.3, signaling rapid progress in China's AI race.
Tim Dettmers, creator of bitsandbytes, is teasing a new quantization method that reportedly runs GLM 5.3 on a single DGX Spark at 7 tokens/s, with cautious optimism from the community.
GLM 5.3 identified 2436 unpatched open-source vulnerabilities, 1097 rated critical or high, many over 26 years old, in what appears to be Z.ai's counterpart to Project Glasswing.
GLM 5.3 has been released. The author compares the speed and intelligence of GLM 5.2 and DeepSeek V4 Flash, and considers whether to switch back to GLM.
Z.ai has released GLM 5.3, an update to its GLM language model family. The official announcement provides details on the new version.
A comparison of DeepSeek 0813 'Pro' against GLM 5.2 and Kimi K3, likely covering benchmark performance and capability differences between these AI models.
The author argues that competition from Chinese AI labs like DeepSeek, Qwen, GLM, and Kimi benefits consumers by pressuring major AI companies to improve quality and keep prices reasonable.
A discussion questioning whether older AI models like GLM 5.2 and Kimi 2.7 remain relevant for coding now that newer models such as Kimi K3, Qwen 3.8 Max, and DeepSeek V4 Pro are arriving.
Mind Lab claims its Macaron-V1 model surpasses GLM-5.2 in benchmarks, using five LoRA expert modules attached to GLM-5.1 with dynamic expert switching and continual learning via distilled LoRA adapters.
The tweet mentions that GLM-5.3 has been launched on GLM's official website, and marvels at the pressure Deepseek v4 flash and Qwen 3.8 max are facing.
Nova versão do DS v4 flash 0731 promete grande melhoria por preço baixo, com desempenho superior ao GLM 5.1 em testes caseiros, embora haja desconfiança sobre os benchmarks.
A tweet highlights DepthFirst Labs' new cybersecurity model dfs-large1, which matches frontier-model performance on vulnerability discovery, built on the open GLM-5.2 model with RL post-training on Fireworks AI, arguing that open weights are a defender's advantage.
Alibaba Cloud launches a unified Token Plan that pools credits for text, image, video, and audio AI models, including Qwen, DeepSeek, GLM, and others, with tiered pricing starting at $4 for the first month.
An analysis of output similarity suggests that Chinese AI frontier models like Kimi K3, DeepSeek V4, and GLM 5.2 show stylistic differences indicating meaningful independent development, challenging claims that their progress relies heavily on distillation from US models.
The author observes that domestic AI models (Kimi, GLM, DeepSeek) improve with each update, believing that domestic model R&D has entered a virtuous cycle: user growth brings abundant data, mature engineering infrastructure, tight compute power forcing efficiency optimization, and clear pricing advantages.
A company shares the practical difficulties of purchasing and maintaining an HGX B300 for AI inference, including high cost (€1.1M), power and space requirements, and limited availability, with throughput estimates for GLM 5.2.
A developer demonstrates adding vision capabilities to the GLM language model, showcasing a significant multimodal extension.
A tweet from Ahmad Osman predicts that August to October 2026 will be a historic period for open-source frontier AI models, including mentions of Kimi K3 and GLM 6, suggesting upcoming significant releases.
Colibri enables running the 744B-parameter GLM 5.2 model locally on CPU, making large-scale AI accessible without GPU hardware.
Upcoming AI model releases include Kimi K3, Deepseek V4, new Liquid and Mistral models, with rumors of GLM 5.5 in August, indicating a busy period for open-weight AI.