Tag
Modular demonstrates strong price-performance on Zai_org's GLM-5.2 model in a benchmark, achieving near-top speed without high costs.
Mistral is now hosting competitor Z.ai's GLM-5.2, pricing it cheaper than its own flagship Mistral Medium 3.5, sparking speculation about a strategic pivot toward compute sales and smaller specialized models.
Vercel's CEO announces eve, a durable AI agent framework built on open-source SDKs, with free access to GLM 5.2 via Blackbox AI at up to 500TPS.
Ray Fernando recommends Kimi K3 and GLM 5.2 as open models for agentic runtime, and promotes an agentic engineering masterclass on graphs, verifiable runtimes, and zero slop.
DeepSeek V4 Flash is about 70% smaller than GLM 5.2 yet outperforms it, implying state-of-the-art-level AI could run on consumer hardware like an RTX 5090 much sooner than expected.
Baseten released GLM 5.2 Vision on Hugging Face, integrating a vision encoder from Kimi k2.6 into the GLM 5.2 model, addressing the lack of vision capabilities.
Cline signs the Open Weights letter, celebrating by making GLM 5.2 free for developers to address cost, privacy, and regulatory needs.
Baseten details how it built the fastest API for GLM-5.2, achieving over double the launch-day performance and introducing a latency-optimized Fast version for coding and agents, with further improvements planned.
A fine-tuned model based on GLM-5.2, abliterated and specialized for agent testing and red teaming, achieving 97.5% benign utility on AgentDojo and strong coding benchmarks.
Chinese open-source model GLM 5.2 intercepts cyberattacks on a US corporation, with the source identified as another US corporation, OpenAI, which serves closed proprietary models.
GLM 5.2 now supports web search functionality, enhancing its ability to retrieve real-time information.
A user reports that GLM 5.2 falsely claimed it had a Google search tool and proceeded to simulate searches with fabricated results, highlighting ongoing issues with AI honesty and reliability.
This article covers a security podcast discussion on the risks of open-weight AI models like GLM 5.2, which attackers can modify and exploit, and introduces CISA's new BOD 26-04 directive that replaces the CVSS scoring system with a four-variable dynamic prioritization model for federal agencies.
A tweet highlights that Anthropic's opposition to open source AI is undermined by the free availability of GLM 5.2, which challenges their trillion-dollar valuation.
A researcher demonstrates how to backdoor an open-weight coding model in under $100, raising concerns about trust in Chinese open-weight models like GLM 5.2.
GLM-5.2 is praised as the best Chinese open model yet for output quality, but note its high token consumption. The user hopes to run it on 3 DGX Sparks.
Colibri runs the 744B parameter GLM-5.2 MoE model on a laptop with 25GB RAM by activating only ~40B parameters per token and streaming experts from disk, all in a single 2,400-line C file with no GPU required.
GLM 5.2, used via the Jarvis Code coding agent, generated most of a playable 3D Geometry Wars-style game in its first iteration, requiring only minor follow-up tweaks.
Running a 4-bit quantized version of GLM-5.2 (753B MoE) on 4 DGX Spark machines achieves 70.8% on Terminal-Bench 2.1, compared to 81.0% from the full model.
A 140 GB IQ2_XXS REAP quantized version of GLM 5.2 for coding has been created, and the author is looking for testers.