glm

Tag

Cards List
#glm

@qingke_ai: https://x.com/qingke_ai/status/2072159674078736556

X AI KOLs Timeline · 2026-07-01 Cached

This article details the latest progress of GLM-5.2 in Agentic RL, including the introduction of slime infrastructure, shifting from GRPO to PPO for handling long trajectories, and an online anti-cheat mechanism; it also explores Qwen's research on verifier quality, proposing three dimensions of scalability, faithfulness, and robustness, and designs multiple verification strategies for different tasks to improve the reliability of reward signals.

0 favorites 0 likes
#glm

DeepSeek-V4-GLM-5.2-PRO🔥

Reddit r/singularity · 2026-06-29

DeepSeek releases version 4 of its GLM model, version 5.2 PRO.

0 favorites 0 likes
#glm

Effect of GLM 5.2 !!

Reddit r/singularity · 2026-06-29

GLM 5.2, a new version of the GLM language model, has been released, demonstrating improved performance.

0 favorites 0 likes
#glm

Kimi and GLM on frontier code

Reddit r/LocalLLaMA · 2026-06-29

Moonshot AI's Kimi and Zhipu AI's GLM have achieved notable results on frontier code benchmarks.

0 favorites 0 likes
#glm

CPU-only GLM 5.2: Epyc and 512GB RAM

Reddit r/LocalLLaMA · 2026-06-29

GLM 5.2 is optimized for CPU-only inference on AMD Epyc processors with 512GB RAM.

0 favorites 0 likes
#glm

GLM 5.2 Q1_S vs Qwen 27B Q8

Reddit r/LocalLLaMA · 2026-06-29

A hobbyist compares a heavily quantized GLM 5.2 (Q1_S) against a high-quant Qwen 27B (Q8) on a code generation task, finding that the lower-quant larger model significantly outperforms the higher-quant smaller model in quality and completeness.

0 favorites 0 likes
#glm

Free GLM 5.2 — ok usage limits

Reddit r/AI_Agents · 2026-06-28

Zhipu AI launches a free tier for GLM 5.2 with usage limits.

0 favorites 0 likes
#glm

huihui-ai/Huihui-GLM-5.2-abliterated-GGUF

Hugging Face Models Trending · 2026-06-28 Cached

A quantized GGUF version of the abliterated GLM-5.2 model is released on Hugging Face, enabling local inference with various tools like Transformers, llama.cpp, and vLLM.

0 favorites 0 likes
#glm

@VukRosic99: GLM 5.2 post-training code is OPEN SOURCE (slime) Megatron-LM trains. SGLang generates the rollouts. A single data buff…

X AI KOLs Timeline · 2026-06-27 Cached

GLM 5.2 post-training code is open-sourced, using Megatron-LM for training and SGLang for rollout generation, forming a continuous RL loop with synchronized weights.

0 favorites 0 likes
#glm

@seclink: 有点意思 ....

X AI KOLs Timeline · 2026-06-26 Cached

TileRT is a tile-based runtime achieving ultra-low-latency LLM inference, with recent milestones including 1000+ tokens/s on a 1-trillion-parameter model. It supports models like DeepSeek-V3.2 and GLM-5, and is available as open-source on GitHub.

0 favorites 0 likes
#glm

@geekbb: Using Hugging Face to access nvidia/GLM-5.2-NVFP4, which is NVIDIA's NVFP4 precision version quantized from the Zhipu GLM-5.2 model. I'm thinking it should at least be stronger than deepseek-v4-flash. Hug…

X AI KOLs Timeline · 2026-06-26 Cached

NVIDIA has released an NVFP4 precision version quantized from the Zhipu GLM-5.2 model, available via the Hugging Face free tier API.

0 favorites 0 likes
#glm

@lqiao: Cursor . GLM5.2 . Fireworks

X AI KOLs Following · 2026-06-25 Cached

GLM 5.2, an open AI model, is now available in the Cursor coding tool via a partnership with Fireworks.

0 favorites 0 likes
#glm

Locked Dell quote for 6x RTX PRO 6000 Max-Q at $8,960 — expires tonight. What would you do?

Reddit r/LocalLLaMA · 2026-06-25

A user discusses a locked Dell quote for 6x RTX PRO 6000 Max-Q GPUs at a discounted price to build an inference cluster for GLM 5.2, asking the community for advice on purchasing strategy before the quote expires.

0 favorites 0 likes
#glm

@TheAhmadOsman: GPT 5.5 > GLM 5.2 But GLM 5.2 > Opus 4.8

X AI KOLs Following · 2026-06-23 Cached

A comparison stating GPT 5.5 outperforms GLM 5.2, but GLM 5.2 outperforms Opus 4.8.

0 favorites 0 likes
#glm

GLM5.2 @7tg on 4x3090 + 192GB on budget motherboard + cpu

Reddit r/LocalLLaMA · 2026-06-22

Running GLM5.2 with 7 trillion tokens on a budget setup using 4x RTX 3090 GPUs and 192GB RAM.

0 favorites 0 likes
#glm

@karminski3: Thinking of buying a Mac to run large models? This is a deterrent post. Actually, the estimation method is simple. Even if you buy a MacStudio to run the Qwen3.6-27B 4bit quantized version, then enable DFlash to use Qwen's built-in speculative decoding, it only reaches 65 token/s. And now most large models can run at 40 token/s…

X AI KOLs Timeline · 2026-06-22 Cached

The author calculates the token cost and break-even period of running large models on a Mac Studio, concluding that it is not cost-effective for ordinary users to buy a Mac for personal large model use, and suggests that using APIs or renting GPUs is more economical.

0 favorites 0 likes
#glm

@guohao_li: yes, it is definitely time to seriously consider buying more GPUs and start building our own local ai stack. i’m curiou…

X AI KOLs Following · 2026-06-22 Cached

A researcher suggests it's time to buy more GPUs and build a local AI stack, referencing Qwen 3.5 27B and GLM 5.2 as models that cancel the threat of a permanent underclass.

0 favorites 0 likes
#glm

@LexnLin: why tf is GLM 5.2 and Kimi 2.7 unlimited in a Devin sub rn? literally a gold mine

X AI KOLs Following · 2026-06-21 Cached

The tweet highlights that GLM 5.2 and Kimi 2.7 are available without limits in a Devin subscription, describing it as a gold mine.

0 favorites 0 likes
#glm

@antirez: First kinda working implementation of GLM 5.2 in DwarfStar. Will take some time to be good enough, but it is a promisin…

X AI KOLs Following · 2026-06-21 Cached

Antirez reports the first working implementation of GLM 5.2 in DwarfStar, using a 433 GB GGUF file on an M3 Ultra with 512GB RAM, though it needs further refinement.

0 favorites 0 likes
#glm

@natolambert: An hour in and first impression is definitely that GLM is really solid (very easy to set up on @FireworksAI_HQ, props t…

X AI KOLs Following · 2026-06-21

Natolambert shares a positive first impression of GLM, noting it is easy to set up on Fireworks AI and works well with Claude Code.

0 favorites 0 likes
← Previous
Next →
← Back to home

Submit Feedback